Skip to content

cookbook: add Cosmos3-Edge TensorRT-Edge-LLM reasoner and policy notebooks - #326

Open
ConstBob wants to merge 3 commits into
NVIDIA:mainfrom
ConstBob:docs/edge-trt-edge-llm
Open

ConstBob wants to merge 3 commits into
NVIDIA:mainfrom
ConstBob:docs/edge-trt-edge-llm

Conversation

@ConstBob

@ConstBob ConstBob commented Aug 19, 2026 •

Copy link
Copy Markdown
Collaborator

Summary

Testing

Tested on an x86 developer GPU (RTX PRO 6000, SM120) with TensorRT-Edge-LLM v0.10.0. Both added notebooks in this PR were run top-to-bottom (run_with_trt_edge_llm.ipynb and run_policy_with_trt_edge_llm.ipynb).

  • Reasoner: export --task reasoning → llm_build + visual_build → llm_inference on robot_153.jpg. Pipeline finished.
  • Policy: export --task policy → cosmos3_policy_build → cosmos3_policy_inference. [1,16,10], finite=true, ~195 ms (runtime smoke).

On this SM120 box the pipeline succeeded for both notebooks (export, engine build, and C++ inference finished; policy returned [1,16,10] with finite=true). Reasoner captions were not usable (text did not match robot_153.jpg); policy was only checked as a finite action tensor, not DROID quality.

Likely cause: TensorRT engines and ViT/FMHA kernels are SM-specific. The device used for this testing is RTX PRO 6000. Official Edge-LLM targets are Thor / Spark / Orin, etc. I do not have an Official Edge-LLM device (Jetson Thor / DRIVE Thor / DGX Spark / Orin), so caption or policy quality on those SKUs are not checked. x86 is Developer-only; on this SM120 box reasoner captions were not usable and should not be treated as gold. Re-run is needed on Thor, or Spark) before treating quality as validated.

The above issue has been resolved (please refer to the comment threads below).

@nvluxiaoz

Copy link
Copy Markdown
Collaborator

Can we add some accuracy benchmark instruction or some references of whether this is correct?

Comment thread README.md
@ConstBob
ConstBob requested a review from nvluxiaoz August 28, 2026 20:17
@ConstBob

Copy link
Copy Markdown
Collaborator Author

Requesting reviews from @nvluxiaoz @qmiao-hub @rickzw so that we can merge this PR. Thanks!

@nvluxiaoz nvluxiaoz left a comment

Copy link
Copy Markdown
Collaborator

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

LGTM as long as ever has validated.

Comment thread cookbooks/cosmos3/generator/action/run_policy_with_trt_edge_llm.ipynb Outdated
@rickzw
rickzw requested a review from lfengad September 1, 2026 07:12
ConstBob and others added 2 commits September 1, 2026 09:43
Align the TensorRT-Edge-LLM policy notebook with the vLLM-Omni/SGLang
observation (640x540 multiview + pick-and-place prompt) instead of the
reasoner still and upstream CLI placeholder.
@ConstBob
ConstBob requested a review from rickzw September 1, 2026 17:36
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

3 participants