Skip to content

Pull requests: NVIDIA/TensorRT-LLM

Author
Filter by author
Loading
Label
Filter by label
Loading
Use alt + click/return to exclude labels
or + click/return for logical OR
Projects
Filter by project
Loading
Milestones
Filter by milestone
Loading
Reviews
Assignee
Filter by who’s assigned
Assigned to nobody Loading
Sort

Pull requests list

[None][fix] Fix one-model MTP KV cache accounting
#17264 opened Aug 4, 2026 by 2ez4bz Collaborator Loading…
1 task done
[TRTLLMINF-237][infra] Re-home L0_Test SLURM finalizer
#17263 opened Aug 4, 2026 by dpitman-nvda Collaborator Loading…
1 task done
[None][feat] DONT REVIEW glm image support VisualGen
#17260 opened Aug 4, 2026 by yibinl-nvidia Collaborator Draft
1 task
[None][fix] Skip no-op MXFP4 weight padding
#17259 opened Aug 4, 2026 by jiaganc Collaborator Draft
1 task done
[TRTLLMINF-40][fix] Throw typed InfraFailure when SLURM submission yields no job ID
#17255 opened Aug 4, 2026 by brnguyen2 Collaborator Loading…
3 tasks done
[None][infra] CBTS code coverage date early save
#17253 opened Aug 4, 2026 by crazydemo Collaborator Loading…
1 task done
[None][test] Remove two GPT-OSS tests from GB200 pre-merge
#17252 opened Aug 4, 2026 by yizhang-nv Member Loading…
1 task done
[None][infra] Align VisualGen CBTS rule with CODEOWNERS scope
#17251 opened Aug 4, 2026 by crazydemo Collaborator Draft
1 task done
[None][chore] Fix guardword
#17250 opened Aug 4, 2026 by tongyuantongyu Member Loading…
1 task done
[None][doc] Explain benchmark output token length
#17249 opened Aug 4, 2026 by zcxGGmu Loading…
[TRTLLMINF-240][infra] L0 job enhancement for automatic nightly release
#17248 opened Aug 4, 2026 by niukuo Collaborator Loading…
1 task done
[None][fix] Announce MTP shapes to attention metadata in layer-wise benchmarks
#17247 opened Aug 4, 2026 by dc3671 Collaborator Loading…
1 task done
[None][doc] Clarify KV cache host offload limits
#17246 opened Aug 4, 2026 by zcxGGmu Loading…
[https://nvbugs/6545424][perf] Enable Qwen3.5 fused ops under torch.c…
#17243 opened Aug 4, 2026 by liji-nv Collaborator Loading…
1 task done
[None][perf] Autotune large-M MXFP8 GEMM tactics
#17238 opened Aug 4, 2026 by peihu-nv Collaborator Loading…
2 tasks done
[None][perf] Use FlashInfer MXFP8 GEMM for MiniMax-M3 decode
#17237 opened Aug 4, 2026 by peihu-nv Collaborator Loading…
1 task done
[None][perf] Fuse MiniMax-M3 MSA block selection
#17236 opened Aug 4, 2026 by peihu-nv Collaborator Loading…
1 task done
[None][doc] Add trtllm-bench LWS launch guidance
#17235 opened Aug 4, 2026 by zcxGGmu Loading…
[https://nvbugs/6525008][fix] Isolate FlashInfer JIT workspaces for MPI workers
#17233 opened Aug 4, 2026 by VALLIS-NERIA Collaborator Loading…
1 task done
ProTip! Exclude everything labeled bug with -label:bug.