Skip to content

Pull requests: NVIDIA/TensorRT-LLM

Author
Filter by author
Loading
Label
Filter by label
Loading
Use alt + click/return to exclude labels
or + click/return for logical OR
Projects
Filter by project
Loading
Milestones
Filter by milestone
Loading
Reviews
Assignee
Filter by who’s assigned
Assigned to nobody Loading
Sort

Pull requests list

[https://nvbugs/6468821][chore] Unwaive GPT-OSS B300 attention backend test
#16616 opened Jul 20, 2026 by yuxianq Collaborator Loading…
1 task done
[None][infra] Waive 1 failed cases for main in pre-merge 48660
#16615 opened Jul 20, 2026 by trtllm-agent Collaborator Loading…
[None][Test] Consolidate dis-agg E2E Tests
#16614 opened Jul 20, 2026 by Shixiaowei02 Collaborator Loading…
1 task done
[None][test] Add deepseek v4 pro cases on the qa side
#16611 opened Jul 20, 2026 by fredricz-20070104 Collaborator Loading…
[None][feat] Support MARLIN MoE with MTP and attention DP + EP
#16597 opened Jul 20, 2026 by Wanli-Jiang Collaborator Loading…
1 task done
[None][perf] Gate NCCL NVLS on NVML fabric state, not IMEX availability
#16595 opened Jul 20, 2026 by Wanli-Jiang Collaborator Loading…
1 task done
[TRTLLM-13409][fix] hard-kill all ranks when one rank's executor loop crashes
#16592 opened Jul 20, 2026 by JunyiXu-nv Collaborator Loading…
1 task done
[https://nvbugs/6305365][chore] Unwaive piecewise cudagraph related tests
#16591 opened Jul 20, 2026 by pengbowang-nv Collaborator Loading…
1 task done
ProTip! Type g p on any issue or pull request to go back to the pull request listing page.