Skip to content

Pull requests: NVIDIA/TensorRT-LLM

Author
Filter by author
Loading
Label
Filter by label
Loading
Use alt + click/return to exclude labels
or + click/return for logical OR
Projects
Filter by project
Loading
Milestones
Filter by milestone
Loading
Reviews
Assignee
Filter by who’s assigned
Assigned to nobody Loading
Sort

Pull requests list

Feat/golden prairie multimodal
#17043 opened Jul 30, 2026 by WeiHaocheng Collaborator Loading…
1 task
[https://nvbugs/6525010][fix] Fix TestLlama3_3_70BInstruct::test_nvfp…
#17040 opened Jul 30, 2026 by liji-nv Collaborator Loading…
1 task done
[https://nvbugs/6481375][test] Unwaive passing DSV3-Lite tests
#17038 opened Jul 30, 2026 by lfr-0531 Collaborator Draft
1 task done
[#17020][fix] DeepSeek-V4: honor thinking_budget in apply_chat_template
#17035 opened Jul 30, 2026 by thorjohnsen Collaborator Draft
1 task done
[None][feat] Native V2 KV cache event publishing
#17023 opened Jul 29, 2026 by tanmayv25 Collaborator Draft
[None][perf] Add CUTEDSL FC2 N-tile tuning override
#17015 opened Jul 29, 2026 by peihu-nv Collaborator Draft
1 task done
[None][fix] Use HND mapping for MiniMax-M3 MSA KV cache
#17011 opened Jul 29, 2026 by peihu-nv Collaborator Loading…
1 task done
[None][perf] MiniMax-M3 MSA: split the KV axis for the decode indexer proxy
#17008 opened Jul 29, 2026 by hyukn Collaborator Draft
1 task done
ProTip! Add no:assignee to see everything that’s not assigned.