Skip to content

Pull requests: SemiAnalysisAI/InferenceX

Author
Filter by author
Loading
Label
Filter by label
Loading
Use alt + click/return to exclude labels
or + click/return for logical OR
Projects
Filter by project
Loading
Milestones
Filter by milestone
Loading
Reviews
Assignee
Filter by who’s assigned
Assigned to nobody Loading
Sort

Pull requests list

Update Kimi K3 GB300 Agentx full-sweep-enabled
#2811 opened Sep 3, 2026 by wzhao18 Collaborator Loading…
perf(agentx): refresh K3 MI355X vLLM recipe with DCP8 MTP arm agentx AgentX benchmarks, recipes, and infrastructure AMD
#2810 opened Sep 3, 2026 by seungrokj Collaborator Loading…
1 task
ci: use Fable 5.1 for all Claude Code invocations
#2805 opened Sep 2, 2026 by cquil11 Collaborator Loading…
perf(agentx): refresh Kimi-K3 MI355X LMCache curve / 刷新 Kimi-K3 MI355X LMCache 曲线 agentx AgentX benchmarks, recipes, and infrastructure AMD
#2804 opened Sep 2, 2026 by hyukjlee Collaborator Loading…
4 of 9 tasks
Validate vLLM P/D cache-source metrics on GB300 full-sweep-enabled
#2797 opened Sep 1, 2026 by cquil11 Collaborator Loading…
[AMD] [AGENTX] Kimi Perf Tuning agentx AgentX benchmarks, recipes, and infrastructure AMD
#2795 opened Sep 1, 2026 by ajith-sirra-amd Collaborator Loading…
[AgentX] Validate vLLM cached-token tier metrics
#2766 opened Aug 28, 2026 by cquil11 Collaborator Draft
Align CI business priority across node counts and model families
#2765 opened Aug 27, 2026 by cquil11 Collaborator Loading…
ProTip! Updated in the last three days: updated:>2026-08-30.