Skip to content

Pull requests: SemiAnalysisAI/InferenceX

Author
Filter by author
Loading
Label
Filter by label
Loading
Use alt + click/return to exclude labels
or + click/return for logical OR
Projects
Filter by project
Loading
Milestones
Filter by milestone
Loading
Reviews
Assignee
Filter by who’s assigned
Assigned to nobody Loading
Sort

Pull requests list

CollectiveX: pass the rank's own routing to MoRI combine (mori#475/#546)
#2760 opened Aug 27, 2026 by Oseltamivir Collaborator Loading…
Migrate MI300X runners to the Barite AMD cluster
#2732 opened Aug 25, 2026 by cquil11 Collaborator Loading…
test: validate upstream-native vLLM Router topologies
#2731 opened Aug 25, 2026 by cquil11 Collaborator Loading…
Remove FlashInfer benchmark backends
#2727 opened Aug 25, 2026 by hbarclay Collaborator Loading…
feat(agentx): retune Kimi-K3 FP4 MI355X ATOM DSpark recipe on _0821 (mirror of #2723) agentx AgentX benchmarks, recipes, and infrastructure AMD full-sweep-enabled
#2725 opened Aug 25, 2026 by seungrokj Collaborator Loading…
[WIP][AMD][AgentX] Add Qwen3.8 FP8 MI355X two-node vLLM agentx-fast Run AgentX throughput with 1 warmup request per lane and a 20-minute profile; not reusable sweep-enabled
#2724 opened Aug 25, 2026 by haic0 Collaborator Loading…
6 of 7 tasks
feat(agentx): retune Kimi-K3 FP4 MI355X ATOM DSpark recipe on _0821 agentx AgentX benchmarks, recipes, and infrastructure AMD full-sweep-enabled
#2723 opened Aug 25, 2026 by zejunchen-zejun Collaborator Loading…
[AgentX] DeepSeek-v4-pro llm-d NVL72-B200
#2719 opened Aug 24, 2026 by ilmarkov Collaborator Draft
[AMD] Refresh DeepSeek-R1 MI355X SGLang image / [AMD] 更新 DeepSeek-R1 MI355X SGLang 镜像 all-evals Expand eval selection to every fixed-sequence config full-sweep-fail-fast
#2691 opened Aug 20, 2026 by Oseltamivir Collaborator Loading…
ProTip! no:milestone will show everything without a milestone.