Uh oh!
There was an error while loading. Please reload this page.
- Notifications
You must be signed in to change notification settings - Fork 95
Pull requests: InfiniTensor/InfiniLM
Author
Uh oh!
There was an error while loading. Please reload this page.
Label
Uh oh!
There was an error while loading. Please reload this page.
Projects
Uh oh!
There was an error while loading. Please reload this page.
Milestones
Uh oh!
There was an error while loading. Please reload this page.
Reviews
Assignee
Assigned to nobodyLoading
Uh oh!
There was an error while loading. Please reload this page.
Sort
Pull requests list
feat: support Ktransformers, CPU-GPU MoE offload via FusedMoE layer
#548
opened Aug 21, 2026 by
whjthuContributorLoading…
37 of 49 tasks
fix(cuda-graph): keep replay metadata dynamic across tensor-parallel ranks
#540
opened Aug 15, 2026 by
junjiewang253-ctrlLoading…
feat: add aclnnMatmulAllReduce fusion in InfiniLM for Ascend RowParallelLinear
#533
opened Aug 11, 2026 by
ShaneWoofContributorLoading…
feat(hygon): add Qwen3-235B-A3B BF16/W8A8 inference support
#532
opened Aug 7, 2026 by
qinyiqunContributorLoading…
49 tasks
feat(engine): overlap decode steps with asynchronous token handoff
#524
opened Aug 3, 2026 by
qinyiqunContributorLoading…
perf(server): coalesce streaming SSE output
#517
opened Jul 28, 2026 by
wooway777CollaboratorLoading…
49 tasks
PreviousNext
ProTip!
Exclude everything labeled
bug with -label:bug.