Skip to content

Pull requests: InfiniTensor/InfiniOps

Author
Filter by author
Loading
Label
Filter by label
Loading
Use alt + click/return to exclude labels
or + click/return for logical OR
Projects
Filter by project
Loading
Milestones
Filter by milestone
Loading
Reviews
Assignee
Filter by who’s assigned
Assigned to nobody Loading
Sort

Pull requests list

fix(ops): select active implementation for configless calls
#910 opened Aug 7, 2026 by voltjia Collaborator Draft
11 of 20 tasks
feat(linked): add vLLM topk sigmoid provider
#909 opened Aug 7, 2026 by voltjia Collaborator Draft
5 of 20 tasks
feat(linked): add vLLM grouped_topk provider
#908 opened Aug 7, 2026 by voltjia Collaborator Draft
5 of 20 tasks
feat(linked): add vLLM moe wna16 gemm provider
#907 opened Aug 7, 2026 by voltjia Collaborator Draft
5 of 19 tasks
refactor(nvidia)!: link vLLM awq marlin repack
#905 opened Aug 7, 2026 by voltjia Collaborator Draft
7 of 20 tasks
feat(linked): add vLLM get cutlass moe data provider
#904 opened Aug 7, 2026 by voltjia Collaborator Draft
5 of 19 tasks
[codex] refactor: consolidate interned Python names
#900 opened Aug 7, 2026 by voltjia Collaborator Draft
feat(ops): add vLLM-aligned paged attention v1
#893 opened Aug 6, 2026 by voltjia Collaborator Loading…
6 of 19 tasks
fix(torch): honor handle streams in generated operators
#880 opened Aug 4, 2026 by voltjia Collaborator Loading…
7 of 19 tasks
feat(auto-tuing): add a new auto-tuning system
#879 opened Aug 4, 2026 by mingdaw689 Loading…
17 tasks
fix: support optional C in Gemm
#870 opened Aug 4, 2026 by voltjia Collaborator Draft
11 of 20 tasks
perf: avoid redundant operator cache lookups
#858 opened Jul 30, 2026 by baominghelly Contributor Loading…
10 of 19 tasks
perf: optimize generic tensor conversion
#832 opened Jul 28, 2026 by baominghelly Contributor Loading…
9 of 19 tasks
feat(moore): support flash_attn_varlen_func
#819 opened Jul 24, 2026 by voltjia Collaborator Loading…
11 of 19 tasks
perf(torch): reuse ATen wrappers in generated operators
#816 opened Jul 24, 2026 by voltjia Collaborator Draft
11 of 19 tasks
feat(triton): add JIT backend with add operator
#800 opened Jul 15, 2026 by fuyou4546 Contributor Draft
2 of 19 tasks
perf: reuse ATen tensors and optimize PyTorch fallback wrappers
#798 opened Jul 13, 2026 by baominghelly Contributor Draft
13 of 19 tasks
feat(ascend): add rotary_embedding operator
#786 opened Jun 30, 2026 by zhangyue207 Contributor Draft
feat(ascend): add add_rms_norm operator
#785 opened Jun 30, 2026 by zhangyue207 Contributor Draft
feat(ascend): add rms_norm operator
#784 opened Jun 30, 2026 by zhangyue207 Contributor Draft
feat(ascend): add flash_attention operator
#783 opened Jun 30, 2026 by zhangyue207 Contributor Draft
feat(ascend): add reshape_and_cache operator
#782 opened Jun 30, 2026 by zhangyue207 Contributor Draft
feat(ascend): add swiglu operator
#781 opened Jun 30, 2026 by zhangyue207 Contributor Draft
feat(ascend): add causal_softmax operator
#780 opened Jun 30, 2026 by zhangyue207 Contributor Draft
feat(ascend): add mha_varlen_fwd operator
#779 opened Jun 30, 2026 by zhangyue207 Contributor Draft
ProTip! Type g i on any issue or pull request to go back to the issue listing page.