-
Notifications
You must be signed in to change notification settings - Fork 6
Pull requests: InfiniTensor/InfiniOps
Author
Label
Projects
Milestones
Reviews
Assignee
Sort
Pull requests list
feat(ops): add vLLM-aligned paged attention v1
#893
opened Aug 6, 2026 by
voltjia
Collaborator
Loading…
6 of 19 tasks
fix(torch): honor handle streams in generated operators
#880
opened Aug 4, 2026 by
voltjia
Collaborator
Loading…
7 of 19 tasks
feat(auto-tuing): add a new auto-tuning system
#879
opened Aug 4, 2026 by
mingdaw689
Loading…
17 tasks
perf: avoid redundant operator cache lookups
#858
opened Jul 30, 2026 by
baominghelly
Contributor
Loading…
10 of 19 tasks
perf: optimize generic tensor conversion
#832
opened Jul 28, 2026 by
baominghelly
Contributor
Loading…
9 of 19 tasks
feat(moore): support
flash_attn_varlen_func
#819
opened Jul 24, 2026 by
voltjia
Collaborator
Loading…
11 of 19 tasks
perf: reuse ATen tensors and optimize PyTorch fallback wrappers
#798
opened Jul 13, 2026 by
baominghelly
Contributor
•
Draft
13 of 19 tasks
feat(ascend): add
rotary_embedding operator
#786
opened Jun 30, 2026 by
zhangyue207
Contributor
•
Draft
feat(ascend): add
flash_attention operator
#783
opened Jun 30, 2026 by
zhangyue207
Contributor
•
Draft
feat(ascend): add
reshape_and_cache operator
#782
opened Jun 30, 2026 by
zhangyue207
Contributor
•
Draft
feat(ascend): add
causal_softmax operator
#780
opened Jun 30, 2026 by
zhangyue207
Contributor
•
Draft
feat(ascend): add
mha_varlen_fwd operator
#779
opened Jun 30, 2026 by
zhangyue207
Contributor
•
Draft
Previous Next
ProTip!
Type g i on any issue or pull request to go back to the issue listing page.