Skip to content

Pull requests: FlashML-org/FreeToken

Author
Filter by author
Loading
Label
Filter by label
Loading
Use alt + click/return to exclude labels
or + click/return for logical OR
Projects
Filter by project
Loading
Milestones
Filter by milestone
Loading
Reviews
Assignee
Filter by who’s assigned
Assigned to nobody Loading
Sort

Pull requests list

feat(models): add Qwen3-Next-80B-A3B support
#212 opened Aug 26, 2026 by akushonkamen Loading…
fix(gguf): chunk the moe_vec z grid past the 65535 cap
#211 opened Aug 26, 2026 by avlp12 Loading…
GGUF: serve DeepSeek-V4-Flash
#210 opened Aug 26, 2026 by vcruz305 Loading…
demo: intentionally duplicate maintenance gate (radar fixture)
#192 opened Aug 25, 2026 by Mitaligrawal Loading…
1 task done
fix(gemma4-gguf): accept scalar attention.head_count_kv
#190 opened Aug 25, 2026 by qxZap Loading…
fix(deps): use PyTorch 2.13 to fix SM89 hangs
#184 opened Aug 25, 2026 by endenis Loading…
fix(engine): load prefill triton kernels at startup, not mid-request
#169 opened Aug 25, 2026 by jason-fxz Collaborator Loading…
GGUF: read multi-shard checkpoints
#154 opened Aug 24, 2026 by vcruz305 Loading…
feat(rocm): serve on AMD GPUs through the HIP toolchain
#137 opened Aug 24, 2026 by paralin Loading…
ProTip! Type g i on any issue or pull request to go back to the issue listing page.