-
Notifications
You must be signed in to change notification settings - Fork 89
Pull requests: Avarok-Cybersecurity/atlas
Author
Label
Projects
Milestones
Reviews
Assignee
Sort
Pull requests list
feat(tui): Benchmarks section, and model state that can actually be torn down
#385
opened Jul 30, 2026 by
tbraun96
Contributor
Loading…
fix(ssm-tier): reap dead tier keys so a capped disk cannot thrash
#382
opened Jul 29, 2026 by
rsafier
Collaborator
•
2/2
Loading…
perf(ssm-tier): 22x cheaper snapshot eviction + a bounded, functional disk tier
#381
opened Jul 29, 2026 by
rsafier
Collaborator
•
1/2
Loading…
chore(deps): Bump the actions-all group with 2 updates
#376
opened Jul 27, 2026 by
dependabot
Bot
Loading…
fix(scheduler): preempt on KV exhaustion instead of failing (decode + prefill paths)
#375
opened Jul 26, 2026 by
rsafier
Collaborator
Loading…
fix(kv): prefix-cache refcount + block aliasing (CUDA-700, pool leak, HSS mis-filing)
#373
opened Jul 26, 2026 by
rsafier
Collaborator
Loading…
chore(deps): Bump the patch-updates group across 1 directory with 9 updates
#355
opened Jul 22, 2026 by
dependabot
Bot
Loading…
perf(strix): Qwen3.6-27B-NVFP4 on Strix Halo (gfx1151) — Atlas beats llama.cpp on decode, prefill & accuracy
#353
opened Jul 22, 2026 by
AzeezIsh
Collaborator
Loading…
feat(mtp): make the drafter context (prefill + cross-turn carry) the default
#352
opened Jul 22, 2026 by
tbraun96
Contributor
Loading…
spark-model: add Laguna-S-2.1 inference support
#350
opened Jul 21, 2026 by
rsafier
Collaborator
Loading…
strix: native-FP8-GDN golden e2e config + TTFT-parity (validated 2026-07-19)
#336
opened Jul 19, 2026 by
AzeezIsh
Collaborator
Loading…
feat(lora): MoE per-expert + router LoRA & embed/lm_head/vocab overlay (correctness-first)
#335
opened Jul 19, 2026 by
rsafier
Collaborator
Loading…
feat: generic GGUF loader + Ternary-Bonsai-27B (2-bit keep-packed on one GB10)
#334
opened Jul 19, 2026 by
rsafier
Collaborator
Loading…
perf(decode): lm_head batched-GEMV (35B + 27B) + Qwen3.6-27B decode bring-up (11->58 tok/s)
#332
opened Jul 18, 2026 by
rsafier
Collaborator
Loading…
perf(decode): w4a16 GEMV 2-chunk ILP + fix K=4 verify crash on native-FP8-GDN ckpts
#330
opened Jul 18, 2026 by
tbraun96
Contributor
Loading…
feat(tool_parser): add native plain-text tool-call streaming parser for gemma4
#325
opened Jul 15, 2026 by
drewhemm
Loading…
fix(jinja): prevent infinite tool-calling loops in gemma4 multi-turn completions
#324
opened Jul 15, 2026 by
drewhemm
Loading…
perf(gdn-prefill): ldmatrix A+B GDN-projection GEMM — 2.1x kernel, warm-TTFT median -5.8% / p90 -10.3%
#296
opened Jul 10, 2026 by
tbraun96
Contributor
Loading…
refactor(loader): gate batched-prefill MLA exclusion on capability, not model_type
#285
opened Jul 10, 2026 by
rsafier
Collaborator
Loading…
feat(gdn-tp): dense-model GDN/SSM TP-shard (dense pure-TP) [stacks on #254+#250]
#284
opened Jul 10, 2026 by
rsafier
Collaborator
Loading…
fix(sampling): preset fallback (not greedy 0) when generation_config absent — #228 review guard
#267
opened Jul 5, 2026 by
rsafier
Collaborator
Loading…
Previous Next
ProTip!
Add no:assignee to see everything that’s not assigned.