-
Notifications
You must be signed in to change notification settings - Fork 90
Pull requests: Avarok-Cybersecurity/atlas
Author
Label
Projects
Milestones
Reviews
Assignee
Sort
Pull requests list
perf(load): GPU-transpose quantized weights at cold load (-44%) — stacks on #388
#389
opened Aug 3, 2026 by
rsafier
Collaborator
Loading…
perf(gb10): enterprise concurrency — one config wins C=1..128 vs vLLM, env recipe replaced by CLI flags
#388
opened Aug 2, 2026 by
tbraun96
Contributor
Loading…
Detect and set active RDMA network interface
#387
opened Aug 2, 2026 by
ngerakines
Loading…
1 task done
chore(deps): Bump the actions-all group across 1 directory with 3 updates
#386
opened Aug 2, 2026 by
dependabot
Bot
Loading…
fix(ssm-tier): reap dead tier keys so a capped disk cannot thrash
#382
opened Jul 29, 2026 by
rsafier
Collaborator
•
2/2
Loading…
perf(ssm-tier): 22x cheaper snapshot eviction + a bounded, functional disk tier
#381
opened Jul 29, 2026 by
rsafier
Collaborator
•
1/2
Loading…
fix(scheduler): preempt on KV exhaustion instead of failing (decode + prefill paths)
#375
opened Jul 26, 2026 by
rsafier
Collaborator
Loading…
fix(kv): prefix-cache refcount + block aliasing (CUDA-700, pool leak, HSS mis-filing)
#373
opened Jul 26, 2026 by
rsafier
Collaborator
Loading…
chore(deps): Bump the patch-updates group across 1 directory with 9 updates
#355
opened Jul 22, 2026 by
dependabot
Bot
Loading…
perf(strix): Qwen3.6-27B-NVFP4 on Strix Halo (gfx1151) — Atlas beats llama.cpp on decode, prefill & accuracy
#353
opened Jul 22, 2026 by
AzeezIsh
Collaborator
Loading…
feat(mtp): make the drafter context (prefill + cross-turn carry) the default
#352
opened Jul 22, 2026 by
tbraun96
Contributor
Loading…
spark-model: add Laguna-S-2.1 inference support
#350
opened Jul 21, 2026 by
rsafier
Collaborator
Loading…
strix: native-FP8-GDN golden e2e config + TTFT-parity (validated 2026-07-19)
#336
opened Jul 19, 2026 by
AzeezIsh
Collaborator
Loading…
feat(lora): MoE per-expert + router LoRA & embed/lm_head/vocab overlay (correctness-first)
#335
opened Jul 19, 2026 by
rsafier
Collaborator
Loading…
feat: generic GGUF loader + Ternary-Bonsai-27B (2-bit keep-packed on one GB10)
#334
opened Jul 19, 2026 by
rsafier
Collaborator
Loading…
perf(decode): lm_head batched-GEMV (35B + 27B) + Qwen3.6-27B decode bring-up (11->58 tok/s)
#332
opened Jul 18, 2026 by
rsafier
Collaborator
Loading…
perf(decode): w4a16 GEMV 2-chunk ILP + fix K=4 verify crash on native-FP8-GDN ckpts
#330
opened Jul 18, 2026 by
tbraun96
Contributor
Loading…
feat(tool_parser): add native plain-text tool-call streaming parser for gemma4
#325
opened Jul 15, 2026 by
drewhemm
Loading…
fix(jinja): prevent infinite tool-calling loops in gemma4 multi-turn completions
#324
opened Jul 15, 2026 by
drewhemm
Loading…
perf(gdn-prefill): ldmatrix A+B GDN-projection GEMM — 2.1x kernel, warm-TTFT median -5.8% / p90 -10.3%
#296
opened Jul 10, 2026 by
tbraun96
Contributor
Loading…
refactor(loader): gate batched-prefill MLA exclusion on capability, not model_type
#285
opened Jul 10, 2026 by
rsafier
Collaborator
Loading…
Previous Next
ProTip!
Follow long discussions with comments:>50.