-
Notifications
You must be signed in to change notification settings - Fork 2.5k
Pull requests: JustVugg/colibri
Author
Label
Projects
Milestones
Reviews
Assignee
Sort
Pull requests list
fix(cli): preserve pasted multiline prompts
#909
opened Aug 9, 2026 by
SulimanAbdulrazzaq
Loading…
5 of 7 tasks
feat(moe): add DEGRADE_ZERO miss-slot zero-fill policy (issue #865)
#906
opened Aug 9, 2026 by
kritikagarg
Loading…
5 tasks
deepseek_v4: honor CTX in the generate CLI
#904
opened Aug 8, 2026 by
Blakeolson21
Loading…
4 of 5 tasks
Fix: make placement accounting backend neutral
#903
opened Aug 8, 2026 by
terrizoaguimor
Contributor
Loading…
Experiment: add offline residency simulator
#902
opened Aug 8, 2026 by
terrizoaguimor
Contributor
Loading…
deepseek_v4: size the OpenMP team around the expert loaders
#897
opened Aug 8, 2026 by
Blakeolson21
Loading…
4 of 5 tasks
fix: report Vulkan expert residency in telemetry and brain map
bug
Difetto verificato nel codice
needs-rebase
Confligge, serve rebase dell'autore
vulkan
Backend Vulkan/AMD
#891
opened Aug 8, 2026 by
MasterCATZ
Loading…
deepseek_v4: emit dashboard telemetry (HWINFO/TIERS/EMAP/PROF/HITS) in serve mode
enhancement
New feature or request
#882
opened Aug 7, 2026 by
PwrBank
Loading…
inkling: ring-buffer KV cache for sliding-window layers (~10x less KV memory at long context)
enhancement
New feature or request
performance
Velocità / tok-s / ottimizzazioni
#830
opened Aug 4, 2026 by
dpanelli
Loading…
fix(metal): int4-g64 (fmt=4) models never used the GPU — enable MoE experts + fused attention
bug
Difetto verificato nel codice
metal
Backend Metal/Apple
needs-rebase
Confligge, serve rebase dell'autore
#829
opened Aug 4, 2026 by
aaristov
Loading…
engine: notice when a tensor's format silently disables the fused Metal decode path
enhancement
New feature or request
metal
Backend Metal/Apple
#827
opened Aug 4, 2026 by
monotophic
Contributor
Loading…
Feat/dstorage transport
enhancement
New feature or request
performance
Velocità / tok-s / ottimizzazioni
#825
opened Aug 4, 2026 by
khalilswdp
Contributor
Loading…
4 of 5 tasks
managed runner, deterministic benchmark, and causal evidence (PR3 of #377 split)
discussion
Proposta / discussione aperta, non un task
enhancement
New feature or request
performance
Velocità / tok-s / ottimizzazioni
#822
opened Aug 4, 2026 by
BColsey
Loading…
ramdisk: headless planning, staging, mounts, recovery + tokenized CLI (PR2 of #377 split)
discussion
Proposta / discussione aperta, non un task
enhancement
New feature or request
performance
Velocità / tok-s / ottimizzazioni
#821
opened Aug 4, 2026 by
BColsey
Loading…
engine: RAMMAP, NUMA, and telemetry (PR1 of #377 split)
discussion
Proposta / discussione aperta, non un task
enhancement
New feature or request
performance
Velocità / tok-s / ottimizzazioni
#820
opened Aug 4, 2026 by
BColsey
Loading…
Add a registry for the 214 environment variables, and check the environment against it
enhancement
New feature or request
#800
opened Aug 3, 2026 by
ZacharyZcR
Contributor
Loading…
Dead code, de-duplication, repo-wide clang-format, and lint gates in CI
enhancement
New feature or request
needs-rebase
Confligge, serve rebase dell'autore
#798
opened Aug 3, 2026 by
ZacharyZcR
Contributor
Loading…
Metal (Apple GPU) backend for Kimi K3
feature
Nuova funzionalità
metal
Backend Metal/Apple
model-support
Supporto a nuovi modelli
#790
opened Aug 2, 2026 by
RDouglasSharp
Contributor
Loading…
[WIP] Vulkan stats
enhancement
New feature or request
vulkan
Backend Vulkan/AMD
#789
opened Aug 2, 2026 by
Neppord
Loading…
5 tasks
feature: Hy3 support
discussion
Proposta / discussione aperta, non un task
model-support
Supporto a nuovi modelli
#775
opened Aug 2, 2026 by
ErikTromp
Loading…
5 tasks done
DeepSeek V4: the CUDA kernels, as a tier rather than an engine
discussion
Proposta / discussione aperta, non un task
model-support
Supporto a nuovi modelli
#772
opened Aug 2, 2026 by
ZacharyZcR
Contributor
Loading…
Vulkan MoE GEMV backend for integrated/AMD GPUs (draft, complements #418)
enhancement
New feature or request
feature
Nuova funzionalità
vulkan
Backend Vulkan/AMD
feat(qwen36): CUDA VRAM expert tier — heat-based placement across GPUs via the shared CUDA backend
cuda
Backend CUDA/NVIDIA
model-support
Supporto a nuovi modelli
feat(qwen36): Qwen3.6-35B-A3B engine (CPU): hybrid Gated Attention + Gated DeltaNet + streaming MoE
model-support
Supporto a nuovi modelli
#712
opened Jul 30, 2026 by
kreuzzelg
Contributor
Loading…
feat(win): fix silent CPU fallback, launcher suite, DirectStorage expert loads
enhancement
New feature or request
needs-rebase
Confligge, serve rebase dell'autore
#670
opened Jul 28, 2026 by
khalilswdp
Contributor
Loading…
8 tasks done
Previous Next
ProTip!
Type g p on any issue or pull request to go back to the pull request listing page.