-
Notifications
You must be signed in to change notification settings - Fork 174
Pull requests: intel/auto-round
Author
Label
Projects
Milestones
Reviews
Assignee
Sort
Pull requests list
unify target_bits/options into bits/schemes for AutoScheme
#2333
opened Sep 10, 2026 by
n1ck-guo
Contributor
Loading…
4 tasks
[ARK] Improve XPU W4A16 WOQ decode with dense S4 DPAS path
#2331
opened Sep 9, 2026 by
Zhenzhong1
Contributor
•
Draft
feat: implement resume functionality for model-free compression
#2330
opened Sep 9, 2026 by
xin3he
Contributor
Loading…
4 tasks
Fix Qwen4 layer_type changed by Transformers, causing vLLM inference failure
#2327
opened Sep 9, 2026 by
wenhuach21
Contributor
Loading…
4 tasks
feat(ark): add INT4 S4 pre-packed Q*K kernel for SageAttention
#2319
opened Sep 8, 2026 by
luoyu-intel
Contributor
•
Draft
Enhance offload cleanup handling for exception cases
#2317
opened Sep 8, 2026 by
lvliang-intel
Contributor
Loading…
1 of 4 tasks
Recurrent Residual Quantization (RRQ) for LLMs
#2308
opened Sep 6, 2026 by
luoyu-intel
Contributor
Loading…
support teq algo
experimental
WIP
#2301
opened Sep 4, 2026 by
WeiweiZhang1
Contributor
Loading…
4 tasks
Improve the low_cpu_mem_usage usage in the doc
#2278
opened Sep 2, 2026 by
hshen14
Contributor
Loading…
Add lagrangian solver in AutoScheme
#2221
opened Aug 24, 2026 by
wenhuach21
Contributor
Loading…
4 tasks
feat: W4A8 ARK XPU MoE kernel (int4 weight / int8 compute) with prefill + decode
#2143
opened Aug 11, 2026 by
Copilot
AI
Loading…
2 of 4 tasks
Support vLLM-based Model Quantization with llm_compressor Export
#1978
opened Jul 1, 2026 by
changwangss
Contributor
Loading…
4 tasks
Add quantization support for DiffusionGemma
#1935
opened Jun 17, 2026 by
lvliang-intel
Contributor
Loading…
1 of 4 tasks
ProTip!
Add no:assignee to see everything that’s not assigned.