-
Notifications
You must be signed in to change notification settings - Fork 174
Pull requests: intel/auto-round
Author
Label
Projects
Milestones
Reviews
Assignee
Sort
Pull requests list
NeUQI grid search for optimized RTN (--enable_neuqi: joint asym scale/zp search, two-stage sym search, frozen-init anchor)
#2341
opened Sep 10, 2026 by
avtc
Collaborator
Loading…
3 of 4 tasks
unify target_bits/options into bits/schemes for AutoScheme
#2333
opened Sep 10, 2026 by
n1ck-guo
Contributor
Loading…
4 tasks
[ARK] Improve XPU W4A16 WOQ decode with dense S4 DPAS path
#2331
opened Sep 9, 2026 by
Zhenzhong1
Contributor
•
Draft
feat: implement resume functionality for model-free compression
#2330
opened Sep 9, 2026 by
xin3he
Contributor
Loading…
4 tasks
Keep the original checkpoint key names and fix Qwen4 vLLM inference
#2327
opened Sep 9, 2026 by
wenhuach21
Contributor
Loading…
4 tasks
feat(ark): add INT4 S4 pre-packed Q*K kernel for SageAttention
#2319
opened Sep 8, 2026 by
luoyu-intel
Contributor
•
Draft
Enhance offload cleanup handling for exception cases
#2317
opened Sep 8, 2026 by
lvliang-intel
Contributor
Loading…
1 of 4 tasks
Recurrent Residual Quantization (RRQ) for LLMs
#2308
opened Sep 6, 2026 by
luoyu-intel
Contributor
Loading…
support teq algo
experimental
WIP
#2301
opened Sep 4, 2026 by
WeiweiZhang1
Contributor
Loading…
4 tasks
Add lagrangian solver in AutoScheme
#2221
opened Aug 24, 2026 by
wenhuach21
Contributor
Loading…
4 tasks
feat: W4A8 ARK XPU MoE kernel (int4 weight / int8 compute) with prefill + decode
#2143
opened Aug 11, 2026 by
Copilot
AI
Loading…
4 tasks done
Support vLLM-based Model Quantization with llm_compressor Export
#1978
opened Jul 1, 2026 by
changwangss
Contributor
Loading…
4 tasks
Add quantization support for DiffusionGemma
#1935
opened Jun 17, 2026 by
lvliang-intel
Contributor
Loading…
1 of 4 tasks
ProTip!
Filter pull requests by the default branch with base:main.