-
Notifications
You must be signed in to change notification settings - Fork 174
Pull requests: intel/auto-round
Author
Label
Projects
Milestones
Reviews
Assignee
Sort
Pull requests list
feat: data-parallel block tuning via --parallel_quantization
#2351
opened Sep 11, 2026 by
avtc
Collaborator
Loading…
3 of 4 tasks
fix: update key cache access for compatibility with newer DynamicCache implementation
#2348
opened Sep 11, 2026 by
chensuyue
Contributor
Loading…
4 tasks
fix: support list inputs in diffusion tuning cache
#2346
opened Sep 11, 2026 by
changwangss
Contributor
Loading…
4 tasks
add fineweb-edu dataset as calibration bakeup
#2345
opened Sep 11, 2026 by
WeiweiZhang1
Contributor
Loading…
3 of 4 tasks
Refine device operations into a unified DeviceManager and ARDevice abstraction
#2344
opened Sep 11, 2026 by
lvliang-intel
Contributor
•
Draft
1 of 4 tasks
feat: retain diffusion tuning samples in GPU cache
#2342
opened Sep 11, 2026 by
changwangss
Contributor
•
Draft
4 tasks
NeUQI grid search for optimized RTN (--enable_neuqi: joint asym scale/zp search, two-stage sym search, frozen-init anchor)
#2341
opened Sep 10, 2026 by
avtc
Collaborator
Loading…
3 of 4 tasks
unify target_bits/options into bits/schemes for AutoScheme
#2333
opened Sep 10, 2026 by
n1ck-guo
Contributor
Loading…
4 tasks
[ARK] Improve XPU W4A16 WOQ decode with dense S4 DPAS path
#2331
opened Sep 9, 2026 by
Zhenzhong1
Contributor
•
Draft
feat(ark): add INT4 S4 pre-packed Q*K kernel for SageAttention
#2319
opened Sep 8, 2026 by
luoyu-intel
Contributor
•
Draft
Enhance offload cleanup handling for exception cases
#2317
opened Sep 8, 2026 by
lvliang-intel
Contributor
Loading…
1 of 4 tasks
Recurrent Residual Quantization (RRQ) for LLMs
#2308
opened Sep 6, 2026 by
luoyu-intel
Contributor
Loading…
support teq algo
experimental
WIP
#2301
opened Sep 4, 2026 by
WeiweiZhang1
Contributor
Loading…
4 tasks
Add lagrangian solver in AutoScheme
#2221
opened Aug 24, 2026 by
wenhuach21
Contributor
Loading…
4 tasks
feat: W4A8 ARK XPU MoE kernel (int4 weight / int8 compute) with prefill + decode
#2143
opened Aug 11, 2026 by
Copilot
AI
Loading…
4 tasks done
Support vLLM-based Model Quantization with llm_compressor Export
#1978
opened Jul 1, 2026 by
changwangss
Contributor
Loading…
4 tasks
Previous Next
ProTip!
no:milestone will show everything without a milestone.