-
Notifications
You must be signed in to change notification settings - Fork 71
Pull requests: intel/llm-scaler
Author
Label
Projects
Milestones
Reviews
Assignee
Sort
Pull requests list
omni: INT4 per-block zero-point GEMM + calling adapter (torchao asymmetric INT4)
#629
opened Aug 19, 2026 by
JWLHS
Loading…
docs(vllm): Qwen3.8-27B DFlash drafter recipe — 72.2 tok/s on Arc Pro B70
#620
opened Aug 16, 2026 by
rmacy
Loading…
docs(omni): add HiDream-O1, Krea2 Turbo, Boogu-Image, Mage-Flow and LTX-2.3 to supported models table
#592
opened Aug 4, 2026 by
KristianZeng
Contributor
Loading…
[XPU] Harden fused GDN recurrent state updates
#570
opened Jul 27, 2026 by
gc-fu
Contributor
Loading…
Fix #533: guard out_proj.weight access in GDN out-projection ESIMD probes
#557
opened Jul 22, 2026 by
joaovgaraujo
Loading…
refactor(omni): multi-stage Docker build with decoupled builder/runtime
#549
opened Jul 17, 2026 by
KristianZeng
Contributor
•
Draft
page_attn: make global-max reduction order-invariant
#535
opened Jul 12, 2026 by
taste-software
Loading…
Fix Qwen3.5/3.6 load_weights stacked-mapping name mutation (gate_gate_up_proj / qkqkv_proj)
#475
opened Jun 12, 2026 by
bongmiin
Loading…
vllm: add MiniCPM-V 4.6 support (MiniCPMV4_6ForConditionalGeneration)
#472
opened Jun 12, 2026 by
Zjq9409
Loading…
Add Lunar Lake Xe2 iGPU compatibility report and benchmarks
#342
opened Apr 1, 2026 by
MegaStood
Loading…
fix: VLLM_SKIP_PROFILE_RUN patch for Lunar Lake iGPU profile_run() hang
#340
opened Apr 1, 2026 by
MegaStood
Loading…
docs: GLM-4.7-Flash MLA bug analysis, patches, and MoE investigation for Lunar Lake XPU
#334
opened Mar 25, 2026 by
MegaStood
Loading…
[Platform] Update oneapi base image to 2025.2.2 in platform docker file
#199
opened Dec 18, 2025 by
james-tang17
Contributor
Loading…
doc(README.md): update 1.6 subsection, adding vLLM-1755756176147.json
#44
opened Aug 21, 2025 by
Arcs-ur
Contributor
Loading…
Previous Next
ProTip!
Add no:assignee to see everything that’s not assigned.