-
Notifications
You must be signed in to change notification settings - Fork 66
Pull requests: intel/llm-scaler
Author
Label
Projects
Milestones
Reviews
Assignee
Sort
Pull requests list
docker: upgrade dev vLLM image to 0.26.0
#614
opened Aug 13, 2026 by
liu-shaojun
Contributor
Loading…
Omni: Account whole LoRA weights in ComfyUI memory budgets
#613
opened Aug 13, 2026 by
xiangyuT
Contributor
Loading…
docs(omni): add HiDream-O1, Krea2 Turbo, Boogu-Image, Mage-Flow and LTX-2.3 to supported models table
#592
opened Aug 4, 2026 by
KristianZeng
Contributor
Loading…
SGL: optimize Qwen3.6 27B and 35B A3B online fp8 quantization performance on BMG w/o XPU Graph and resolve acc issue.
#577
opened Jul 29, 2026 by
lalalapotter
Contributor
Loading…
[XPU] Harden fused GDN recurrent state updates
#570
opened Jul 27, 2026 by
gc-fu
Contributor
Loading…
Fix #533: guard out_proj.weight access in GDN out-projection ESIMD probes
#557
opened Jul 22, 2026 by
joaovgaraujo
Loading…
refactor(omni): multi-stage Docker build with decoupled builder/runtime
#549
opened Jul 17, 2026 by
KristianZeng
Contributor
•
Draft
page_attn: make global-max reduction order-invariant
#535
opened Jul 12, 2026 by
taste-software
Loading…
Fix Qwen3.5/3.6 load_weights stacked-mapping name mutation (gate_gate_up_proj / qkqkv_proj)
#475
opened Jun 12, 2026 by
bongmiin
Loading…
vllm: add MiniCPM-V 4.6 support (MiniCPMV4_6ForConditionalGeneration)
#472
opened Jun 12, 2026 by
Zjq9409
Loading…
Add Lunar Lake Xe2 iGPU compatibility report and benchmarks
#342
opened Apr 1, 2026 by
MegaStood
Loading…
fix: VLLM_SKIP_PROFILE_RUN patch for Lunar Lake iGPU profile_run() hang
#340
opened Apr 1, 2026 by
MegaStood
Loading…
docs: GLM-4.7-Flash MLA bug analysis, patches, and MoE investigation for Lunar Lake XPU
#334
opened Mar 25, 2026 by
MegaStood
Loading…
Previous Next
ProTip!
no:milestone will show everything without a milestone.