Skip to content

Keep qwix out of tpu-inference's fused MoE kernel and pre-quantize weights - #5136

Open
sierraisland wants to merge 1 commit into
AI-Hypercomputer:mainfrom
sierraisland:sierraq/vllm-fused-moe-qwix-boundary
Open

Keep qwix out of tpu-inference's fused MoE kernel and pre-quantize weights#5136
sierraisland wants to merge 1 commit into
AI-Hypercomputer:mainfrom
sierraisland:sierraq/vllm-fused-moe-qwix-boundary

Keep qwix out of tpu-inference's fused MoE kernel and pre-quantize we…

7b8a2e9
Select commit
Loading
Failed to load commit list.
Google CLA / cla/google succeeded Sep 4, 2026 in 16s

✅ All contributors are covered under a CLA with Google

See https://cla.developers.google.com/ for more info about Google's Contributor License Agreement (CLA).

ℹ️ Googlers: Go here to view more details and manage scans for this pull request.

Details

The following contributors were found for this pull request:

7b8a2e9 Author: @sierraisland <shen********1225​@gmail.com>

(Only the first commit for a unique contributor is listed.)