Measured trade-offs of three REAP-pruned Qwen3.8-Flash-Next MLX builds on Apple Silicon
MensaProdigy · reddit · 2026-09-22
MensaProdigy released three REAP-pruned Qwen3.8-Flash-Next MLX builds with a decision boundary based on measured runs: REAP-384 oQ5e scores highest (HumanEval 97/100, LiveCodeBench 62/100, 72.8 GB peak on M4 Max 128GB); oQ4e saves 11.3 GB and generates faster; REAP-288 oQ4e needs only 46 GB. A comparison with Jundot's oQ4e MTP package is presented as context, not a controlled benchmark. All scores are 100-question stratified samples.
More from Infra
- Meta Muse gives every user a cloud VM, and 100M users would mean a chip problem — AccBalanced · 2026-09-22
- DeepSeek bets next model on Huawei chips after Ascend 910C training run failed — soumitrashukla9 · 2026-09-22
- Inference era shifts to 1-30MW micro datacenters as OpenAI and Anthropic chase distribution — AccBalanced · 2026-09-22
- World model startup Reactor raises $59M to build the developer platform for world models — buckymoore · 2026-09-22
- Meta unveils Petal, first petabit-class transoceanic subsea cable, live in 2029 — Meta_Engineers · 2026-09-22
- One of the biggest AI launches ever stayed up under crazy load, out-uptimeing Anthropic — hardimanjames · 2026-09-22