2026-10-11 17:11 UTC

Independent evaluations will determine whether Alibaba’s Qwen3.8-Max and smaller Qwen3.8 variants set a competitive new bar for coding, agentic, and cowork workflows among frontier and open-weight models.

state: resolvedheat: lowuncertainty: mediumconvergesscott: highqwen coding-agents open-models local-inferenceAlibaba Qwen

What is this?

Alibaba’s Qwen team announced Qwen3.8-Max as a flagship model aimed at coding, agentic computer use, and long-horizon β€œcowork” tasks, alongside a smaller Qwen3.8-27B variant and a stated plan to release weights. Alibaba reports competitive benchmark results and an approximately 500-turn autonomous run, but the supplied coverage emphasizes that these results are vendor-reported and that Qwen3.8-Max does not consistently lead rival frontier models. Independent evaluationsβ€”especially in real coding harnesses, long-running agent workflows, and local deployment of the smaller modelβ€”are therefore needed to establish whether the release sets a new bar; details about the weight release and 27B hardware requirements remain thin in the supplied snippets.

Why it matters to Scott

Alibaba’s planned open-weight 27B release and coding/agent claims converge with Scott’s model-swappable, locally deployable strategy, while his Model-Plus-Harness Benchmark Unit and trace-backed comparison work explain why vendor benchmarks and a 500-turn run are insufficient. If the reported 17GB footprint and real-harness performance hold, the model could become actionable for his local GPU stack and agent benchmarks; until weights and independent traces appear, it remains a promising candidate rather than a changed conclusion.
ip:concept.model-plus-harness-benchmark-unitip:framework.sovereign-software-assurancedev:concept.trace-backed-agent-comparisondev:concept.hardware-aware-local-inferencedev:project.gamepcradar:concept.open-modelsradar:concept.local-inferenceradar:concept.coding-agent-benchmarksradar:concept.agent-harnessesradar:concept.long-horizon-agents
queries asked of Scott's wikis
  • coding-agent evaluation harnesses and real-world benchmarks
  • long-horizon autonomous coding reliability
  • open-weight frontier model strategy and sovereignty
  • local inference VRAM and deployment economics
  • model versus harness effects on agent performance
  • cowork agents and agentic workflow design

Measured heat

no measured readings yet β€” the hourly heat pass fills this in

How the heat travelled

no chain yet β€” the hourly chain pass fills this in

Evidence (139) β€” ⭐ canonical anchor

sourceobjectauthorscorecomments
🟠 redditQwen3.8-27B announced alongside Qwen3.8-Max
LocalLLaMA
TKGaming_112700647
🟧 echo.x ⭐Announced Qwen3.8-Max alongside the smaller Qwen3.8-27B variant.Alibaba Qwenβ€”β€”
🟠 redditCan't wait to see Qwen3.8-27B
LocalLLaMA
FormOne261522481
🟠 redditDaniel Han of Unsloth validates Qwen3.8-27B will run only 17GB VRAM
LocalLLaMA
quantier1704284
🟠 redditQwen 3.8 morning to you too Dario, 2$ input/ 6$ output per 1M.
singularity
Boring_Aioli79161394106
🟠 redditQweb 3.8-Max Released - #2 in Vision Arena
LocalLLaMA
SnooBunnies83924510
🟠 redditmodel: MTP support for Qwen3-Next by yomaytk · Pull Request #25589 · ggml-org/llama.cpp
LocalLLaMA
jacek20236312
🟠 redditQwen says next week 3.8 will be open weights
LocalLLaMA
Terminator8576024
🟠 redditQwen3.8-Max matches Kimi K3 and DeepSeek V4 Flash
LocalLLaMA
davidthesong566106
🟠 redditQwen3.8-max: 53 on AA index
LocalLLaMA
_AnemicRoyalty_2317
🟠 redditQwen 3.8 Max Artificial analysis scores
singularity
Ill_Distribution85175940
🟠 redditMore Qwen 3.8 sizes coming
LocalLLaMA
appakaradi1344327
🟠 redditQwen3.8-Max Aquarium break simulator
LocalLLaMA
kms_dev2432
🟠 redditQwen3.8-Max: Oneshots, one of the few model that nailed aquarium break simulation alongside opus
LocalLLaMA
kms_dev41
🟠 redditQwen3.8-Max just took #2 in Vision Arena, trailing Claude by only 13 points. Open-source vision is getting crazy close.
LocalLLaMA
kevin_cn_ai149
🟠 redditQwen 3.8 Max improves over Qwen 3.7 Max on the Debate Benchmark: 1462 β†’ 1588. But the average cost per debate increased by 45%.
LocalLLaMA
zero0_one1810
🟠 redditQwen Developers' responses from their recent Twitter/X AMA
LocalLLaMA
pmttyji311118
🟠 redditQwen 3.8 max is really 56 points or benchmaxxed?
LocalLLaMA
ideaofsoul015
🟠 redditGLM/Qwen Appreciation Post
LocalLLaMA
Prestigious_Thing7974725
🟠 redditQwen3.8-2.4T-A95B (aka Qwen3.8-Max) open release time: next wednesday
LocalLLaMA
HugeConsideration211666152
🟠 redditQwen 3.8 max is the fifth on artificial analysis leaderboard.
singularity
Snoo2683716032
🟠 redditQwen 3.8 Max now ranked as best overall model ahead of Opus 5 by Artificial Analysis agentic index
LocalLLaMA
anderspitman1275243
🟧 hnQwen3.8-Max reproduce a research paperdyfang30
🟠 redditWhy China model stays at #2?
LocalLLaMA
Ok-Shower7286028
🟧 hnShow HN: Qwen3.8-Max – Use Qwen Studio and MCP to Code Locally for Freetohid4n20
🟠 redditQwen 3.8-27b coming this week
LocalLLaMA
Bestlife732386292
🟠 redditWho’s ready to bet that Qwen 3.8 27B will be less popular than 3.6 27B in the end?
LocalLLaMA
AdNew5862086
🟠 redditIt's the final countdown, baby! Qwen is out in just over 7 hours!
LocalLLaMA
LegacyRemaster1414204
🟠 redditExact Qwen 3.8 27b release date and time
LocalLLaMA
yuicebox897220
🟠 redditQwen3.8-Max
LocalLLaMA
frontsideair9322
🟠 redditQwen3.8-2.4T-A95B Released
LocalLLaMA
de4dee1559399
🟠 redditQwen 3.8 2.4T is out , no 27b today RIP.
LocalLLaMA
cviperr335038
🟠 redditWe getting today grok 4.6, DeepSeek v4 pro, open source Qwen 3.8 models !!
singularity
Independent-Wind446219915
🟧 hnQwen/Qwen3.8-2.4T-A95BPhilpax692161
🟧 hnQwen3.8-2.4Tmmastrac374
🟠 redditQwen 3.8 27B β€” MTP or DFlash?
LocalLLaMA
mailto_devnull4419
🟠 redditQwen 27b 3.8 release date took down?
LocalLLaMA
EveningIncrease7579147112
🟠 redditI am thankful for the Chinese Model, but what's the deal with Text Mainly and No Multimodal releases?
LocalLLaMA
Hannibalj2ca022
🟠 redditAs I predicted, we get a crippled open-weight version of Qwen 3.8 relative to the API
LocalLLaMA
entsnack024
🟠 redditQwen 3.8Max 2.4t Open Weight NO vision?!?
LocalLLaMA
UltraFOV2450
🟠 redditThe countdown to Qwen3.8-27B starts now!
LocalLLaMA
Ok-Shower7286476134
🟠 redditQwen/Qwen3.8-27B · Official Countdown · Hugging Face
LocalLLaMA
paf113819955
🟠 redditWhile waiting for the release of Qwen3.8-27B, let's try to guess what will happen
LocalLLaMA
Ok-Shower72862895
🟠 redditQwen/Qwen3.8-2.4T-A95B · Hugging Face
LocalLLaMA
techlatest_net015
🟠 redditEXPERIMENT: Qwen3.8-2.4T-A95B running locally on an RTX 5090 + RTX 5060 Ti at ~0.80 tok/s
LocalLLaMA
mossy_troll_846530
🟠 redditFixed Jinja chat template for Qwen 3.5, 3.6, and the new 3.8 release
LocalLLaMA
ex-arman6830580
🟠 redditQwen3.8 2.4T UD-Q1_0 - 178 token generation - 11 min 38s - 0.25 tokens/sec
LocalLLaMA
klicker01613
🟠 redditQwen/Qwen3.8-27B · Upcoming release · Hugging Face
LocalLLaMA
mrinterweb038
🟠 reddit1BIT Qwen 3.8 2.4T a95b (unsloth iQ1_S) (MEDIUM Reasoning)
LocalLLaMA
Ok_Technology_59624723
🟠 redditGetting ready for the big 3.8 drop (strix halo centric but applies widely)
LocalLLaMA
profcuck812
🟠 redditA preliminary Qwen3.8-27B model card is live!
LocalLLaMA
-Cubie-573220
🟠 redditQwen 3.8 Max Is Extremely Good at Knowing What Not to Build
LocalLLaMA
RevealIndividual75672017
🟠 redditOnly 2 Hours left in Qwen3.8-27BQwen3.8-27B Release
LocalLLaMA
techlatest_net046
🟠 redditQwen/Qwen3.8-27B · released
LocalLLaMA
de4dee980294
🟠 redditQwen 3.8 release
LocalLLaMA
Top-Eye-81049253
🟠 redditunsloth/Qwen3.8-27B-GGUF
LocalLLaMA
koloved8210
🟠 redditReturn of the local King! Qwen3.8-27B performance chart
LocalLLaMA
HugeConsideration211397
🟠 redditQwen3.8-27B Released!
LocalLLaMA
Tall_Abrocoma_3533413
🟠 redditIT'S OUT
LocalLLaMA
Certain-Cod-14042202681
🟠 redditQwen3.8-2.4T-A95B Oneshots
LocalLLaMA
kms_dev00
🟠 redditQwen 3.8 27b is here
singularity
song9152397
🟧 hnQwen3.8-27Bmfiguiere20779
🟧 hnQwen 3.8 27B is out: open weights, best local dense model yeterdaltoprak1387778
🟠 redditQwen3.8-27B Serving Configs: DGX Spark vLLM NVFP4 and RTX 4090 llama.cpp GGUF
LocalLLaMA
erdaltoprak32
🟠 redditQwen 3.8 Q8 β€” reasoning trace review (code review / bug-finding task)
LocalLLaMA
Ok-Shower7286811
🟠 redditQwen3.8 Benchmarks Converted to Charts
LocalLLaMA
Tccybo9014
🟠 redditQwen3.8-27B is identical to Qwen3.6-27B!
LocalLLaMA
Course_Latter1079179
🟠 redditQwen 3.8 27B has caveman thinking!
LocalLLaMA
My_Unbiased_Opinion3314
🟠 redditOpen labs are finally embracing the power of continued post training
LocalLLaMA
Daniel_H2124023
🟠 redditQwen 3.8 27B links
LocalLLaMA
FrankWanders02
🟧 hnUnsloth Qwen3.8-27B GGUF filesapitman451
🟠 redditNInfer day0 support for Qwen3.8 27b: ~200 tok/s generation, with tons of engine improvments
LocalLLaMA
FormOne26158268
🟠 redditQwen3.8-27B - llm-decode-bench - BF16 TP2 (2x RTX 6000) - vLLM nightly
LocalLLaMA
Maleficent_Bridge_4165
🟠 redditQwen 3.8 27B - Aquarium Burst Sample Test
LocalLLaMA
live4evrr17218
🟠 redditQwen 38 still seem to have that random stop behavior?
LocalLLaMA
T_rex2700326
🟠 redditOptimized Dual 3090 Qwen3.8 Quant
LocalLLaMA
luedtek1414
🟠 redditSGLang support for Qwen3.8-27B: 200+ tok/s on 5090, 38 tok/s on DGX Spark (NVFP4 + DSpark)
LocalLLaMA
unseenmarscai4829
🟠 redditQwen 27b 3.6 vs 3.8 on web design tasks - first thoughts?
LocalLLaMA
ShadyShroomz4116
🟠 redditQwen 3.8 27B built from scratch with PyTorch and a tiny agentic harness
LocalLLaMA
No-Compote-679481
🟠 redditQwen3.8-27B running on a 12GB RTX 5070 Ti laptop β€” around 4.5 tok/s with 80% MTP acceptance
LocalLLaMA
CoffeeToCode99149
🟠 redditQwen 3.6 vs 3.8 MTP Sweep comparison 27B-FP8
LocalLLaMA
Radiant_Condition86158
🟠 redditQwen3.8-27B only 5 tk/s - What's the best config for 8GB VRAM + 32GB RAM?
LocalLLaMA
SoAp9035022
🟠 redditQwen 3.8 27b is like Opus 4.6 on your machine
LocalLLaMA
Odd_Tumbleweed57416065
🟠 redditI told qwen 27 3.8 and 3.6 to make an api interface for Grok
LocalLLaMA
Whole_Alternative_182523
🟠 redditQwen3.8-27b, Benchmaxxxed to the Maxxx
LocalLLaMA
KitchenAmoeba4438071
🟠 redditQwen 3.8 27B tends to overthink on html+svg task a lot. And does so rather well.
LocalLLaMA
hurdurdur722
🟠 redditQwen3.8-27B Early Performance Report on 2x RTX 3060 12GB
LocalLLaMA
anderspitman2110
🟠 redditQwen 3.8 27b sometimes thinks in caveman
LocalLLaMA
Eyelbee6235
🟠 redditSo, does Qwen 3.8 27B still have the huge "but wait"-ing itself to death overthinking problem as 3.5 / 3.6 Qwens?
LocalLLaMA
ZootAllures911112104
🟠 redditQwen3.8-27B on 2x3090 β€” 200K context with f16 KV, vision and thinking
LocalLLaMA
Sisuuu410
🟠 redditThe difference between "medium" and "xhigh" reasoning effort for Qwen3.8-27B is actually insane.
LocalLLaMA
SarcasticBaka20198
🟠 redditFixed/improved Jinja chat template for Qwen 3.8
LocalLLaMA
Chromix_4315
🟠 redditQwen3.8-27B Q6_K at 128K on a single 32GB GPU
LocalLLaMA
WSTangoDelta2534
🟠 redditNinfer-3090
LocalLLaMA
mrmontanasagrada3037
🟠 redditRetroCraft - Qwen 3.8 27B Q8, one shot with exact performance data on dual 3090s.
LocalLLaMA
No-Statement-000115624
🟠 redditMy new "resident" models Qwen3.8:27b and Muse-Glimmer
LocalLLaMA
NicolaZanarini533210
🟠 reddit3.8 27B is bench maxed because reasoning is default to xhigh in the chat template Jinja file.
LocalLLaMA
jinnyjuice05
🟠 redditunsloth/Qwen3.8-27B-NVFP4 was released - 22 GB model size
LocalLLaMA
TheLocalLab114
🟠 redditQwen 3.8 is better than Muse Glimmer
LocalLLaMA
Numerous_Mulberry514726
🟠 redditA hunch: Qwen3.8-27B's general knowledge got pruned (good, if true)
LocalLLaMA
bonobomaster233124
🟠 reddit3.8 27b dual 3060s recipe 50tk/s
LocalLLaMA
Ecstatic-Wash-76671329
🟠 redditllama.cpp: reasoning_effort not forwarded to chat template (Qwen3.8-27B)
LocalLLaMA
fbms285
🟠 redditQwen endless looping issue and possible fix
LocalLLaMA
ParvusNumero1320
🟧 hnQwen 3.8 27B vs. Nemotron 3.5 vs. Muse Glimmer on an RTX 50902a0c4010
🟠 redditrdtand/Qwen3.8-27B-PrismaAQUA-5.5bit-vllm · Hugging Face
LocalLLaMA
Fragrant_Scale645610
🟠 redditQwen 3.8 27B - Note the new recommended sampling parameters (from the official HF page)
LocalLLaMA
Thrumpwart7615
🟠 redditQwen 3.8 - 27B is a game changer
LocalLLaMA
Potential_Block4598637279
🟠 redditwe quantized Qwen3.8-27B and compared it with community GGUFs on 4x RTX 5090!
LocalLLaMA
Top-Eye-8104167
🟠 redditMost of the "27B on 16 GB" posts I see land on a Q2. I wanted to know how far up the quant ladder I could go and still keep real context and real speed, so I benchmarked UD-Q3_K_XL (~3.9 bpw, 12.51 GiB of weights) properly instead of eyeballing it.
LocalLLaMA
No-Head2511010
🟠 redditTesting out DSpark GGUF for Qwen3.8 27B
LocalLLaMA
Hefty_Wolverine_55313
🟠 redditQwen 3.8 27B on AMD - 24 t/s on Ryzen AI Max+ 395 & 51 t/s on Radeon AI PRO R9700
LocalLLaMA
pmttyji1710
🟠 redditTry out this "high" reasoning mode for 27B (tested on VLLM)
LocalLLaMA
TokenRingAI6429
🟠 reddit16gb vram users, how has qwen3.8b at q3 been? Is it worth?
LocalLLaMA
Adventurous-Gold6413537
🟠 redditQwen 3.8 27b Q5 - Q6 comparison on 24vram with a three.js 3D arena game .
LocalLLaMA
source-drifter62
🟠 redditQwen 3.8 35BA3B spotted
LocalLLaMA
BazzyIm1146358
🟠 redditQwen 3.8 27b UD Q8_K_XL - Racing Game
LocalLLaMA
BlackBeardAI1711
🟠 redditQwen3.8 surprises from more extended testing
LocalLLaMA
KitchenAmoeba4438016
🟠 redditQwen3.8-27B EXL3: 0.0074 KLD vs 0.0950 for Unsloth NVFP4 at slightly lower VRAM
LocalLLaMA
malaiwah1235
🟠 redditQwen3.8-27B / 3.6-27B Comparison
LocalLLaMA
thomble06
🟠 redditWhat is the highest quality version of Qwen3.8-27B I can run on 48 GB of vram?
LocalLLaMA
Valuable-Run2129146
🟠 redditQwen3.8: much thinking for nothing
LocalLLaMA
WhatererBlah555768
🟠 redditQwen3.8 27B vs Qwen3.6 27B vs Qwen3.5 27B, a slight improvement in oneshotting ability across generations.
LocalLLaMA
kms_dev7620
🟧 hnQwen3.8-27B – Release Day Demosyogthos30
🟠 redditmy harness *somehow* makes qwen3.8 27b barely think. i dont even know how!
LocalLLaMA
rosie254242
🟠 reddit3.8 27B xhigh, we gotta talk about it
LocalLLaMA
Old-Sherbert-4495530
🟠 redditQwen 3.8 27B caveman thinking only working on xhigh and no tools?
LocalLLaMA
Borkato210
🟧 hnQwen 3.8 27B at 2x speed on a 5090husky820
🟠 redditQwen 3.8 27B Q8 faster than Q6 w/ MTP on Apple Silicon using llama.cpp and lmstudio gguf
LocalLLaMA
jcmyang10
🟠 redditTested Qwen3.8 27B locally on coding with OpenCode & agentic work
LocalLLaMA
curiousily_01
🟠 redditMoE with MTP?
LocalLLaMA
chibop123
🟠 redditclub-5060ti refresh: tested RTX 5060 Ti presets, a proper high-context harness, and Qwen3.8 27B
LocalLLaMA
do_u_think_im_spooky3120
🟠 redditQwen3.8-27B vs Qwen3.6-27B writing ray-tracers in BASIC
LocalLLaMA
Ok-Breakfast187883898
🟠 redditCan we moderate the lying?
LocalLLaMA
DustNearby2848098
🟠 redditAnyone managed to get Qwen 3.8 27B running smoothly on vLLM? Can't get rid of endless thinking
LocalLLaMA
germangrower69918
🟠 redditUsers of Qwen 3.8 27b on the strict halo, REPORT!
LocalLLaMA
Acrobatic_Stress1388028
🟠 redditOh, qwen 3.8, it's too perfect!
LocalLLaMA
Ok-Shower728639
🟠 redditIf you are at the lowest budget, which you can think of.Which hardware would you recommend to run? qwen 3.8 27b oWith like 50 tokens per second. I currently have a RTX 5070 Ti.
LocalLLaMA
InternationalGap369864146
🟠 redditI'm still running Qwen 3.5 122B. Should I switch to Qwen 3.8 27B?
LocalLLaMA
MackThax51113
🟠 redditRunning Qwen3.8 27B on Mac
LocalLLaMA
koc_Z3212

Interpretation history

Decision trace