2026-10-11 18:04 UTC

Independent evaluations will determine whether Qwen3.8-27B delivers frontier-competitive tool use, visual QA, and reverse-engineering performance in locally run agent workflows.

state: resolvedheat: lowuncertainty: highconvergesscott: mediumqwen38-27b local-inference coding-agents agent-harnessesQwen

What is this?

Alibaba’s Qwen3.8-27B is a roughly 27B-parameter, deployment-oriented model for local coding and agent workflows, with native image/video understanding, configurable reasoning, and a reported 262,144-token context window. Launch coverage and user anecdotes claim strong tool use and reverse-engineering performance on consumer hardware, potentially approaching closed frontier systems. However, the supplied sources emphasize that benchmark claims still require independent validation, and they do not yet establish frontier competitiveness in real-world visual QA, tool use, or reverse-engineering tasks.

Why it matters to Scott

The release creates a direct dated-receipts test of Scott’s model-plus-harness evaluation doctrine and could materially alter his Model Barbell if a local 27B model can reliably take reverse-engineering, visual, and tool-using work previously reserved for frontier systems. It also bears directly on his gamepc local-serving substrate and Ask agent, although the supplied anecdotes are not yet strong enough to establish that routing change or contradict the barbell.
ip:concept.model-barbellip:concept.model-plus-harness-benchmark-unitip:concept.model-perishabilitydev:concept.trace-backed-agent-comparisondev:project.gamepcdev:project.askradar:concept.local-inferenceradar:concept.agent-evaluationradar:concept.agent-harnessesradar:concept.multimodal-modelsradar:deepseek-v4-flash-harness-efficiencyradar:meta-muse-open-weights-local-inference
queries asked of Scott's wikis
  • local models replacing frontier APIs in coding agents
  • independent evaluation of agent tool use
  • local inference economics for agent workflows
  • model routing between local and frontier systems
  • multimodal local agents and visual QA
  • agent harness effects on model performance

Measured heat

no measured readings yet β€” the hourly heat pass fills this in

How the heat travelled

no chain yet β€” the hourly chain pass fills this in

Evidence (79) β€” ⭐ canonical anchor

sourceobjectauthorscorecomments
🟠 reddit ⭐Qwen 3.8 27B with xhigh thinking is awesome
LocalLLaMA
awitod134
🟠 redditI gave Qwen 3.8 27B a reverse-engineering job I assumed needed a frontier model, and it finished in 30 minutes
singularity
yogthos67347
🟠 redditBenchmark results: what is the best and fastest engine to run Qwen3.8-27B on macOS
LocalLLaMA
ex-arman681921
🟠 redditWeird speed gap between LM Studio vs raw llama.cpp + questions on Reasoning Effort (Qwen3.8-27B on dual GPU)
LocalLLaMA
MkGod11
🟠 redditNew qwen3.8:27b on a 39k line C to single-file HTML / three.js port
LocalLLaMA
codehamr45898
🟠 redditQwen 3.8 27b helped me with something unique that Opus 4 couldn't - Firmware + Software preservation and emulation on an early 2000's ARM based POS system
LocalLLaMA
maxwell32111518
🟠 redditWe quantized Qwen 3.8 27B and compared the quants on an RTX 6000
LocalLLaMA
Fun-Meaning-6474233109
🟠 redditQwen3.8-27B β€” One Week Later: The r/LocalLLaMA + r/LocalLLM Verdict
LocalLLaMA
Jonathan_Rivera09
🟠 redditMy Qwen3.8 Setup So Far
LocalLLaMA
BopSupreme05
🟠 redditreasoning-budget for Qwen3.8-27b
LocalLLaMA
Thin_Pollution8843214
🟠 redditQwen3.8-27b is making my messages up
LocalLLaMA
Thin_Pollution8843014
🟠 redditQwen 27B 3.8 quants: How low can you go?
LocalLLaMA
jeremyckahn1034
🟠 redditQwen 3.8 27B, just wanted to say thanks to you guys
LocalLLaMA
sshwifty26085
🟠 redditThose of you running qwen 3.8 27b with 16GB VRAM, what pi.dev plugins or skills are you using?
LocalLLaMA
radlinsky1120
🟠 redditR9700 AI Pro TP=2 Qwen3.8-27B-FP8 low speed? Need Advice.
LocalLLaMA
YehowaH120
🟠 redditQwen 3.8 27B Aider score
LocalLLaMA
Baldur-Norddahl4827
🟠 redditGetting ~11.7 tok/s from Qwen3.8 27B across an RTX 4070 Ti and M5 MacBook Air. Any ideas to push it further?
LocalLLaMA
zannix018
🟠 redditLimits to Harnessing Qwen3.8-27B
LocalLLaMA
norenEnmotalen015
🟠 redditOptimizing Qwen 3.8 27B FP8 or BF16 on two RTX 6000 Pro?
LocalLLaMA
EkbatDeSabat060
🟠 redditThis is what Qwen 3.8 27b is capable of
LocalLLaMA
BlackBeardAI13554
🟠 redditQwen 3.8 27B in 9th position on code arena. Gemma 4 31B is 80th.
LocalLLaMA
tarruda638172
🟠 redditHow would you test Qwen3.8-27B inside a coding agent?
LocalLLaMA
Binary_orchid75
🟠 redditqwen38-27b-rtx3090 (https://github.com/syv-ai/qwen38-27b-rtx3090) is extremely good with deepseek harness.
LocalLLaMA
politefella01824
🟠 redditPlanning to spend ~$100 benchmarking differnet Qwen3.8-27B quants and kv cache and looking for input before I start
LocalLLaMA
m_mukhtar2342
🟠 redditIs Qwen3.8-27B half baked?
LocalLLaMA
BitterProfessional7p021
🟠 redditQwen 3.8 27B thinks too much, so...turn it off?
LocalLLaMA
N34257019
🟠 redditThe journey of letting Qwen 3.6/3.8 autonomously coding a c compiler.
LocalLLaMA
Naiw803619
🟠 redditQwen 3.8 27b with tools and directed search on a non-coding professional suite
LocalLLaMA
offgridai116
🟠 redditToday I merged the first feature branch written entirely by my 4060Ti 16GB!
LocalLLaMA
o0genesis0o1925
🟠 redditDynamic quants of Qwen3.8 27B not using caveman thinking
LocalLLaMA
Glad_Claim_628739
🟠 redditQwen 3.8 Flash Next day 0 support from unsloth
LocalLLaMA
jacek2023736183
🟠 redditQwen3.8-27B on an IGX Thor with an RTX PRO 6000 Blackwell (Max-Q)
LocalLLaMA
ahstanin53
🟠 redditQwen3.8-27B on an RTX 3060 + RTX 2060
LocalLLaMA
sheriffoftiltover215
🟠 redditQwen 3.8 27b has ThreeJs locked down.
LocalLLaMA
Both_Opportunity53276424
🟠 redditPeak Portable Personal Datacenter
LocalLLaMA
Special-Wolverine6143
🟠 redditQwen3.8-27B IQ3_XXS wrote a correct multilayer TMM on a 16 GB Quadro β€” after 100 minutes, 3 compactions, and 108k output tokens
LocalLLaMA
1000_bucks_a_month3626
🟠 redditAnyone directly compare 3.8 27B at INT4 vs INT8?
LocalLLaMA
OvertaxedOne711
🟠 redditFully quantized NVFP4 Qwen3.8-27B with QUASAR QAD
LocalLLaMA
arty_photography17188
🟠 redditGetting Qwen3.8-27B with decent speed on my 4080 with 16Gb card
LocalLLaMA
Apprehensive_Bar66092034
🟠 redditOpenCode with Qwen3.8-27B for Small Games or Browsing the Web With 16GB VRAM
LocalLLaMA
Due-Project-75073210
🟠 redditA 27b model beating latest frontier models was not on my 2026 bingo card
LocalLLaMA
Gohab200126772
🟠 redditwe made Qwen 3.8 27b MLX vision quants and compared them against other popular community publishers (lm-studio, lukaskremla, mlx-community and etc)
LocalLLaMA
Top-Eye-8104112
🟠 redditA minecraft clone I fully vibecoded with Qwen3.8-27b Q4
LocalLLaMA
liright30968
🟠 redditForget the Pelican, it's Weevil-Time! / Benchmaxxing-Proof SVG and Vision Benchmark
LocalLLaMA
bonobomaster15158
🟠 redditQwen 3.8 Flash Next: Beating DS V4 Flash at half the parameters, stronger than Opus 4.6
singularity
elemental-mind18527
🟠 redditWhoever the fuck predicted we would have gpt 5.5 performance in coding on consumer hardware a couple months ago now, i applaud you
LocalLLaMA
GrokiniGPT714180
🟠 redditGLM 5.3 FLASH vs QWEN 3.8 FLASH NEXT
LocalLLaMA
Nota_ReAlperson1214
🟠 redditBenchmarking Qwen3.8 27B quantizations: 4-bit holds up, 1-bit collapses
LocalLLaMA
pmigdal24547
🟠 redditCompared Qwen 3.8 27B community quants on RTX 6000 vs Claude Opus 4.6
LocalLLaMA
Top-Eye-81045622
🟠 redditQwen3.8 27B C8 at 972 TG / 5,680 PP on 4x MI100 rig ($6.5k) using my new INT8 vLLM fork
LocalLLaMA
1ncehost3230
🟠 redditIs Qwopus3.5-9B-Coder still worth using?
LocalLLaMA
Ammargok014
🟠 redditQwen3.8 Garbage Outputs after a few hours of use
LocalLLaMA
trashacct383014
🟠 redditQuestions on optimism speed/intelligence on this rig
LocalLLaMA
Ambitious_Fold_287418
🟠 redditopenrouter is a hop. local 27b on the mac is on-prem for the tool loop
LocalLLaMA
conifer_v1105
🟠 redditOver 200k context on 16GB VRAM with Qwen 3.8 27B UD-IQ3_XXS
LocalLLaMA
abskvrm13046
🟠 redditQwen3.8 27b: UD Q_K_XL vs W4A16-AutoRound
LocalLLaMA
Lower-Ad6101828
🟠 reddityall are sleeping on qwen 3.8 27b q2 + q2 dflash + q5 kv
LocalLLaMA
Square_Light144111594
🟠 redditTepid take: qwen 3.8 is less suitable as a daily-driver compared to 3.6
LocalLLaMA
mailto_devnull034
🟠 reddit5090 + 96GB RAM, any better choice than Qwen3.8-27B for coding?
LocalLLaMA
a9udn9u1754
🟠 redditDual 5060 ti running Qwen 3.8 27b UD 3.0 Q4_K_XL
LocalLLaMA
SellToOpen811
🟠 redditThe RTX 3060 12GB: An unsung hero of the current Local AI Climate (3.8 27B 30t/s)
LocalLLaMA
I_Play_Zed066
🟠 redditQwen3.8-27b q8 KV cache does seem to actually hurt model performance
LocalLLaMA
maddie-lovelace5774
🟧 hnRun Qwen3.8 27B locally: real numbers from my Mac Studiospeckx12696
🟠 redditQwen3.8 27B int4 with Dflash2 at 165t/s and 18M kv cache pool on dual 3090
LocalLLaMA
Old_Ad_60331610
🟧 hnHow to run Qwen3.8-27B on a single 16GB cardCuriositry10
🟠 redditLocal agentic coding Benchmark : Qwen3.8-Flash-Next NVFP4 vs 27B (and the others...)
LocalLLaMA
WonderRico2921
🟠 redditI benchmarked 9 open models on spotting fake sources during agentic search (DeepSeek V4, Qwen 3.8, Nemotron 3 Ultra)
LocalLLaMA
RevealIndividual75678516
🟠 redditIs Qwen 3.8 27B more sensitive than 3.6 to quantization in general?
LocalLLaMA
draetheus06
🟠 reddit[Release] SOTA GGUFs for Qwen3.8-27B: GSQ-RCO at 2.5 to 3.0 bpw
LocalLLaMA
Loginhe10827
🟠 redditSpeed boost for Qwen 3.8 27B on Apple Silicon
LocalLLaMA
koc_Z311
🟠 reddit~ 2x Speed Boost for Qwen3.8 27B on Apple Silicon
LocalLLaMA
koc_Z32618
🟠 redditdgx sparks and new models my tests and results
LocalLLaMA
jtsaint33357
🟠 redditQwen 3.8 27B at 50 tok/s with 100k Context on a 16GB GPU! (beellama.cpp)
LocalLLaMA
qaf23603182
🟠 redditQwen3.8-27B-QAT-Q2: Surprisingly solid for quick answers, but unusable for long context work
LocalLLaMA
kirisoraa012
🟠 redditSpeeds of a local Qwen3.8-27B on an M4 Max with 5 runtime/quant stacks and context from 32K to 256K
LocalLLaMA
vitordeas33
🟠 redditQwen3.8-27B thinking xhigh Vs. thinking off - Apple M5 Max
LocalLLaMA
DerTomsn724
🟠 redditBenchmarking Qwen3.8-27B at Q4/Q5/Q6 on a laptop GPU + eGPU of a completely different tier
LocalLLaMA
CoffeeToCode9910
🟠 redditNew local claude code?
LocalLLaMA
Square_Light144108
🟠 redditan unscientific qwen 3.8 flash next and glm 5.3 flash comparison
LocalLLaMA
nomorebuttsplz14338

Interpretation history

Decision trace