TWITTER_POST

sudoingX recommends Qwopus as a strong locally hosted coding model on consumer…

Brief

sudoingX recommends Qwopus as a strong locally hosted coding model on consumer hardware after a full day of testing on an RTX 3090. The post argues that while base Qwen 3.5 MoE wins on raw throughput, Qwopus is a better fit for agentic coding workflows in Claude Code because its distilled reasoning behavior, native thinking mode, and lack of Jinja failures make the model cooperate more smoothly with the agent.

Source evidence

title: @sudoingX: spent the entire day testing Qwopus (Claude 4.6 Opus distilled into Qwen 3.5 27B...
author: sudoingX
contenttype: twitterpost
published: 2026-03-07T11:03:47+00:00
source_url: https://x.com/sudoingX/status/2030344904774402115/photo/1

word_count: 213

Tweet by @sudoingX

spent the entire day testing Qwopus (Claude 4.6 Opus distilled into Qwen 3.5 27B) on a single RTX 3090 through Claude Code. this is my new favourite to host locally. no jinja crashes. thinking mode works natively. 29-35 tok/s. 16.5 GB. the harness matches the distillation source and you can feel it. the model doesn't fight the agent. my flags: llama-server -m Qwopus-27B-Q4KM.gguf -ngl 99 -c 262144 -np 1 -fa on --cache-type-k q40 --cache-type-v q40 if you want raw speed, base Qwen 3.5 MoE still wins at 112 tok/s. but for autonomous coding where the model needs to think, wait for tool outputs, and selfcorrect without stalling, Qwopus on Claude Code is the cleanest setup i've found on this card. i want to see what everyone else is running. drop your GPU, model, harness, flags, and tok/s below. doesn't matter if it's a 3060 or a 4090, nvidia or amd. configs help everyone. let's push these cards to their ceilings. let's make this thread the reference.


Posted: 2026-03-07T11:03:47.000Z
Engagement: 2454 likes, 202 retweets, 111 replies


Quoting @sudoingX

Qwopus on a single RTX 3090. Claude Opus 4.6 reasoning distilled into Qwen 3.5 27B dense, running through Claude's own coding agent (claude code). 29-35 tok/s with thinking mode on.

the jinja bug that kills thinking on base Qwen doesn't carry over. harness and model matched. x.com/sudoingX/statu…