Qwen's 2.4T flagship goes open-weights next week — and DeepSeek codes at 3 cents a test
Qwen3.8-Max opens up, DeepSeek matches Gemini 3.6 Flash at 3¢ a test, 'don't be a meat proxy', Cursor hides dollar costs, qm's shared agents, Fable-os, and Copilot's Gemini deadline.
Alibaba shipped Qwen3.8-Max — a 2.4-trillion-parameter coding flagship whose weights go public next week, a first for its Max class — while DeepSeek's updated V4-Flash posted Gemini-3.6-Flash-level scores at three cents a test: the price of frontier-class coding keeps collapsing. Meanwhile the weekend's biggest dev conversation was an essay asking people to stop pasting Claude's answers into Slack, a 'multiplayer' agent harness cleared 8,600 GitHub stars in five days, and Cursor removed dollar costs from its usage page on purpose. Plus: GitHub's Gemini deadline landed, and a hobbyist handed Claude an entire kernel. The daily pulse of AI coding tools — what shipped, what matters, what's next.
• Qwen3.8-Max: 2.4T flagship, open weights next week
• DeepSeek V4-Flash matches Gemini 3.6 Flash at 3¢/test
• 'Don't be a meat proxy' — stop relaying Claude verbatim
• qm: team-steered shared agents, 8,600 stars in five days
• Cursor removed dollar costs — 'deliberate design'
• Copilot dropped Gemini 2.5 Pro and 3 Flash Thursday
• Fable-os: Claude runs the kernel, writes its own driver
Frontier-class coding models keep getting cheaper
Alibaba ships Qwen3.8-Max, a 2.4T coding model — open weights next week
Qwen's own pitch is 'a new bar for coding and cowork,' priced at $2 in / $6 out per million tokens — and for the first time for a Max-class flagship, the weights (the model itself, downloadable to run on your own hardware) go public next week, with the small Qwen3.8-27B going open alongside it. One honest caveat from Qwen's own benchmark table: it still trails Fable 5 and GPT-5.6 Sol on several coding rows — the news is the price and the openness, not a new crown.
📢Meet Qwen3.8-Max — our most capable model to date.
— Qwen (@Alibaba_Qwen) August 3, 2026
Next week, the open weights of Qwen3.8-Max will be released, and Qwen3.8-27B is also going open-weights to meet you all!🎉
Qwen3.8-Max, a new bar for coding and cowork at 2.4T parameters:
- Autonomous coding: 10+ days of… pic.twitter.com/e3YFj2hqcT
Qwen3.8-27B announced alongside Qwen3.8-Max
by u/TKGaming_11 in LocalLLaMA
Link: Qwen's announcement
DeepSeek's updated V4-Flash matches Gemini 3.6 Flash — at 3 cents a test
Artificial Analysis scores the 0731 update at 50 on its Intelligence Index — a 10-point jump that puts it level with Gemini 3.6 Flash — and Reuters, citing the same firm, puts its average benchmark cost at 3 cents a test, versus 86 cents for Kimi K3 and $1.86 for GPT-5.6 Sol; the weights are on Hugging Face under MIT. The open-weights crown is still Kimi K3's (57 on the same index) — and the weekend's wildest K3 story is WASTE, from the SQLite author's company, which runs the full 2.78-trillion-parameter model on a 64 GB MacBook at about half a token a second by streaming expert weights off the SSD: the maintainer's own numbers, experts re-quantized to 3-bit, and nobody who didn't build it has run it yet.

Read: news.ycombinator.com/item?id=49123386
Link: Artificial Analysis' writeup
Working with agents: relaying output, sharing sessions
'Don't be a meat proxy': HN's top essay says stop relaying Claude's answers
The argument, in the author's words: 'I can talk to Claude myself… I don't need a meat proxy in between' — prompt AI all you want, but 'read it, understand it, validate it, and then write a response in your own words.' It cleared 1,000 points and 440-plus comments on Hacker News today — if your team's Slack has an AI-paste habit, this is the link that will circulate.
Link: the essay
qm lets a whole team steer shared coding agents — 8,600 stars in five days
Most agent setups are one developer, one session; qm's pitch is 'multiplayer' — several people watching and steering the same agents at once. It launched Wednesday and sat at 8,615 GitHub stars plus a 668-point HN thread as of Monday evening — a launch pitched at the team, not the individual developer, as the unit of work.
Link: the repo
What moved in your billing pages and model menus
Cursor removed dollar costs from its usage page and CSV — 'deliberate design'
Cursor staff confirmed it in the forum: self-serve plans (Teams included) now see tokens only — 'The Spend metric and Cost column were removed, and the Usage CSV no longer contains dollar costs' — and there is no toggle back: 'This is deliberate design.' GitHub, the same week, retired its Copilot Billing Preview app (today) in favor of more dollar detail in native billing settings — two vendors moving in opposite directions on whether you get to see what a request cost.

Link: the forum thread
GitHub cut Gemini 2.5 Pro and 3 Flash from Copilot on Thursday
The deprecation announced July 2 landed on schedule across every Copilot surface — chat, inline edits, ask and agent modes, completions — with GitHub pointing at Gemini 3.1 Pro (Preview) and Gemini 3.6 Flash as replacements (not the 3.5 Flash the original announcement named). If a workflow, eval harness, or org model policy pointed at either retired model, it broke Thursday; Copilot Enterprise admins may need to enable the alternatives under their model policy.

Link: GitHub's changelog
From the builders
Fable-os gives Claude the kernel — in its demo it writes a sound driver
It's a real C kernel where the only interface is an agent with full kernel privileges — no shell, raw syscalls as its tools — and the launch demo has it noticing it lacks audio, finding the (emulated) sound card, and writing itself a driver. The usual launch-post caveat applies (nobody who didn't build it has run it, and the thread's top technical objection calls it 'an agent harness that uses sudo for everything') — but the 211-comment thread's most-upvoted line is the real artifact: 'Yes, I did in fact delete your load bearing kernel environment, and you're right to push back on that.'
I built a real self-evolving operating system: Fable-os
by u/robi0t in ClaudeAI
Link: the repo
That's Monday. If this kept you ahead of the group chat, forward it to the teammate who always hears it a day late — and if we missed something or got something wrong, reply and say so; corrections make the next issue.
Also worth your time
• Karpathy's viral 'Pelican' post (588 points of HN discussion)
• Simon Willison on the new stateless MCP spec — plus two new tools built on it
• JFrog: a critical CVE was issued for a hallucinated SQLite vulnerability
• WSJ: how OpenAI fell behind Anthropic by prioritizing chatbots over coding tools
The New Way is human-curated — a person picks every story. The summaries are written with AI (Claude) and reviewed before we hit send.

