
This is incredible -- Qwen 3.8 Max is seriously revolutionary.
A new open-source model that beats both Claude Fable and Opus 4.8 and GPT-5.6 Sol -- despite being several times cheaper. It's even cheaper than Kimi K3.
Only the second open-source model in history to rank Top 5 in the AI Arena Leaderboard.
Beats Fable comfortably — and look at the massive gap it has ahead of Opus 4.8 — next version 3.9 will very likely beat Opus 5 too:
Open-source has really started to dominate the AI ecosystem. The tide is turning rapidly. Long gone are the days of seeing them as weaker variants of the real deal.
We've seen the shocking results from Kimi K3, DeepSeek v4 -- and now this.
Don't be surprised if a few months from now all the top 5 models on any major leaderboard are all open-source.
Qwen 3.8 Max scored 86.6 on Terminal-Bench 2.1, outperforming Claude Fable 5 (84.6) in one of the industry's toughest evaluations of real-world, agentic terminal coding workflows
And open-source is also becoming contagious -- before now the Max model in the Qwen series was never open-source. It was always the lesser models that had their weights made publicly available.
But things have changed completely now -- now with Qwen 3.8 even the top-tier Max model is open-source.
AI is getting a lot cheaper and more accessible -- both for coding and building AI-powered apps.
10+ days of non-stop coding...
We've gotten wild reports of the model being able to sustain 10+ days of autonomous project development -- managing hundreds of Git commits while continuously improving software.
Qwen 3.8 Max has been heavily optimized for agentic software development -- enabling it to tackle all your most complex engineering tasks from start to finish.
On PaperBench, which measures an AI's ability to reproduce and understand machine learning research, Qwen 3.8 Max posted an impressive 93.0, comfortably ahead of Claude Opus 4.8 (88.8)
It's been designed from the ground up for:
Continuous scientific research
Multi-hour data science tasks
Long-horizon planning
Extended agent workflows
Qwen 3.8 Max achieved 92.6 on GPQA Diamond—a benchmark of graduate-level scientific reasoning—surpassing Claude Opus 4.8 (91.0) while remaining competitive with the strongest frontier models
Perhaps most importantly, Qwen 3.8 Max maintains context remarkably well across lengthy sessions -- reducing the looping and repetitive behavior that often appears during large refactoring projects.
Scary for the competition
Major improvements across software engineering and agent benchmarks:
The docs platform Anthropic, HubSpot, and Coinbase trust
Your documentation is a product decision. The companies building the most-used developer tools in the world chose Mintlify because great docs reduce support load, accelerate adoption, and convert evaluators into customers. Backed by a16z and Salesforce Ventures, Mintlify powers 20,000+ companies with AI-native documentation that keeps pace with your product — without pulling engineers off the roadmap to maintain it.
One Account. Every Market. No Closing Bell.
Markets don't wait for Monday. News breaks Saturday morning, and most traders can only watch.
Not on Liquid. Trade equities, commodities, forex, crypto, and prediction markets from one account, 24/7.
Log in with Google, deposit in minutes, trade anywhere. While everyone else refreshes headlines, you're already positioned.



