Analysis Summary
Qwen Plus 0728 thinking is a text model with a 1M token context window, tool use, and function calling. That context capacity is valuable for large document sets, extensive research material, and codebases where keeping source information available can simplify orchestration.
The thinking variant has no independent benchmark results in this record, so its reasoning and writing reliability cannot be confirmed despite the product positioning. Its higher price than the standard Qwen Plus version also makes it better suited to difficult, context-heavy tasks than routine traffic. Test it for long-document analysis and structured agents, with output checks before using it in client-facing work.
Assessed August 9, 2026
Editorial notes
Qwen Plus 0728 thinking provides a 1M token context window, tool use, and function calling for large-document workflows, but this variant lacks independent benchmark data and costs more than the standard version.
Rankings consider pricing, capabilities, benchmarks, and real-world applicability and are refreshed as new models launch. Feedback?
DFO Verdict
Qwen Plus 0728 thinking provides a 1M token context window, tool use, and function calling for large-document workflows, but this variant lacks independent benchmark data and costs more than the standard version.
How Qwen: Qwen Plus 0728 (thinking) compares
Its 1M-token context window is larger than 86% of the models we list. At $0.26 per million input tokens it is cheaper than 49% of comparable models.
Dark bar = input · light bar = output, scaled to the priciest peer.
1M tokens ≈ 8 full-length novels or ~2,500 pages of business documents in a single request.
Strongest on value. The pulled-in technical corner is the trade-off, and if the shape matters more than the price, this is your model.
Compare shapes side-by-side →Pricing
| Token Type | Cost per 1M tokens | Cost per 1K tokens |
|---|---|---|
| Input | $0.26 | $0.000260 |
| Output | $0.78 | $0.000780 |
What would Qwen: Qwen Plus 0728 (thinking) cost your business?
Pick the job that looks most like yours, then fine-tune with the sliders. Estimates update live.
A website chatbot handling around 100 customer conversations a day, a few short messages each.
Full calculator with 704 models → Price Calculator
These numbers get smaller with the right architecture.
We route routine calls to cheap models and save Qwen: Qwen Plus 0728 (thinking) for the hard ones. Most clients cut their estimate by 60-80%.
Talk to our teamAbout Qwen: Qwen Plus 0728 (thinking)
Qwen Plus 0728, based on the Qwen3 foundation model, is a 1 million context hybrid reasoning model with a balanced performance, speed, and cost combination.
Explore Related Models
Frequently asked questions about Qwen: Qwen Plus 0728 (thinking)
How much does Qwen: Qwen Plus 0728 (thinking) cost?
Qwen: Qwen Plus 0728 (thinking) costs $0.26 per million input tokens and $0.78 per million output tokens.
What is the context window of Qwen: Qwen Plus 0728 (thinking)?
Qwen: Qwen Plus 0728 (thinking) has a context window of 1,000,000 tokens (1M).
What can Qwen: Qwen Plus 0728 (thinking) do?
Qwen: Qwen Plus 0728 (thinking) supports tool use and function calling.
Who created Qwen: Qwen Plus 0728 (thinking)?
Qwen: Qwen Plus 0728 (thinking) is developed by Qwen and was released on September 8, 2025.
Data sourced from the OpenRouter API, Artificial Analysis, the Hugging Face Open LLM Leaderboard and our own internal testing. Scores are editorially curated by our team.
Last updated: August 18, 2026 8:38 pm