Granite 3.3 8B (Non-reasoning)

Granite 3.3 8B (Non-reasoning)

IBM · Released Apr 16, 2025
Intelligence #574 / 699
22.1 our score
Speed #320 / 320
14.3 tok/s
Input Price #154 / 714
$0.030 per 1M tokens
Output Price #226 / 714
$0.250 per 1M tokens
Context
Not reported

Analysis Summary

Granite 3.3 8B Non-reasoning is a compact IBM model with a very low measured intelligence result and no additional benchmark evidence in the supplied record. Its pricing is clear at $0.03 per million input tokens and $0.25 per million output tokens, giving it a low-cost position despite limited demonstrated capability.

That profile may fit basic classification, routing, extraction, and repetitive transformations at scale. It is not appropriate as a primary model for nuanced client content, complex reasoning, coding, or autonomous agents. Use it only for narrowly defined tasks with representative quality testing and automatic or human review.

Assessed August 9, 2026

Editorial notes

Granite 3.3 8B is a low-cost IBM model with limited measured intelligence and no supporting workflow benchmarks, making it suitable for simple automation rather than demanding content or coding work.

Rankings consider pricing, capabilities, benchmarks, and real-world applicability and are refreshed as new models launch. Feedback?

DFO Verdict

Granite 3.3 8B is a low-cost IBM model with limited measured intelligence and no supporting workflow benchmarks, making it suitable for simple automation rather than demanding content or coding work.

#574 of 699 overall

How Granite 3.3 8B (Non-reasoning) compares

Granite 3.3 8B (Non-reasoning) ranks #414 of 427 AI models we track for overall intelligence. At $0.03 per million input tokens it is cheaper than 78% of comparable models.

Position in the field
Intelligence: smarter than 18% of models #574
Speed: faster than 0% of models #320
Price: cheaper than 78% of models #154
worst in fieldmedianbest in field
Price vs frontier peers · $ per 1M tokens
Granite 3.3 8B (Non-reasoning) $0.03 in $0.25 out
Claude Opus 5 $5.00 in $25.00 out
SpaceXAI: Grok 4.6 $2.00 in $6.00 out
Qwen: Qwen3.8 Max $2.00 in $6.00 out

Dark bar = input · light bar = output, scaled to the priciest peer.

Context window vs peers · tokens
Granite 3.3 8B (Non-reasoning) 0

1M tokens ≈ 8 full-length novels or ~2,500 pages of business documents in a single request.

Intelligence0.2Technical0Value7.3Content2.1
Performance profile

Strongest on value. The pulled-in technical corner is the trade-off, and if the shape matters more than the price, this is your model.

Compare shapes side-by-side →

Pricing

Token Type Cost per 1M tokens Cost per 1K tokens
Input $0.03 $0.000030
Output $0.25 $0.000250

What would Granite 3.3 8B (Non-reasoning) cost your business?

Pick the job that looks most like yours, then fine-tune with the sliders. Estimates update live.

A website chatbot handling around 100 customer conversations a day, a few short messages each.

3,000
One request is one message, email, draft or automation call.
1,200 tokens

$0/mo Granite 3.3 8B (Non-reasoning)

Full calculator with 714 models → Price Calculator

DFO AI AUTOMATION

These numbers get smaller with the right architecture.

We route routine calls to cheap models and save Granite 3.3 8B (Non-reasoning) for the hard ones. Most clients cut their estimate by 60-80%.

Talk to our team

Frequently asked questions about Granite 3.3 8B (Non-reasoning)

How much does Granite 3.3 8B (Non-reasoning) cost?

Granite 3.3 8B (Non-reasoning) costs $0.03 per million input tokens and $0.25 per million output tokens.

Who created Granite 3.3 8B (Non-reasoning)?

Granite 3.3 8B (Non-reasoning) is developed by IBM and was released on April 16, 2025.

© 2026 Design for Online Ltd. Registered in England and Wales No. 10328553. VAT Registered. Design for Online® and Forerunner® are registered trademarks.