- Organizations
- IBM
- Granite 3.3 8B Instruct
Granite 3.3 8B Instruct: API Pricing, Context Window & Benchmarks
Granite 3.3 8B Instruct is a language model from IBM, released in April 2025, with multimodal input.
Granite 3.3 models feature enhanced reasoning capabilities and support for Fill-in-the-Middle (FIM) code completion. They are built on a foundation of open-source instruction datasets with permissive licenses, alongside internally curated
Granite 3.3 8B Instruct benchmarks
Rankings
Quality Tracker
Granite 3.3 8B Instruct Performance Across Datasets
Scores sourced from the model's scorecard, paper, or official blog posts
Granite 3.3 8B Instruct pricing
Providers
Granite 3.3 8B Instruct starts at $0.500 per million input tokens and $0.500 per million output tokens via Replicate.
| Provider | Input $/M | Output $/M | Context in / out | TTFT p50 / p95 s | Output avg / p5 c/s | Success 7d | Modalities in / out |
|---|---|---|---|---|---|---|---|
| $0.500 | $0.500 | 128.0K/8.2K | —/0.30 | 50/— | — | / |
Cached input is the discounted price for prompt tokens served from a provider cache. TTFT is time to first token. Output is characters per second; p5 is the sustained floor exceeded by 95% of observed requests. Success is calculated from completed versus failed requests over the trailing seven days.
Granite 3.3 8B Instruct model size
Granite 3.3 8B Instruct has 8 billion parameters. See how it compares to other models in the same parameter range.
Granite 3.3 8B Instruct context window
Input and output token limits for Granite 3.3 8B Instruct, plus how it ranks on long-context understanding.
Granite 3.3 8B Instruct API
Available from the model provider
Granite 3.3 8B Instruct has an official provider API. It is not currently routed through the LLM Stats gateway.
Read the official API documentationGranite 3.3 8B Instruct latency
Granite 3.3 8B Instruct time to first token, sustained output throughput, and failed-request rate from live API traffic over the trailing 7 days.
Granite 3.3 8B Instruct examples
Recent arena outputs from Granite 3.3 8B Instruct, picked from the highest-ranked matchups.
Granite 3.3 8B Instruct license
Granite 3.3 8B Instruct is released under the Apache 2.0 license, which permits commercial use, has 8.0B parameters, has a knowledge cutoff of April 2024.
- License
- Apache 2.0
- Commercial use allowed
- Parameters
- 8.0B
- Knowledge cutoff
- April 2024
Apache License 2.0 - allows commercial use
Granite 3.3 8B Instruct resources
Official sources for Granite 3.3 8B Instruct: api documentation, official playground, official launch post, source repository.
Granite 3.3 8B Instruct vs other models
The most-compared alternatives to Granite 3.3 8B Instruct are Phi 4 Reasoning Plus, GPT-4o, Qwen3 30B A3B. Open any pair side-by-side for benchmarks, pricing, context, and latency.
Models like Granite 3.3 8B Instruct
Models ranked just above and below Granite 3.3 8B Instruct by LLM Stats score.
FAQ
Common questions about Granite 3.3 8B Instruct.