API Provider1 active models2 organizations

Sambanova: API pricing, speed & models

Sambanova hosts 1 active AI models, with input pricing from $0.40 per 1M tokens. Compare Sambanova's API speed, pricing, and reliability against other inference providers.

Catalog throughput328 c/s
Catalog latency1.08s
From$0.40 /M tok

Catalog

Type
Price
1 model
At a glance

Sambanovapricing, performance & catalog

The citable facts about Sambanova's 1 model — sourced from provider APIs and refreshed continuously.

Lowest price
Qwen3 32B at $0.400 per 1M input tokens
Largest context
Qwen3 32B at 128K tokens
Catalog
1 active models from 2 organizations

Most affordable

  1. 1Qwen3 32B$0.400/M

Fastest median throughput

No throughput data yet.

Largest context

  1. 1Qwen3 32B128K

FAQ

Common questions about Sambanova.

What is Sambanova?

Sambanova is an API provider that hosts large language models. Active models: 1; From (input): $0.40 / 1M tok; Catalog throughput: 328 c/s.

How many models does Sambanova offer?

Sambanova currently serves 1 active models out of 6 historical offerings on LLM Stats.

What is Sambanova's API pricing?

Sambanova input pricing starts from $0.40 per 1M tokens, with the most expensive offering at $0.4 per 1M tokens. See the Pricing tab above for the full per-model breakdown.

Does Sambanova support function calling?

Yes. 1 of 1 models on Sambanova support function calling (tool use). The Capabilities tab lists which specific models accept tool definitions.

Does Sambanova support JSON mode and structured output?

Yes. 1 of 1 models on Sambanova support structured output (JSON mode / schema-constrained generation). The Capabilities tab shows which specific models accept response_format or json_schema parameters.

Does Sambanova offer batch inference?

Yes. 1 of 1 models on Sambanova support batch inference for cheaper, asynchronous workloads.

Whose models does Sambanova host?

Sambanova hosts models from Alibaba Cloud / Qwen Team and Meta. See the Models tab for the full catalog grouped by creator.

How do I start using Sambanova?

Sign up at https://sambanova.ai/ to get an API key, then call Sambanova's API directly from your application. Most clients work out of the box by pointing the OpenAI SDK at Sambanova's base URL with your key. Use the Models and Pricing tabs above to pick the right model for your latency, cost, and context-window requirements.