Ibm logo

Ibm: Granite-4.0-H-Small

Open-weight hybrid model for enterprise chat, coding, retrieval-augmented generation, and tool-calling workloads

Source-linked granite Open weights Released 2025-10-02
API record Report
InputT
OutputT
Input price$0.064/M
Output price$0.265/M
Context131.072K
Max output131.072K
Providers1
Inference availability

Providers

Provider-specific identifiers, limits, and listed prices per million tokens. Every row links back to the provider's own documentation.

Report a price
Providers offering Granite-4.0-H-Small
ProviderProvider model IDContextMax outputInputOutputCache readCapabilitiesDocs
watsonx.ai ibm/granite-4-h-small 131.072K 131.072K $0.064 $0.265 ToolsJSON Docs ↗

Capability badges appear only where the provider catalog explicitly lists support. A blank cell means the source is silent, not that the feature is absent.

Listed rates

Price across providers

Input price per million tokens as published by each provider. Bars are drawn from listed rates only — no traffic weighting, since the catalog observes no requests.

Lowest input $0.064/M

Across 1 priced provider

Median input $0.064/M

Midpoint of listed rates

Highest input $0.064/M

Same as the lowest listed rate

Output range $0.265 – $0.265

Per million output tokens

watsonx.ai $0.064/MLowest
Cost calculator

Estimate a workload

$0.00
Excludes taxes, non-token charges, and tiered discounts.
Specification

Capabilities

Recorded from the source catalog and provider listings.

× Reasoning No
Tool calling Yes
Structured output Yes
× Attachments No
× Vision input No
Open weights Yes
Creator
Ibm
Model family
granite
Knowledge cutoff
Not documented
License
Not documented
Release date
2025-10-02
Model ID
ibm/granite-4-h-small

Weights: Hugging Face ↗

Published evaluations

Benchmarks

Every result stays attached to its source, version, metric and harness. Scores from different versions are never merged, and the profile below plots each benchmark against its own population rather than on a shared scale.

Benchmark registry
No published benchmark results

ModelBench does not infer quality from price, context size, or model name. When a source publishes a comparable result, it appears here with its version and link.

Read the methodology
Catalog activity

Change log

Field-level changes detected between successful source imports.

Full change log
No changes recorded

This record has not changed within the retained import history.

Public API

Use this record

Fetch the complete source-linked model record. No key, no account, no rate-limited tier.

API documentation
Endpoint
GET https://www.modelbench.lol/api/v1/models/ibm/granite-4-h-small
curl
curl "https://www.modelbench.lol/api/v1/models/ibm/granite-4-h-small"
Common questions

Frequently asked questions

Answered directly from the stored record — nothing here is generated beyond the catalog's own fields.

What is Granite-4.0-H-Small?

Open-weight hybrid model for enterprise chat, coding, retrieval-augmented generation, and tool-calling workloads. It is published by Ibm and catalogued here from Models.dev.

How much does Granite-4.0-H-Small cost?

Listed input pricing starts at $0.064 per million tokens from watsonx.ai.

What is the context length of Granite-4.0-H-Small?

Granite-4.0-H-Small accepts up to 131.072K tokens of context and returns up to 131.072K output tokens.

Does Granite-4.0-H-Small support tool calling and structured output?

Provider catalogs list support for tool calling, and structured output.

Which providers serve Granite-4.0-H-Small?

1 provider list this model: watsonx.ai.

Are the weights for Granite-4.0-H-Small open?

Yes. The weights are published and downloadable from Hugging Face.

When was Granite-4.0-H-Small released?

The catalog records a release date of 2025-10-02, last verified Aug 09, 2026.