Google: Gemma 4 31B (free)

google/gemma-4-31b-it:free

Free token ratesTool callingVision inputReasoning
View on OpenRouter ↗

Price Published

Input
$0
per million tokens
Output
$0
per million tokens

A $0 token rate isn't unlimited. OpenRouter applies free-tier rate and daily limits, and needs an account and API key.

Other listings of this model

Specifications Published

Context window
262,144 tokens
Max output
32,768 tokens
Tool calling
Supported
Vision input
Supported
Reasoning
Supported
Knowledge cutoff
Not reported
Listed since
Apr 2, 2026

Measured by the Club Signed cheapoS reports

Request success
5%
Likely 0.9–23.6% · 1 of 20 known outcomes
Avg request time
6,763 ms
20 timed reports · includes generation
Rate-limit / quota failures
19
of 19 reported failures

20 request reports from 1 member · OpenRouter via OmniRoute · Sep 20, 2026 to Sep 29, 2026

What these numbers mean

0 cancelled and 0 unknown outcomes are excluded from the success percentage. Request success means the provider answered; it doesn't mean the code or task succeeded. The likely range is a 95% Wilson interval: with few outcomes it stays wide, and one bad day can move the rate a lot. Rates need at least 20 known outcomes and averages at least 5 timed reports.

  • rate limit quota: 19

How we measure · Reports come from cheapoS members who share model details.

Benchmarks Artificial Analysis via OpenRouter

Intelligence index
Not reported
Coding index
43.4
Agentic index
4.2

Scores from Artificial Analysis, as distributed in the OpenRouter catalog. They measure test suites, not request reliability on real work.

Try it

  1. Get an OpenRouter API key. Create one here (free models need an account, not credits). Other free tiers are compared on Providers.
  2. Add it to cheapoS. Open Models → Open OmniRoute, add OpenRouter with your key, then refresh models. Get cheapoS
  3. Pick this model. Use the ID google/gemma-4-31b-it:free as a worker or reviewer, and start with a small task you can check.

About this model OpenRouter description

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...

Other free models with tool calling

Catalog data from the OpenRouter model catalog, retrieved Sep 29, 2026, 12:37 PM UTC. A listing isn't a live availability check.