Google (native)reference

generateContent — Google's own protocol · serves gemma-4-31b-it as gemma-4-31b · probed live 2026-07-28 · back to the matrix

This is Google's native API. Google's OpenAI-compat shim is graded separately as its own provider: google-openai.

Harness
B−
A − 1 red − 1 yellow, per the ladder; flags below.
Billing
Reported usage consistent with bytes on the wire.
Caching
Fail
A repeat 20K-token submission waits out the free tier's per-minute token quota (~46s of 429 retries) before it is served — and gemma-4 has no paid tier, so there is no quota to buy. Whether context is reused is unobservable behind the wait. Speed numbers →
Faithfulness
reference
This provider is the reference implementation the battery compares against.

Flags

Every flag links to its finding — evidence, repro, and disclosure records live there; the deepest findings have full writeup pages.