# Lightning - gemma-4-31b grade card

OpenAI chat/completions · serves `lightning-ai/gemma-4-31B-it` as gemma-4-31b · probed live 2026-08-25

HTML version: [https://inferencecanary.com/lightning/gemma-4-31b](https://inferencecanary.com/lightning/gemma-4-31b) · matrix: [https://inferencecanary.com/index.md](https://inferencecanary.com/index.md)

## Grades


| Axis | Grade | Basis |
|---|---|---|
| Inference | A− | A − 0 high − 1 medium severity, per the published ladder |
| Billing | OK | Reported usage consistent with bytes on the wire. |
| Caching | Yes (3/3 hits) | Reused already-seen context on 3/3 cold→warm trials. Speed numbers: [https://inferencecanary.com/speed.md](https://inferencecanary.com/speed.md) |


## Flags

- **[medium · inference] Lightning leaks template reasoning markers into content** - On an image-bearing tool turn, the channel parse intermittently fails in both directions at once: the reasoning field comes back empty and literal template marker text leads the answer field. Two of four identical runs leak, byte-identically, with the response nowhere near its token budget; the same provider's text-tool and user-image turns parse clean every time. (details: [/lightning-marker-leak.md](https://inferencecanary.com/lightning-marker-leak.md))

