German Language LLM Index
a PeerBench project

Individual experiment result

SB10K (DE)

Gemma 4 26B A4B Google

Matthews correlation coefficient reasoning off provider-internal
15.2%
Matthews correlation coefficient

Run details

Date
2026-06-08
Test cases
1,024
Median latency
0.70s
Total cost
$0.12
Avg. prompt tokens
758.8
Avg. answer tokens
2.8
Avg. reasoning tokens
Quantization
provider-internal

About this benchmark

Native German Twitter sentiment (positive/neutral/negative), generate_until variant; primary metric MCC.

3-class sentiment native · Native German Source ↗

Provenance

Run ID
2026-06-08T16-39-47__vertex_ai__openai__google__gemma-4-26b-a4b-it-maas__sb10k_de