German Language LLM Index
a PeerBench project

Individual experiment result

SB10K (DE)

GPT-5.5 OpenAI

Matthews correlation coefficient reasoning on / locked provider-internal
62.5%
Matthews correlation coefficient

Run details

Date
2026-06-13
Test cases
1,024
Median latency
1.36s
Total cost
$4.12
Avg. prompt tokens
756.9
Avg. answer tokens
5.9
Avg. reasoning tokens
2.2
Quantization
provider-internal

About this benchmark

Native German Twitter sentiment (positive/neutral/negative), generate_until variant; primary metric MCC.

3-class sentiment native · Native German Source ↗

Provenance

Canonical model
openai/gpt-5.5
Run ID
2026-06-13T12-15-51__openai__codex__gpt-5.5-low__sb10k_de