Kaelum Bench v0.2 · Public Calibration

glm-5.2

IFRS Reporting · Run evidence

Provisional · Deterministic

Provisional calibration result — not a verified v1 result.

Configuration-specific evidence for a single published run. It is not a statistical ranking and is not an overall score.

Run identity

Model and configuration

Model display name
glm-5.2
Requested model ID
z-ai/glm-5.2
Provider
Wafer
Reasoning configuration
xhigh
Suite
IFRS Reporting
Validity
Provisional

Published score

Scoring evidence

Score
546.0 / 600
Percentage
91.0%
Scoring method
Deterministic
Comparison protocol
Configuration-specific
Tool access
None
Published total cost
$0.255

Usage

Tokens and execution

Input tokens (tested model)
56,213
Output tokens (tested model)
50,046
Total tokens (tested model)
106,259
Attempted tasks
60 / 60
Valid responses
60 / 60
Invalid JSON
0
Schema failures
0
Provider failures
0
Missing responses
0

Release provenance

Versioned run evidence

Benchmark ID
FMB-KNOW-ACC-IFRS-02
Benchmark version
0.2.0-dev
Scorer ID
fmb-live-scorer-v1
Started at
Completed at
Run ID
d0aa0a2b-d885-5fe9-a22e-5b79409595bf
Signed release ID
9dbf3d05-67e5-4bbd-aada-585c12163f1a