Kaelum Bench v0.2 · Public Calibration
glm-5.2
IFRS Reporting · Run evidence
Provisional calibration result — not a verified v1 result.
Configuration-specific evidence for a single published run. It is not a statistical ranking and is not an overall score.
Run identity
Model and configuration
- Model display name
- glm-5.2
- Requested model ID
- z-ai/glm-5.2
- Provider
- Wafer
- Reasoning configuration
- xhigh
- Suite
- IFRS Reporting
- Validity
- Provisional
Published score
Scoring evidence
- Score
- 546.0 / 600
- Percentage
- 91.0%
- Scoring method
- Deterministic
- Comparison protocol
- Configuration-specific
- Tool access
- None
- Published total cost
- $0.255
Usage
Tokens and execution
- Input tokens (tested model)
- 56,213
- Output tokens (tested model)
- 50,046
- Total tokens (tested model)
- 106,259
- Attempted tasks
- 60 / 60
- Valid responses
- 60 / 60
- Invalid JSON
- 0
- Schema failures
- 0
- Provider failures
- 0
- Missing responses
- 0
Release provenance
Versioned run evidence
- Benchmark ID
- FMB-KNOW-ACC-IFRS-02
- Benchmark version
- 0.2.0-dev
- Scorer ID
- fmb-live-scorer-v1
- Started at
- Completed at
- Run ID
- d0aa0a2b-d885-5fe9-a22e-5b79409595bf
- Signed release ID
- 9dbf3d05-67e5-4bbd-aada-585c12163f1a