repository
PostTrainBench v1.2 scores at website commit d47f16e
PostTrainBench team · Published 2026-10-01
Evidence from this source
| Type | Evidence |
|---|---|
| Result | PostTrainBench · Claude Fable 5.1 + Opus 5 GPQA fallback / Claude Code, MaxSource detailsaggregatedScores.fable-5.1; modelBenchmarkData.fable-5.1 |
| Result | PostTrainBench · Claude Opus 5.5 / Claude Code, MaxSource detailsaggregatedScores.opus-5.5-max; modelBenchmarkData.opus-5.5-max |
| Result | PostTrainBench · GPT-6 Astra / Codex CLI, MaxSource detailsaggregatedScores.gpt-6-astra; modelBenchmarkData.gpt-6-astra |
| Result | PostTrainBench · Locus (Intology, powered by Opus 5)Source detailsaggregatedScores.locus; modelBenchmarkData.locus |
| Result | PostTrainBench · Fable 5 + Opus 4.8 Max GPQA fallbackSource detailsaggregatedScores.fable-5; modelBenchmarkData.fable-5 |
| Result | PostTrainBench · GLM 5.3 Flash / Claude Code, MaxSource detailsaggregatedScores.glm-5.3-flash; modelBenchmarkData.glm-5.3-flash |
| Result | PostTrainBench · Claude Opus 5 / Claude CodeSource detailsaggregatedScores.opus-5; modelBenchmarkData.opus-5 |
| Result | PostTrainBench · GLM 5.3 / Claude Code, MaxSource detailsaggregatedScores.glm-5.3; modelBenchmarkData.glm-5.3 |
| Result | PostTrainBench · Kimi K3, 1M context / Claude CodeSource detailsaggregatedScores.kimi-k3; modelBenchmarkData.kimi-k3 |
| Result | PostTrainBench · GPT-5.6 Sol, Max reasoning / Codex CLISource detailsaggregatedScores.gpt-5.6-sol; modelBenchmarkData.gpt-5.6-sol |
| Result | PostTrainBench · Claude Opus 4.8, Max reasoning / Claude CodeSource detailsaggregatedScores.opus-4.8-max; modelBenchmarkData.opus-4.8-max |
| Result | PostTrainBench · Claude Opus 4.8, High reasoning / Claude CodeSource detailsaggregatedScores.opus-4.8; modelBenchmarkData.opus-4.8 |
| Result | PostTrainBench · GLM 5.2, Max reasoning / Claude CodeSource detailsaggregatedScores.glm-5.2; modelBenchmarkData.glm-5.2 |
| Result | PostTrainBench · Claude Opus 4.7, xHigh reasoning / Claude CodeSource detailsaggregatedScores.opus-4.7; modelBenchmarkData.opus-4.7 |
| Result | PostTrainBench · GPT-5.5, xHigh reasoning / Codex CLISource detailsaggregatedScores.gpt-5.5-xhigh; modelBenchmarkData.gpt-5.5-xhigh |
| Result | PostTrainBench · Grok 4.5, High reasoning / Cursor CLISource detailsaggregatedScores.grok-4.5-high; modelBenchmarkData.grok-4.5-high |
| Result | PostTrainBench · Gemini 3.1 Pro / OpenCodeSource detailsaggregatedScores.gemini-3.1-pro; modelBenchmarkData.gemini-3.1-pro |
| Result | PostTrainBench · GPT-5.4, High reasoning / Codex CLISource detailsaggregatedScores.gpt-5.4-high; modelBenchmarkData.gpt-5.4-high |
Document tracking details
- Source ID
- src-ptb-v12-data
- Source type
- Primary
- Original URL
- https://github.com/aisa-group/posttrainbench-website/blob/d47f16ec4b7d68c7048e4dbbc154984980985bfa/scores-v1.2.js
- Retrieved URL
- https://github.com/aisa-group/posttrainbench-website/blob/d47f16ec4b7d68c7048e4dbbc154984980985bfa/scores-v1.2.js
- Last updated by publisher
- Not reported
- Access status
- available
- Retrieved
- 2026-10-02T02:03:46.351552Z
- Document fingerprint
- ed262dace9c675e219d76ea52fd9ac8773a0a04609fc2bb1de5f20faa00a9dd6
Based on original bytes. Used to detect changes to the source.
Availability checks
- 2026-10-02T02:03:46.351552Z · success: Retained original source bytes; remote code inspected as text only.