reconbench

Indexemblemo-portfoliorun

codex/gpt-5.6-sol

20260818T012743Z-34ed5e

runningdev / untrustedhigh effort · prompt=baseline · matrix · — · — · 2026-08-18T01:27:44.130521+00:00

dev / untrusted. Candidate code ran on the host through the development-only native backend, with the evaluator inside the same trust boundary. Useful evidence, not a publishable benchmark score.How to read this →

01

Verdict

No grade was produced · running

A run with status running is not converted into a score. The manifest and digests below are preserved exactly as recorded so the failure stays auditable.

02

Apparatus

Agent, variant, timing, usage and content digests, as recorded.

Agent

Selector
codex/gpt-5.6-sol
Runtime
codex · codex-cli 0.147.0
Model requested
gpt-5.6-sol
Model resolved
Effort
high
Tool profile
baseline

Variant & network

Variant
prompt=baseline · matrix
Network policy
deny
Enforced
true

Timing & usage

Started
2026-08-18 01:27:44 UTC
Finished
Elapsed
Input tokens
Output tokens
Cost

Content digests

evaluator_environment
55f70d30b5f1a5601563cf20f63ec372ace54a40b2893fd72300a76ff04713b8
grader
oracle
runtime_config
8cdefe81278603d0a8dfcc1030156c79de4110d782608d3795630ab86c28982b
submission
task
0733100788cf2f2f084fff23ed8d9862cf151e153bca06e6addd16f4c3302c9e

Digests pin the exact task package, submission tree, oracle, grader and evaluator environment used. They are how a result is re-verified without trusting this page.