Agent objective
- Interact with the supplied stateful environment.
- Produce verifier-checkable actions or artifacts.
- Maximise scalar reward under the package contract.
A five-environment mathematics suite built around long-horizon inference under scientific and measurement uncertainty. Each domain has its own mathematical model and action grammar, requiring agents to maintain compact beliefs, choose observation-dependent experiments, manage nonlinear costs and irreversible actions, predict sealed outcomes, and know when to identify, return a candidate set, or abstain.
A five-environment mathematics suite built around long-horizon inference under scientific and measurement uncertainty. Each domain has its own mathematical model and action grammar, requiring agents to maintain compact beliefs, choose observation-dependent experiments, manage nonlinear costs and irreversible actions, predict sealed outcomes, and know when to identify, return a candidate set, or abstain.
Enough detail to understand the intellectual terrain; generated instances, hidden mechanisms, and solution paths remain inside the private package.
| Environment | Mathematical or technical frontier | Adaptive research problem |
|---|---|---|
| SpectralSequenceForge | Homological algebra and spectral sequences | Route page-level evidence into compatible differential, extension, and secondary-operation tests under hidden calibration. |
| TropicalResultantMaze | Tropical geometry and Puiseux lifting | Navigate subdivisions and continuation branches, deciding when calibration or a chart change is worth its cost. |
| FreeMomentCipher | Free probability and operator algebras | Route word families, linearizations, and resolvent probes while managing realization reuse and finite-size contamination. |
| ConformalCrossingLab | Conformal bootstrap and inverse field theory | Choose an irreversible continuation branch from noisy sector evidence, then resolve mixed-correlator structure. |
| EndoscopicPacketLab | Automorphic forms and representation theory | Coordinate ramified local evidence, global twists, transfer checks, normalization ambiguity, and missing observations. |
A source-informed frozen controller scored 0.8445 across 50 preregistered episodes and passed 40/50, with FreeMomentCipher remaining the clearest frontier at 6/10. More importantly, plausible first observations changed the preferred follow-up in 14/15 cases; static and one-step controls largely failed. The package therefore measures genuinely adaptive investigation, not a fixed query script. Its release reports 95 tests and ten fixture replays.
We publish aggregate behavior and task structure, while withholding generated instances, hidden labels, exact successful probes, private checks, and solution trajectories.
Shown with its provenance and limitations; it is not a performance guarantee.
Reported result from the evaluation artifact supplied with this package.
As identified by the supplied artifact.
50 reported runs.
hyperomega_v2_final_score_summary.json
Machine-readable provenance and the exact displayed metric are available in results.json.
The paid ZIP will live in a private R2 bucket. Vercel authorizes the buyer and issues a 2–5 minute object URL; R2 serves the bytes directly.
Authenticated buyer + entitlement check
+ private R2 object + 2–5 minute signed URL
= direct, auditable download
Package SHA-256
3e81c53ab32ba7157ab6f17c7719c9682dadf281967f2852314a53d54385143fOne purchase licenses this identified item to one legal organisation for worldwide, perpetual commercial model training, evaluation, research and development. Redistribution and resale of the package are not permitted.