To the line
Capabilities on the main line · VERIFIABLE BEHAVIOUR

Reliable autonomy

In one viewBefore trusting AI with an operating theatre, a power plant or a company's books, you must be able to prove it will not fail silently. No such method exists today.

Autonomy is limited by guarantees, not intelligence. Without formal behavioural verification, AI will not enter critical loops.

StatusUNSOLVED
TypeCapabilities on the main line
Marker?
Events in dossier5
Development chronology

In progress

препятствие

Silent failures

About this eventThe core problem of long tasks: failing without signalling failure — confidently returning a wrong result.

An agent feeds the output of one step into the next, so small mistakes compound along a long trajectory. Reliability therefore needs more than a stronger model: explicit state checks, action limits, audit logs and a safe way to stop the process.

Source: METR · длинные задачи
2026

Dangerous capability evals

About this eventBio and cyber risk testing became a standard part of frontier releases rather than a goodwill gesture.

Frontier labs now publish frameworks that connect measured dangerous capability levels to required safeguards. This is not yet a safety certificate: the evaluations are still evolving, and results depend on test scenarios and the quality of external review.

Source 1: OpenAI · Preparedness FrameworkSource 2: Anthropic · Responsible Scaling Policy
сейчас

Mechanistic interpretability

About this eventThe attempt to see inside the model rather than judge it by its output. So far it works at toy scale.

Planned

в планах

Audit as a requirement

About this eventRegulation turns verifiability from a virtue into a shipping requirement.

Distant horizons

впереди

A behavioural certificate

About this eventA formal guarantee that a system stays inside drawn boundaries — the equivalent of aircraft certification.

Sources and research

Primary material behind this dossier: papers, lab publications and official reports.

Capabilities on the main line