Discovery Loop launched in August 2026 as a Delaware public benefit corporation with only its four founders and a generic engineering role posted 44 sources. Crucially, the resolution criteria require a specific, named framework with operational commitments and thresholds—analogous to Anthropic's Responsible Scaling Policy or Google DeepMind's Frontier Safety Framework 2 sources. Launch-day artifacts, which include mission statements about being a 'positive force for humanity' and standard PBC language 44 sources, explicitly do not qualify. With zero safety or policy headcount currently, building a document of this rigor from scratch will take substantial time.
The primary drivers pushing estimates into 2029 are Discovery Loop's deployment posture and the historical reference class. Because the company plans to act as its own first customer, it lacks the external model-deployment triggers that typically force the publication of safety frameworks. Furthermore, imminent regulations like California's SB 53 or New York's RAISE Act target large frontier developers with revenue exceeding $500 million wiley.law—thresholds a pre-revenue startup will not hit for years. The reference class for peer startups is sobering: Anthropic took over 2.5 years to publish its RSP, xAI took well over a year for a draft, and peers like Safe Superintelligence (SSI) have published nothing two years post-launch ssi.inc. Even Thinking Machines took roughly 18 months to publish a 'high-level framework' that explicitly deferred concrete stop conditions and access criteria, which would likely fail this question's strict bar 55 sources.
Conversely, the timeline is accelerated by the intense 2026 focus on exactly what Discovery Loop aims to build: automated AI research and recursive self-improvement 2 sources. The July 2026 'Pacing the Frontier' letter, signed by over 1,100 researchers, explicitly demands pacing mechanisms for recursive self-improvement 44 sources. Combined with scrutiny from FLI’s indices futureoflife.org and Discovery Loop's joint-research partnership with Google—whose Frontier Safety Framework already outlines Tracked Capability Levels for ML R&D deepmind.google—there will be substantial reputational pressure and ready-made templates for the founders to eventually formalize their safety commitments.
The median estimate of March 2029 (roughly 31 months post-launch) balances the time required to build an internal platform, hire safety staff, and encounter concrete evaluation needs against the heightened external pressure for RSI governance. The wide intervals, particularly the long right tail (p75 in January 2031, p90 in September 2034), reflect a roughly 25-30% probability that a qualifying public document is never published. This 'never/very late' mass accounts for scenarios where Discovery Loop remains an internal research shop, publishes only vague non-qualifying blog posts, pivots, fails to scale, or is simply reabsorbed into Google before ever needing a public operational safety framework.
Assessed against expectations for early lean headcount scaling, the earliest bound shifted slightly later to reflect the low likelihood of drafting a complex policy framework during an initial bootstrapping phase.