OY Labs Achieves A Leading ARC-AGI-3 Public Score In Its Mission To Build The World's Smartest AI
| System and configuration | Test set | Score | Run cost | |||
| OY Labs OY1-AGI | Public | 100.0% | $415.37 | |||
| GPT-6 Astra · High reasoning | ||||||
| OpenAI GPT-6 Astra | Semi-private | 99.95% | $18,817 | |||
| Provider Adapter · High reasoning | ||||||
| Retrodict | Public | 99.9% | $654 | |||
| NVIDIA NOOA | Public | 85.1% | $332 |
Selected results checked on 10 September 2026. Public and semi-private evaluations are not directly comparable; systems and cost accounting also differ. OY1's leaderboard application remains under review.
OY1-AGI also exceeds Retrodict's 99.9%, the next-highest community score below 100%, and NVIDIA NOOA's 85.1%. The NVIDIA comparison concerns different systems, rather than the same model and reasoning configuration.
Three publicly available OY1-AGI scorecards
All three scorecards are linked below. The 9 September result is the main run highlighted in this release, with a recorded API cost of $415.37.
| Published scorecard | Score | Games won | Levels solved | Actions | ||||||||||||||||||||||
9 September 2026 · Main result | 100.0% | | 25 / 25 | | 183 / 183 | | 6,732 | | 7 September 2026 | 100.0% | | 25 / 25 | | 183 / 183 | | 6,714 | 6 September 2026 | 100.0% | | 25 / 25 | | 183 / 183 | | 6,659 | OY Labs' initial priority is developing and training its own large language models for mainstream AI users and institutions, including models trained to meet institutions' specific requirements. OY Labs reports tremendous interest in its product as it pursues leading AI capability at the lowest possible cost per unit of intelligence. “Our goal is to make advanced intelligence affordable enough to use everywhere,” said Laurin Bylica, Cofounder and CEO of OY Labs. OY2-AGI built on ethical values from the outset A core differentiator for OY2-AGI will be its foundation in strong ethical values from the outset. OY Labs plans to embed a hard-coded constitution into the model's design, with safeguards intended to protect user privacy and prevent harmful use. “We are building powerful AI models to advance human welfare. That purpose should shape both the capabilities we develop and the boundaries we build into them,” said Loïc Bellez, Cofounder and CTO of OY Labs. The current benchmark system The benchmark result was produced by GPT-6 Astra at High reasoning with OY Labs' OY1 harness, which guides the model's problem-solving process. OY Labs has open-sourced the benchmark implementation under the Apache 2.0 license, including the evaluated source, reproduction instructions and results documentation. Explore the OY1 benchmark repository. About OY Labs OY Labs is an AI research company founded by Laurin Bylica and Loïc Bellez and operated by Orca Labs sp. z o.o., registered in Poland. Sources and notes to editors The scorecards substantiate scores, levels, environments and action counts. OY Labs reports cost, runtime, normal completion and 37,704 evidence consistency checks; these software checks are not ARC Prize certification. The repository states that public games were used during development. These results do not establish a held-out ranking or general superiority across AI tasks. Product interest and application-review status are company statements. 9 September · Main: 7 September: 6 September: Benchmark comparisons: ARC Prize's OpenAI results · Community Leaderboard Media contact: ... · oylabs.ai OY Labs Contact Details: ![]() |
Legal Disclaimer:
MENAFN provides the
information “as is” without warranty of any kind. We do not accept any
responsibility or liability for the accuracy, content, images, videos,
licenses, completeness, legality, or reliability of the information
contained in this article. If you have any complaints or copyright issues
related to this article, kindly contact the provider above.


Comments
No comment