
OpenAI is reportedly building Astra, a model family designed to work on problems for hours or days
Quick Answer
OpenAI is developing a new model family called 'Astra' aimed at enhancing long-duration task performance, potentially alongside existing models like Sol and Terra.
Quick Take
CEO Sam Altman showcased Astra's capabilities to coordinate multiple agents for complex problem-solving, with a long-term vision of autonomous AI research by 2028.
Key Points
- Astra aims to outperform existing models in long-running tasks like advanced math.
- The model is currently in testing and may be the first under new US AI regulations.
- OpenAI plans to showcase its current models solving ten unsolved math problems soon.
- Astra could significantly enhance human researchers' productivity by 2028.
- The development requires substantial computational resources, raising funding questions.
DeepSignal Analysis
What happened
OpenAI is developing a new model family called Astra, which aims to enhance performance on long-duration tasks. CEO Sam Altman demonstrated Astra's capabilities to coordinate multiple agents for complex problem-solving. The models are currently in testing and will be the first to undergo a new U.S. regulatory framework.
Key evidence
- Astra is designed to be more capable at long-running tasks than previous models from OpenAI, according to CEO Sam Altman.
- The models are expected to be the first tested under a new U.S. regulatory framework that requires AI models to be submitted to the federal government before public release.
- OpenAI's long-term goal is to develop a fully autonomous AI researcher by March 2028, which would rely on long-running AI processes.
Why it matters
The development of Astra reflects OpenAI's ambition to create AI systems that can handle complex, long-term tasks, which current models struggle with. This could lead to significant advancements in AI capabilities, particularly in research and problem-solving. However, the success of Astra will depend on overcoming existing limitations in multi-agent coordination and error correction.
📖 Reader Mode
~3 min readOpenAI is working on a new model family tentatively called "Astra" that's meant to be far more capable at long-running tasks than anything the company has shipped so far.
CEO Sam Altman demoed Astra to politicians and regulators in Washington, D.C., this week. OpenAI stressed the system's ability to coordinate multiple agents over extended periods to tackle especially hard problems. The company pointed to complex projects and advanced math as potential use cases. The Information reported the details, citing three people familiar with the plans.
According to the report, Astra would form a new model class alongside OpenAI's existing Sol, Terra, and Luna families. Whether it ships as GPT-6 or as a variant within the GPT-5 line, something like GPT 5.7, hasn't been decided yet. There's no release date either.
OpenAI also plans to publish a report soon showing how the company used its most advanced AI to solve ten previously unsolved math problems. The goal is to show what its current models can already do.
Astra would be the first model tested under a new US regulatory framework
The models are already in testing, according to The Information. They're expected to be the first to go through the Trump administration's planned new AI framework, which would require AI models to be submitted to the federal government before public release. The administration aims to finalize the framework by the end of this week.
One key question is whether the models can avoid compounding errors during long-running workflows and correct themselves when a process drifts off course as the context keeps growing. That remains a major weakness in today's agentic systems. Multi-agent setups like Astra can also perform worse on tightly linked tasks such as planning because coordination overhead and compounding errors can wipe out any gains.
The long-term goal is autonomous AI research
The Astra rumors line up with earlier statements from the company. Chief Scientist Jakub Pachocki said on OpenAI's official podcast last summer that the company wants to build AI systems that can work on a problem for hours or days. Current systems are often limited to short tasks, but OpenAI wants models that can plan, reason, and experiment over longer time horizons. Late last year, the company even raised the question of how to think about systems that could solve tasks a human would need centuries to complete.
By March 2028, OpenAI wants to have a fully autonomous AI researcher that can run research projects on its own, and that system would also depend on long-running AI processes. As early as this September, the company plans to have an AI system with research-intern-level skills that would significantly speed up human scientists. Astra could end up being that system.
Pachocki also said these systems will need far more compute. OpenAI's long-term infrastructure plans reflect that ambition. Whether the startup's revenue grows fast enough to fund that massive buildout remains an open question.
— Originally published at the-decoder.com
Want this in your inbox every morning?
Daily brief at your local 8am — bilingual EN/中文, free.
More from The Decoder
See more →
An AI model programmed nonstop for 19 days on a single MirrorCode task that cost $2,600 to run
Epoch AI's MirrorCode benchmark reveals Claude Opus 4.7 as the leader with a 56% solve rate, reconstructing a 16,000-line toolkit in 14 hours. Despite this, all models tested struggle with the most complex tasks, highlighting limitations in current AI capabilities. The single task consumed $2,600 over 19 days, raising questions about cost-effectiveness in AI development.

