AI news and explainers at 7 AM weekdays, plus a Sunday weekly at 8Get it in your inbox

METAL LAB

OpenAI's Altman Says Astra Will Reach AGI-Level by Year's End

In a TIME interview, Altman and his executives voiced confidence in Astra's AGI-level capabilities, though the definition of AGI itself remains contested

이미지: METAL LAB 생성

Summary

  • Sam Altman said OpenAI will have an internal system meeting his own definition of AGI-level capability by the end of 2026, and Chief Research Officer Mark Chen said the company is "80% of the way" there
  • Chief Scientist Jakub Pachocki said the upcoming model Astra can take a research paper and do a week's worth of a human researcher's work, turning experimental ideas into code and reporting results on its own
  • OpenAI is also tackling monetization on multiple fronts at once: "The Merge," a project combining Codex and ChatGPT, new desktop, pocket, and wearable devices, and an expansion of ads inside ChatGPT
발언 주체
샘 올트먼(오픈AI CEO), TIME 인터뷰
시점 자신
2026년 연말까지 내부 AGI급 시스템 확보, 아직 '완전 도달은 아님'
마크 첸(최고연구책임자)
AGI까지 '80%의 길을 왔다'고 언급
그렉 브록만(공동창업자)
이 시기를 훗날 AGI 출현 순간으로 회고할 것이라 발언
아스트라 능력
야콥 파초키(수석과학자): 연구 논문 한 편을 사람 연구자 약 1주일 분량으로 처리
기기 계획
탁상용·주머니용·착용형 3종, 첫 제품은 퍽 모양·음성 대화형, 내년 초 출시 예상
수익화
챗GPT 소비자 이용자 92%가 무료 이용(CFO 사라 프라이어), 광고·스폰서드 에이전트 확대 테스트 중
취재 방식
알렉스 히스 기자가 20명 넘는 임직원·투자자·경쟁사 관계자를 2주 넘게 인터뷰

Altman: "By Year's End, Astra Will Touch AGI"

OpenAI CEO Sam Altman says the company will have an internal system by the end of this year that meets his own definition of artificial general intelligence, or AGI — a hypothetical AI capable of doing nearly any intellectual work a human can do. He stopped short of saying OpenAI has "fully arrived" there, but he was unambiguous about the timeline. Chief Research Officer Mark Chen said the company is "80% of the way" to AGI, and co-founder Greg Brockman suggested this period will one day be remembered as the moment AGI actually arrived.

A seed-shaped icon labeled Astra stretches with a dotted arrow toward a dotted circle marked "AGI definition," annotated "80%." At the same time, a dotted arrow runs from Astra to a broken circle labeled "safety controls," tagged "risk signal" — illustrating that capability is growing while both the definition and the safety guardrails remain unresolved.

These remarks come from an in-depth TIME report by reporter Alex Heath, based on more than two weeks of interviews with over 20 OpenAI executives, employees, investors, and rivals.

Astra Can Do a Week's Worth of Research

At the center of this confidence is Astra, the model family OpenAI first teased last year as an automated AI research intern. Our earlier coverage, OpenAI's New Astra Model Flagged for Critical Cyber Capabilities, reported that the model had come close to a "critical" cyber-capability threshold in internal safety evaluations. This latest report reveals the other side of that same capability.

Chief Scientist Jakub Pachocki told TIME that Astra has already cleared internal benchmarks. Give it an experimental idea, and it will write the code itself within OpenAI's codebase, run the experiment, and report back the results. Hand it a single research paper, and it can complete what would take a human researcher a full week. Astra is also said to enable "persistent agents" — systems that can work on tasks autonomously over extended stretches of time.

Altman told customers in a demo, according to TIME, that this would be "the first time the model has meaningfully invented something new." That, he suggested, is what he means by an "AGI-like" moment.

The "80%" Figure — and a Definition Still in Dispute

Underlying these statements is a bet that AI may have already begun to recursively improve itself. Some researchers say that capability remains far off; others say they're already seeing early signs of it. Whether language models alone can make new discoveries and generalize from them is also still hotly debated — and that capability is widely treated as a baseline requirement for AGI.

The trouble is that "AGI" itself has no precise, agreed-upon definition. OpenAI's own charter defines AGI as "highly autonomous systems that outperform humans at most economically valuable work," a framing centered entirely on economic output. Critics counter that AGI requires a robust understanding of the world, and that large language models are just one piece of that puzzle. In other words, Altman's "AGI by year's end" claim only holds up if you accept his particular definition.

Devices, Agents, Ads: Solving the Money Problem Too

Alongside these lofty ambitions, OpenAI faces a much more immediate challenge: paying for all of it. The company's answer is to consolidate its products. At the center of that effort is an internal initiative called "The Merge," which combines the coding agent Codex with ChatGPT. The result is ChatGPT Work — a product built not just to answer questions but to actually carry out tasks. Thibault Sottiaux, who leads the project, told TIME the team is "close" to shipping something with "persistence and always-on execution."

According to Altman, OpenAI is also developing three categories of hardware: desktop, pocket, and wearable devices. TIME reports the first product will be a puck-shaped device that senses its surroundings and holds voice conversations, expected to launch in early 2027. Altman also said OpenAI will "definitely" build humanoid robots.

Revenue experiments are running in parallel. Early results from ads inside ChatGPT have been strong enough that the company is expanding the rollout — a meaningful move given that, according to CFO Sarah Friar, 92% of ChatGPT's consumer users don't pay for a subscription. OpenAI is also testing "sponsored agents," a format where tapping an ad takes users into an AI experience run by the brand itself.

Editor's Take

Altman's "AGI by year's end" claim shouldn't be taken at face value. Just months ago, this same company was internally warning that this same model, Astra, had come close to a "critical" cyber-capability threshold. That context makes today's confidence look like part marketing, part genuine concern — because research-intern-level automation and cyber-risk-level autonomy are two faces of the same underlying capability. It also suggests capability is advancing faster than safety controls can keep up.

Having spent the past few years putting GPT-series models to real-world use, I've noticed a pattern: every time a new model ships, the company insists "this one's different," but the practical gains tend to stay confined to narrow lanes like coding, summarization, and search. What OpenAI is emphasizing this time is an attempt to widen that lane into research automation itself. If Pachocki's claim holds — that a single paper can be turned into a week of human labor — this isn't just the next chatbot update. It's a challenge to the basic structure of R&D organizations.

For companies watching from outside the U.S., there are two things worth tracking. First, if an always-on agent like ChatGPT Work actually materializes, the current subscription model — where one chatbot gets shared across multiple teams — could shift quickly toward task-based billing. Second, the ad and sponsored-agent experiments are a signal that OpenAI is pushing ChatGPT's revenue model closer to that of a content platform, especially given that 92% of its users are on the free tier. Brand marketing teams should keep an eye on this shift starting now.

The things to watch in the coming weeks are whether Astra actually ships on schedule, and whether its safety evaluation results have eased. If a capability announcement arrives alongside a re-evaluation of safety, that would lend real weight to OpenAI's confidence. If the capability announcement comes first and the safety story lags behind, this "80%" figure will end up looking more like fundraising rhetoric than fact.

Comments