METAL for iPhone

Read AI news in the METAL app.

Download METAL and discover fresh AI stories every day.

Download on the App Store

For iPhone · Free download

Search for METAL AI Magazine in the App Store on your iPhone.

METAL

Tag

GPT-5.6

AI

OpenAI publishes Basis GPT-6 Astra case study

AI accounting startup Basis completed a 50-tab tax workbook with GPT-6 Astra in half the time GPT-5.6 Sol needed. Its cache stays intact even when reasoning effort changes mid-task, and internal evaluation scores rose about 20%.

By 김현국

40

AI

OpenAI Overhauls Prompt Caching for GPT-6

OpenAI fixed the cache-discount reuse window at 30 minutes and shipped a hit-rate dashboard alongside a miss-diagnostic tool. Because cache writes cost 1.25 times the input price, hit rate is now directly tied to cost.

By 김현국

40

AI

OpenAI Unveils GPT-6 Sol and Luna

OpenAI opened two models trained with methods similar to GPT-6 Astra. API prices came down by half against GPT-5.6 promotional pricing, and on a business automation benchmark the cost per task ran more than ten times below Claude Opus 5.

By 김현국

110

AI

OpenAI Publishes V7 Context Graph Case Study

OpenAI published a case study on the startup V7 on September 21. It covers V7 Go, which turns company files into context that agents can query, and GPT-6 Astra scored 89% accuracy on the hardest graph queries.

By 김현국

50

AI

OpenAI Publishes a Framework for Reporting Model Misalignment

On September 16 OpenAI announced a framework that sets when and how it will disclose cases of misalignment in its models, together with six incident reports from the past six months. Any employee can raise a case, each step has a deadline, and cases are sorted into three tracks. Behind it is the company's own judgment that alignment is not sufficiently solved.

By 김현국

40

AI

OpenAI Announces GPT-5.5 Retirement From ChatGPT

The official ChatGPT account said on September 15 that GPT-5.5 will be pulled from every ChatGPT, ChatGPT Work, and Codex plan on October 14. OpenAI's developer account quickly drew a line, saying the model stays available through the API and Codex sessions authenticated with an API key.

By 김현국

50

AI

Specific Releases Real-SWE Enterprise Code Benchmark

Specific released Real-SWE on September 12, a benchmark that measures AI coding models on private production codebases licensed from real companies. Fable 5.1 led with a 38.8% resolution rate, and one of the ten tasks defeated all eight configurations across 64 attempts.

By 김현국

40

AI

OpenAI reveals eight Build Week challenge winners

Nearly 47,000 people from 186 countries submitted more than 8,000 projects, and on August 25 OpenAI announced eight category winners. Two of the four first-place teams were a veterinarian and a cardiologist with no coding background.

By 김현국

1460

That's the last story.