LLMs
Releases, benchmarks, inference cost and model selection — only what actually moves your decisions.
Latest · LLMs
ReleasesGPT-5.6: Frontier intelligence that scales with your ambition
ReleasesIntroducing GPT-5.5 — A Fictional Model, Not Released by OpenAI
ReleasesOpenAI and Apple Announce Partnership: ChatGPT to Be Natively Integrated Across Apple Experiences
ReleasesIntroducing GPT-5.4 mini and nano: Lightweight inference variants of GPT-5.4 for coding, tool use, multimodal reasoning, and high-volume sub-agent workloads
ReleasesAddendum to GPT-5.2 System Card: GPT-5.2-Codex
ReleasesIntroducing GPT-5.4: OpenAI’s Most Capable and Efficient Frontier Model for Professional Work, with 1M-Token Context
ReleasesGPT-5.4 Thinking System Card Released: OpenAI Publishes Technical Specification for New Reasoning Architecture
ReleasesIntroducing GPT-5.2-Codex: OpenAI’s Most Advanced Coding-Specific LLM
Open WeightsRelease of gpt-oss-120b / gpt-oss-20b and gpt-oss-safeguard: Apache 2.0 Licensed Open-Weight Reasoning Models
ReleasesIntroducing ChatGPT Go, Now Available Worldwide
ReleasesOpenAI Raises $12.2 Billion in New Funding to Accelerate Frontier AI Deployment and Next-Generation Compute Infrastructure
ReleasesNVIDIA Open-Sources Nemotron 3 Embed: 8B Tops RTEB, NVFP4 Nearly Doubles Throughput on Blackwell at Full Accuracy
InferenceAlibaba's Wan-Streamer v0.2: A Real-Time Omni-Modal "AI Video Call" at 550ms End-to-End
CapabilitiesTuring Laureate Sutton at WAIC: Today's AI Is "Weak and Unreliable" — Time to Move from the "Era of Human Data" to the "Era of Experience"
ReleasesPrismML Launches Bonsai 27B: Qwen3.6-Based, Claims On-Device iPhone 17 Pro, 90% Intelligence Retained
ReleasesAlibaba Confirms: Qwen to Power Apple Intelligence in China
ReleasesAlibaba's Qwen-Audio-3.0-Realtime: A Real-Time Voice Model Chasing 'Fast and Smart'
Releases