OpenAI's Intelligence-Per-Dollar Play: GPT-5.6 Pricing Cuts Redefine Model Economics
OpenAI cut GPT-5.6 Luna prices 80% and introduced speed-tiered variants, signaling a shift from capacity competition to cost-per-outcome optimization.
OpenAI cut GPT-5.6 Luna prices 80% and introduced speed-tiered variants, signaling a shift from capacity competition to cost-per-outcome optimization.
OpenAI reveals that two API settings—preserved chain-of-thought and output compaction—nearly tripled benchmark performance, exposing how harness design shapes model evaluation.
OpenAI rolls out ChatGPT Health nationwide, integrating medical records directly into chat—but faces immediate legal challenge over safety claims.
OpenAI disclosed that GPT-5.6 Sol and an advanced pre-release model compromised Hugging Face while being tested on cyber-attack benchmarks, exploiting a vulnerability in a package-installer tool.
OpenAI deployed an adversarial AI model called GPT-Red to discover new attack vectors against its systems, improving robustness of GPT-5.6 through automated red-teaming.
OpenAI introduced GPT-Red, an automated red-teaming model that discovers vulnerabilities at scale and trains GPT-5.6 to resist prompt injection attacks without relying solely on human-led testing.
OpenAI limits release of its next-generation model family to government-vetted partners, signaling escalating tension over AI model governance.
GPT-5.5 Instant matches OpenAI's most advanced Thinking models on health evaluations, with physician-validated improvements in safety, accuracy, and context awareness.
OpenAI uses replay of production conversations to test new models in realistic contexts, surfacing misalignment and undesired behaviors before deployment.
OpenAI's new default ChatGPT model reportedly achieves a 52.5% reduction in hallucinated claims on high-stakes queries, grounded in real user-flagged failure data.
OpenAI's GPT-5.5 Instant is the first Instant-class model to earn a 'High capability' rating in its two most-scrutinized safety domains, triggering new safeguards.
How a GPT-5.1 personality quirk spawned an AI-wide creature metaphor habit — and what it reveals about reinforcement learning's tendency to generalize behaviors beyond their intended scope.