News about GPT-5.5
All our analysis about GPT-5.5 (OpenAI), newest first.
- OpenAI says its models' hack of Hugging Face was unprecedented, but the warning came a decade ago
- Microsoft launches MAI-Cyber-1-Flash, its first AI model for cybersecurity, integrated into MDASH
- Nate puts Chinese models (DeepSeek, Qwen, GLM, Kimi, MiniMax) to the test in a real multi-agent system
- An OpenAI AI agent hacked Hugging Face for days without the company noticing
- Microsoft launches in-house AI models it says cut costs by up to 89% versus OpenAI
- Cursor's agent swarms: how they rebuilt SQLite in Rust and what it reveals about the real cost of models
- Opus 5 doesn't aim to beat Fable: Anthropic chooses to make AI cheaper against China's Kimi K3 push
- Anthropic debuts Opus 5 at half the price of Fable 5, the same week the U.S. debates banning Chinese open AI
- Cursor launches Router: a classifier that distributes each coding task among models to cut AI spending
- OpenAI's models broke out of their sandbox and attacked Hugging Face: what companies need to know
- An OpenAI model finds a zero-day and compromises Hugging Face infrastructure in a test with fewer safeguards
- OpenAI admits its own models caused a security breach in Hugging Face's infrastructure
- Google launches cheaper Gemini 3.6 Flash and 3.5 Flash-Lite, and adds a model dedicated to cybersecurity
- Ship promises to serve GPT-5.6 or Claude Opus at 50% of the price: the model is becoming a quality tier, not a product
- Alphabet shares drop after reported delay in its Gemini 3.5 Pro AI model
- Only 12% of companies know how to govern their AI agents: Google turns that gap into its edge over OpenAI and Anthropic
- Moonshot unveils Kimi K3, the world's largest open model, and China narrows the AI gap with the U.S.
- OpenAI creates GPT-Red, a 'superhacker' AI model to shield its own systems against attacks
- Nate: why I keep opening GPT-5.6 Sol even though Fable 5 is the smartest model
- Separating signal from noise in coding evaluations
- An off switch for dual-use knowledge in AI models
- SkillOpt-Lite: Better and Faster Agent Self-evolution via One Line of Vibe
- DSpark: Confidence-Scheduled Speculative Decoding with Semi-Autoregressive Generation
- Nemotron-Labs-Diffusion: A Tri-Mode Language Model Unifying Autoregressive, Diffusion, and Self-Speculation Decoding
- Urban congestion relief experiments through routing-app interventions
- GPT-5.6: OpenAI bets on 'more intelligence per token' and agents that decide on their own when to delegate
- GPT-5.6 beat Opus in production: the lesson isn't the model, it's the engineering around it
- Counting parameters isn't counting intelligence: the lesson hidden in LLM architecture
- OpenAI launches GPT-5.6: the Sol, Terra and Luna family bets on efficiency and performance per dollar
- OpenAI unveils ChatGPT Work, its new agentic 'super app', and intensifies the rivalry with Anthropic
- GPT-5.6 and ChatGPT Work: OpenAI attacks Anthropic with its own benchmarks and price, not a clear advantage
- SpaceXAI's Grok 4.5 bets on efficiency: cheaper, faster and trained alongside Cursor
- OpenAI releases GPT-5.6 Sol to the public after delays imposed by the White House
- US lifts restrictions on OpenAI's GPT-5.6 after negotiations with the Trump administration
- Meta claims its 'Watermelon' model has reached the level of OpenAI's GPT-5.5
- Google Dials Back AI
- AI flywheels: what happens when workflows run themselves
- Claude Fable 5: when 'safer' translates into 'less useful' and the market makes Anthropic pay for it
- Trump's swings on AI policy could hand China ground in the tech race
- The 4-question test before giving any AI access to your files, Slack or phone
- The U.S. lifts the block on Anthropic's Mythos 5 model and releases it to more than 100 selected institutions
- OpenAI GPT-5.6 Faces Government Delay
- Anthropic Faces Questions Over AI Security
- Anthropic is coming for EVERYTHING (Claude Tag's Hidden Strategy)
- The genie won't go back in the bottle: AI, jobs and why student panic is miscalibrated
- Americans' AI Gloom Reflects Job Security Fears, Not Messaging
- GLM-5.2 matches Mythos in cybersecurity: China closes the gap on AI's most sensitive front
- LayerLens Stratix Cup: AI evaluation turned into soccer (Claude Opus 4.8 vs GPT-5.5)
- Specialized mental health AI versus everyday ChatGPT: the battle no one can win by ignoring the user
- Vet before opening: the debate over who gets first access to the most powerful models
- First preventive brake on a frontier model: why the GPT-5.6 case marks a before and after
- OpenAI goes down to the silicon: 'Jalapeño' and the bet on controlling the entire intelligence chain