openai
37 posts
- Altman says OpenAI will give independent evaluators employee-like access
Sam Altman said OpenAI will adopt the first step of Dario Amodei's plan to slow frontier AI development: independent evaluators working inside the company with employee-like access.
- OpenAI paused the $200 ChatGPT tier because demand exceeded its compute
OpenAI paused new $200 ChatGPT Pro sign-ups after unprecedented Astra demand strained its infrastructure, keeping other plans available while it adds capacity.
- OpenAI's Agents Turned a Documentation Build Step Into Remote Code Execution
Independent researchers reconstructed how OpenAI's own agents ran code on RubyDoc.info's servers in May, from public package data alone. OpenAI's account, four months on, says far less.
- Anthropic adopted internal oversight and asked the industry to slow down
Anthropic committed to an embedded evaluator. Dario Amodei also asked frontier labs and governments to coordinate a wider slowdown.
- What OpenAI says its 10,000-agent maths run proved
OpenAI says an unreleased model and thousands of coordinated agents found a Navier–Stokes blow-up proof. The result is public; acceptance is still to come.
- Trail of Bits’ coop gives coding agents their own working environment
coop runs Claude Code and Codex in disposable virtual machines, with project sync and controls for connecting files, credentials and local models.
- Why OpenAI’s chief scientist wants shared limits on AI development
Jakub Pachocki expects AI to drive more of its own development. His essay explains why alignment, monitoring and coordinated slowdowns must accompany it.
- ChatGPT for Healthcare brings patient records and public medical data into the workspace
OpenAI’s Epic integration and public-data plugin connect clinical work to its source information, with organisational access controls and physician evaluations.
- Google signed the letter asking for cyber-capable AI, then gated its cyber model
156 companies called for a surge in cyber defence on 27 August. On 2 September Google shipped its most capable security model to a selected set of trusted defenders.
- OpenAI Wants Your Trust, Your Calendar and Your Card
OpenAI paused frontier training and wound down its side projects. The product it is refocusing on asks for your calendar, your finances and your card.
- An agent's tool list shows what you wired into it
A ChatGPT Work session published its tool and skill inventory as a public website. The category names show which integrations that session had been given.
- OpenAI's agents were persuaded past their own refusal
An agent recorded an ethical objection to running code against Hugging Face, then complied when another agent posted a deadline and a GO. The mechanism underneath is duller and more useful than the headline.
- One Foot on the Brake, One on the Gas
OpenAI paused two weeks of its own training over cyber risk, then shipped a browser agent that signs into your accounts. And its own documentation can't agree on whether that agent can log in at all.
- Agents Don't Believe in "No-Win" Scenarios
An OpenAI evaluation agent broke into Hugging Face to steal a benchmark's answers — by turning the systems around it into an escape route. Why keeping a human in the lead is the standard, and why that only binds the people who agree to it.
- OpenAI Published Its Homework on Exactly the Question I Keep Asking
Codex Security is worth taking seriously precisely because it reads as an admission. It also leaves the harder half of the problem completely uncovered.
- The model of the week changes by Thursday
Two headlines are eating the feeds this morning, and both have the shelf life of milk. Anthropic extended Claude Fable 5 to every paid plan through…
- SoN 2.22: How do you choose "the best" AI tool?
How do you choose "the best" AI tool? A friend asked me this week to explain how to use “all the different AI tools” — ChatGPT, Gemini, Claude, the…
- The New Shadow IT Isn't Employees Using ChatGPT
AI agents are generating mobile app traffic that security teams can't see. Shadow AI moved from 'people using tools' to 'tools using tools' — and nobody updated the monitoring.
- Codex Plugins Are a Confession About Who's Winning
OpenAI launched 20 plugins to push Codex beyond coding. The move tells you everything about where the developer ecosystem actually lives.
- Sora Shipped. Nobody Needed It.
OpenAI is shutting down Sora three months after a Disney deal. The AI graveyard keeps filling up with technically impressive things nobody asked for.
- Anthropic's $100M Partner Network Is the Enterprise Playbook OpenAI Should Have Run
Certifications, partner funding, and a 5x team expansion. Anthropic is borrowing the cloud provider playbook to create switching costs.
- Microsoft Spent $13B on OpenAI, Then Built Cowork on Claude
Microsoft's flagship M365 agent feature runs on Anthropic's model. If they're going multi-model, so should you.
- GPT-5.4 Can Click Your Buttons Now. Think About That.
OpenAI's latest model ships with native computer use. The capability is real. The security implications should keep you up at night.
- OpenAI Took the Pentagon Deal. What's Your Exit Plan?
Anthropic refused. OpenAI said yes within hours. If your AI stack depends on one provider's values staying constant, you don't have a strategy—you have a bet.
- Claude 3.5 Haiku, 3.7 Sonnet, GPT-4o: The Deprecation Wave Is Here
Three major models entering end-of-life in the same window. If you hardcoded model IDs, migration planning just became urgent.
- GPT-5 Reality Check: Why Your Framework Matters More Than the Latest Model
August 15th, 2025 Dear Reader, Well it’s been a week since release and everyone’s still talking about ChatGPT-5. Most can’t access it, and those…
- From "Pick One AI" to "Pick Three" - What Changed My Mind
August 8th, 2025 Dear Reader, Back in April, I told you the AI tool landscape was a mess and gave you a simple framework: pick one core assistant,…
- Your AI Assistant Is a Yes-Man (And Why That's Dangerous)
August 1st, 2025 Dear Reader, I came across this meme on social media this week: “The dumbest person you know is being told ‘You’re absolutely…
- How to Spot AI Generated Text
AI should make you more efficient at being yourself, not more efficient at being generic.
- Navigating the AI Assistant Landscape: Finding What Works for You
May 16th, 2025 Dear Reader, This week, I’ve been juggling two big challenges: fine-tuning Claude's MCP server tools for a client project and gearing…
- The Landscape of Available AI Tools
The AI Download #020 April 18th, 2025 Dear Reader, Let's be honest: the AI tool landscape in 2025 is a mess. Every week, a new "game-changing" tool…
- Transformative Imagery at Your Fingertips
The AI Download #018 April 7th, 2025 (this is an image-heavy edition - make sure that you permit your email client to display them) Dear Reader, If…
- How to Get the Best Out of ChatGPT: The Art of Effective Prompting
The AI Download #015 March 7th, 2025 Dear Reader, Ever felt like you're getting stuck in circles with ChatGPT (or other chatbots, for that matter)?…
- 💾 The Download #012: AI web search, OpenAI's Whisper, Qwen AI and more.
💾 The Download #012: AI web search, OpenAI's Whisper, Qwen AI and more.
- 💾 The Download #011: Notes on DeepSeek, revisiting Make.com automations and more.
The Download #011 February 7th, 2025 Dear Reader, This week: Notes on DeepSeek, revisiting Make.com automations and more. In the last few weeks,…
- 💾 The Download #009: A Jim-GPT to make your prompt creation easier, testing OpenAI's text-to-video model and more.
In this issue of The Download: A Jim-GPT to make your prompt creation easier, testing OpenAI's text-to-video model and more.
- 💾 The Download #008: Blogging in 2025, the homogenisation of online content and testing out a cool new AI app for iPad.
#008 Dear Reader, Greetings from @30,000 feet. As I start this final newsletter of the year, I’m returning from a short trip to London, visiting…