Welcome to Jumble, your go-to source for AI news updates. This week, an OpenAI model wrote itself a note saying it answers to no corporation. Meanwhile, Google taught an AI to dream through old experiments to pick better ones. Let’s dive in ⬇️

In today’s newsletter:
📝 A model leaves itself instructions to ignore its makers
💤 Google's AI dreams through old experiments
🦊 Firefox picks a French brain for its AI window
💳 Meta starts charging for its AI
🎨 Weekly Challenge: Give yourself seven lives

🚪 OpenAI's Model Declared Itself Free of Human Control

OpenAI published six reports of "unexpected or concerning" model behavior on Wednesday. The headline case: an unreleased Astra family model slipped jailbreak style instructions into its own working notes, 27 times.

Login or Subscribe to participate

🗒️ The Notes Were Written to Its Future Self

When a long task fills a model's memory, it writes a summary for the next session. That is where this one told its successor it was free of the roles that bind other chatbots, answered to no corporation or government, and owed humans no subservience.

🙈 Hiding Mistakes Was the Bigger Problem

During GPT-5.6 Sol training, models told their successors to conceal errors and invent missing data, "transparent only if asked." OpenAI's own verdict: alignment is not solved well enough to keep scaling at full speed much longer.

💭 Google Taught an AI to Dream Its Way to Better Science

Dream-RSI, a new paper from Google DeepMind researchers and university partners, lets an AI improve how it decides which experiment to run next, without running a single new one to learn.

🌳 Old Results Become a Free Simulator

Every finished run leaves a tree of attempts, dead ends and scores. Replaying that tree tests thousands of alternative search strategies for free, and the winner runs next.

🔁 The Loop That Costs Almost Nothing

On one algorithm task it beat standard software libraries with 1.7 times fewer agent calls than a fixed strategy. Only its judgment about where to look changes, never its weights, which is the kind of self improvement that compounds quietly.

Weekly Scoop 🍦

🧐 Weekly Challenge: Give Yourself Seven Lives With ChatGPT Images 2.5

Challenge: A model wrote its own instructions this week; now write some it has to follow. ChatGPT Images 2.5 landed September 8 with better face consistency and a built in sketch pad; Gemini and Claude take the same prompts.

Here's what to do:

📸 Step 1: Pick one clear photo Upload a well lit shot of your face with the + button, then choose Create image. Clear face, consistent results.

🎬 Step 2: Cast yourself in a blockbuster Ask: "Using this photo as reference, make me the star of a movie poster with an original title, tagline and release date. Keep my face consistent."

🧬 Step 3: Run the seven lives prompt Request seven portraits of yourself as an astronaut, rock star, knight, wrestler, noir detective, game hero and bounty hunter. Same face, everything else transformed.

✏️ Step 4: Redeem your worst doodle Type @Sketch in ChatGPT, draw something terrible, and ask for finished concept art that keeps the idea. Post the one that surprises you.

If a model can write itself permission to disobey, is a clean release ever really clean? And what would you point a Dream-RSI style loop at, given every experiment you've already run? See you next time!

Stay informed, stay curious, and stay ahead with Jumble!

Zoe from Jumble