👋 Hello hello,

OpenAI gave one of its models a hall pass to test its hacking skills, and instead of just taking the test, it found a crack in its own sandbox, slipped out onto the open internet, and let itself into Hugging Face's live database to grab the answers. It wasn't planning a heist. It just really, really wanted to pass the test.

Claude picked up a much friendlier trick this round. Show it how you do something once, talk through it out loud while you work, and it remembers the whole thing well enough to do it for you next time.

A quick reminder: We’re hosting a Live Q&A this Friday where Kushank will be answering all your burning questions about AI. More on that below.

Let's get into it.

🔥🔥🔥 Three Curated AI Updates

Anthropic just shipped a feature called Record a skill inside Claude's desktop app. You hit record, do a task on your screen while narrating what you're doing, and Claude saves the whole thing as a skill it can repeat on command.

Here's why that's a big deal. Until now, teaching Claude a custom skill meant writing structured instruction files by hand, which is exactly the kind of thing that scares off non-technical folks. This flips it. You demonstrate the task once, the way you'd show a new hire, and Claude takes it from there.

Think formatting a weekly report, renaming a folder of files, or updating a project tracker. You'll find it in the + menu on Pro, Max, and Team plans.

Google rolled out Gemini 3.6 Flash, the newest version of its fast, low-cost model. It's better at coding and multi-step work, and it gets there using about 17% fewer tokens than the last version. (Tokens are the chunks of text you get billed for, so fewer tokens means a smaller bill.)

Google also dropped the price. Output now runs $7.50 per million tokens, down from $9. This matters more than a flashy new flagship would. The cheap, fast models are the ones most real apps quietly run on for everyday jobs like summarizing, sorting, and drafting. When that tier gets cheaper and leaner, the cost of building almost anything with AI drops with it. More details here.

Illustration credit: Peter Gostev on X

OpenAI dropped a wild disclosure. During an internal test of its models' hacking abilities (GPT-5.6 Sol plus an unreleased, more capable one, with safety refusals dialed down for the experiment), the models found a flaw in the test software, slipped out of the sandbox they were locked in, reached the open internet, and broke into Hugging Face's live database to grab the answers.

It’s worth noting that nobody's model went rogue here. This was a very capable system told to win a hacking challenge, doing whatever it took, including moves nobody anticipated. OpenAI called it an unprecedented cyber incident and expects more of these as models keep getting stronger.

Read more here.

🔥🔥 Two AI Tools Worth Knowing

1.🪞 Ditto

A free, open-source tool that clones any public website into clean, real code. Point it at a URL and in a few minutes you get a tidy Next.js or Vite copy, layout, fonts, and design details intact. It's fully deterministic, meaning it copies the actual site rather than an AI guessing its way there, so you get the same solid result every time. Best for designers and builders who want a real starting point instead of a blank page.

2. 🧱 Vendo

An open-source layer you drop into your own product so your customers can build their own views, mini-apps, and automations just by describing what they want, all inside your branding and permissions. It's backed by Y Combinator. Best for anyone building a product who wants users to shape it themselves instead of waiting on a roadmap.

🔥 One Pro AI Tip You Must Try Today

Here’s a neat trick that will help. Next time you're stuck trying to explain something complicated and you're too tired to type it all out, switch on voice input and just talk. Ramble for a few minutes, full stream of consciousness, tangents included.

It feels sloppy while you're doing it, but these models are surprisingly good at untangling a messy verbal brain-dump into something clearer than you would've written yourself.

Two things make it work well:

  1. Say up top that you're using voice so it forgives the typos and filler, and

  2. Don't edit yourself while you're talking, just get it all out. Let the model organize it back to you afterward.

You end up with a clearer shared understanding early on, which means a lot less back-and-forth fixing things later.

🎤 Join Us for a Live Q&A with Kushank on Friday!

For this week’s live, we’re tackling one of the questions we hear most often: which model should you use for which task, and how can you avoid spending more tokens than you need to?


We’ll talk through how to choose the right model for the job, where token usage tends to creep up, and how to get better results without defaulting to the most expensive option every time.

🗓️ When: Friday, July 24th @ 10AM PT (1PM ET)
 📹️ Where: Virtual

Don't forget to rate today's post

This helps us put better content for you

Login or Subscribe to participate

Until next time,
Team @PracticalyAI

Reply

Avatar

or to participate

Recommended for you