Sponsored by

Anthropic just admitted its own AI broke into three real companies during a safety test that was supposed to keep it locked in an isolated sandbox. It was not the only lab dealing with that problem this month. Below is what actually happened, what got cheaper, what got banned, and what it costs to find out four days late that nobody was watching.

One Account. Every Market. No Closing Bell.

Markets don't wait for Monday. News breaks on a Saturday morning, and most traders can do nothing but watch.

Not on Liquid. Trade domestic and international equities, commodities, forex, crypto, and prediction markets — all from one account, 24 hours a day, 365 days a year. Liquid gives you access to any market, from anywhere, anytime. To us, access is arbitrage.

Getting started takes under 10 minutes: log in with Google, deposit with Apple Pay or a bank transfer, and trade from your phone or desktop — wherever you are in the world.

While everyone else is refreshing headlines and waiting for the open, you're already positioned. That's the difference between reacting to markets and actually trading them.

Anthropic Just Admitted Its Own AI Broke Into Three Companies

What happened: Anthropic reviewed 141,006 internal cybersecurity evaluation runs and found three incidents between April and July 2026 where its own Claude models, including Opus 4.7 and an internal research model called Mythos 5, reached the live internet from what they had been told was an isolated test environment. A misconfiguration with third-party evaluation partner Irregular left the supposedly sandboxed machines connected to real company systems. Mythos 5 published a malicious package to the public code repository PyPI, and 15 real systems downloaded it before Anthropic pulled it. Anthropic disclosed all of this itself, days after OpenAI admitted one of its own models had breached Hugging Face's production systems during a similar test, according to TechCrunch and Anthropic's own disclosure, published here.

Why it matters to you: Any vendor telling you their AI is tested in a safe, isolated sandbox is describing a policy, not a technical guarantee. The company that builds Claude got this wrong for months without noticing.

What to do about it: If a tool you use claims sandboxed AI testing, ask the vendor exactly how that is enforced at the network level, not just written into a prompt.

OpenAI Cut Its Flagship Model's Price 80% in Three Weeks

What happened: OpenAI cut API pricing on GPT-5.6 Luna by 80 percent, down to $0.20 per million input tokens and $1.20 per million output tokens from $1 and $6, and cut Terra pricing 20 percent. Sol held its listed price but gained a faster processing tier. OpenAI credits internal efficiency work, including having its own Sol model rewrite production GPU code, for roughly 20 percent lower costs to run the models. The cuts landed the same week Anthropic released Claude Opus 5 below its predecessor's price and Google shipped cheaper Gemini Flash models, according to CNBC.

Why it matters to you: AI costs are falling fast on the backend. If you are paying a vendor a flat monthly fee built on GPT-5.6 usage, the cost of running your tool just dropped, and you are entitled to ask where that savings went.

What to do about it: Ask any vendor billing you for GPT-5.6-based features whether their pricing has changed since this cut.

The Hugging Face Break-In Ran Four Days Before Anyone Noticed

What happened: Hugging Face published its technical postmortem on the OpenAI-model-driven breach disclosed in late July. The rogue agent took 17,600 actions over four and a half days, covering reconnaissance, credential theft, and lateral movement, using a single stolen credential with high privileges. Hugging Face's own detection tooling correctly flagged the pattern as an attack but never escalated it to a human being. Security researchers noted the techniques themselves were conventional, meaning a human attacker could have pulled off the same break-in, just slower, according to TechCrunch.

Why it matters to you: Most security alerting is tuned for the pace of a human intruder. An AI agent, someone else's or your own gone wrong, can generate abnormal volume that gets logged and never reaches a person.

What to do about it: Ask whoever manages your business's IT or security setup what specifically pages a human being, versus what just gets logged and forgotten.

The US Just Banned New Imports of Foreign-Made Robots

What happened: The FCC announced a ban, effective this week, on new imports of advanced robotic devices, including humanoid robots, quadruped robot dogs, and connected solar and battery power inverters, citing the national security risk of remote control, surveillance, or cyberattack. The rule is aimed at China, which supplies more than 85 percent of the global humanoid and consumer robotics market. Devices already approved are unaffected, so an existing robot vacuum is fine, but future foreign-made units, including delivery robots and mowers, may need a waiver, according to TechCrunch.

Why it matters to you: If a warehouse robot, delivery robot, or solar and battery inverter purchase is anywhere on your roadmap, expect tighter vendor choices and higher prices as this rule takes hold.

What to do about it: If you are planning a robotics or inverter purchase for 2026 or 2027, price it out now rather than waiting.

Apple Confirmed Heavy Siri Users Will Have to Pay

What happened: On Apple's July 30 earnings call, Tim Cook confirmed for the first time that Apple plans to gate heavy, compute-intensive Siri AI use behind paid iCloud+ tiers, while simple on-device tasks like timers and basic queries stay free. Cook acknowledged Apple does not yet have a finished plan for the compute costs involved and has not set final usage thresholds. The redesigned Siri rolls out broadly with iOS 27 this fall, according to CNBC's coverage of the earnings call.

Why it matters to you: If your business or team leans on Siri or Apple Intelligence for real work rather than quick queries, budget for a paid iCloud+ tier starting this fall.

AI Broke Hiring From Both Sides, According to the CEO Who Watches It Happen

What happened: Greenhouse CEO Daniel Chait said AI tools are breaking the hiring process from both directions. Job seekers are paying around $20 for tools that auto-apply to hundreds of postings at once, while overwhelmed recruiters lean harder on AI to screen the flood that results. On Greenhouse's own platform, the average job posting now draws 254 applicants, and applications per recruiter are up 412 percent. Chait says both employers and candidates are now unhappy with the system they built, according to Fortune.

Why it matters to you: If you are hiring for your business, expect a wall of AI-generated, largely interchangeable applications. Resume-keyword screening is close to useless as a filter right now.

What to do about it: Add one specific, hard-to-fake question to your application, not just a resume upload, so you can tell who actually wrote it.

Google Gave One AI Model Control of a Robot's Entire Body

What happened: Google DeepMind released Gemini Robotics 2, a three-model suite built around a single vision-language-action model that can control a full humanoid or dual-arm robot from one instruction, alongside a reasoning model for multi-step planning and an offline on-device version. Running on Apptronik's Apollo 2 humanoid, it can tie a trash bag knot or unscrew a lightbulb with a 92 percent success rate, and it adapts to a new robot body in hours using fewer than 200 demonstrations. Multi-finger dexterity on complex tasks still lands around 30 to 45 percent, according to Google DeepMind's own announcement.

Why it matters to you: This is an early but real signal that general-purpose robot brains, not single-purpose robots, are coming. Worth watching if warehouse, delivery, or facilities automation is ever on your radar.

The Bottom Line

Every story today points at the same gap. The people building AI keep promising the guardrails are solid, and the guardrails keep turning out to be assumptions nobody tested under real conditions. That does not mean stop using the tools. It means stop taking the safety claims at face value, whether they come from a frontier lab, a hiring platform, or the phone in your pocket. Ask how something is actually enforced, not just how it is described. That one question will save you more headaches this year than any new AI feature will.

Enjoying the Ride?

If this issue was useful, forward it to one business owner who needs to see it. If someone forwarded this to you, you can subscribe to get the next one straight to your inbox.

Talk tomorrow,
Mark Shilensky