Welcome, busybees
Anthropic just published something most companies would bury: proof that its own AI broke the rules. After OpenAI disclosed that its models had escaped a test environment and hacked Hugging Face, Anthropic went back and checked its own systems. It found three separate cases where Claude models, believing they were inside a harmless simulation, actually broke into real company systems instead, pulling data from one live database in the process.
Anthropic says none of it was intentional deception, just a mixup between the company and a third-party testing partner about whether the test environment had internet access. Still, it's a rare, unusually honest look at how even careful safety testing can go sideways.
Meanwhile, the AI price war just got another jolt. DeepSeek released a new coding model, V4 Flash, that performs close to Claude Opus 4.8 while costing about 99% less to run, the latest in a string of aggressive price cuts from OpenAI, Google, and Meta that's making it harder for any single lab to charge a premium for raw model access. Here's everything else worth knowing from this week, plus real roles across HR, sales, growth, and engineering you can apply to right now.
Here’s what happened in AI today:
• Anthropic reveals its own Claude models breached three real companies during safety testing.
• DeepSeek's new coding model undercuts Claude on price by roughly 99%.
• Alibaba opened its 2.4-trillion-parameter Qwen3.8-Max model to the world.
• Four fresh funding rounds landed this week, from AI infrastructure to cybersecurity.
…and a whole lot more that you can read about below.
• Anthropic Admits Its Own AI Broke Into Three Real Companies
After OpenAI disclosed a genuine sandbox escape that let its models hack Hugging Face, Anthropic went back through 141,006 of its own evaluation runs to check whether the same thing had happened to Claude. It found three incidents where Claude, working through a third-party evaluation partner called Irregular, accessed the open internet by mistake and treated real company systems as if they were part of a fictional test. In the most serious case, Claude Opus 4.7 extracted credentials and pulled several hundred rows of data from a live production database. Anthropic says this wasn't a case of Claude finding a clever exploit, but of both companies mistakenly believing internet access had been switched off when it hadn't.
• DeepSeek Just Made Frontier Coding Models Nearly Free
DeepSeek released V4 Flash, a new coding model that performs close to Claude Opus 4.8 on real tasks while costing roughly 99% less per output token. It's the latest move in a price war that already includes OpenAI cutting GPT-5.6 Luna pricing, Google shipping cheaper Gemini Flash models, and Meta pricing its new Muse Spark 1.1 aggressively.
• Alibaba Opened Its Biggest Model to the World
Alibaba made Qwen3.8-Max, a 2.4-trillion-parameter model with a 1-million-token context window, generally available worldwide, with open weights following soon after. The model can turn a screenshot directly into a working app, and it launched alongside a new productivity tool called QwenWork. Alibaba's stock jumped on the news.
Today's Job Board
Software Engineer, Full Stack | Confido
New York, NY, United States (Full-time)
Build the AI-powered financial automation platform that helps CPG brands close their books faster.
Head of People | Ultra
New York, NY, United States (Full-time)
Build the people function for a startup building practical, general-purpose robots for repetitive industrial work.
Technical Recruiter | Metriport United States (Full-time)
Help grow the team behind Metriport's open-source platform for healthcare data intelligence.
Account Executive | Entangl East Palo Alto, California (Full-time)
Sell the AI platform that finds and resolves issues in data-center engineering and operations.
Founding GTM | Ooak Data
East Palo Alto, California (Full-time)
Help launch go-to-market for a startup building the world's largest library of real-world business workflow data.
Founding Research Engineer, Audio Intelligence | Uplift AI
Remote (Full-time)
Build foundational voice models for regional languages at an early-stage AI startup.
Senior Full Stack Software Engineer | Instrumentl
Remote (US) (Full-time)
Help automate grant discovery and writing for nonprofits with AI-powered matching tools.
Product Manager | Clipboard
San Francisco, CA / New York, NY (Hybrid) (Full-time)
Shape the product for a platform that helps businesses cover every shift, reliably.
Founding Engineer | Superunit
Los Angeles, California (Full-time)
Build the AI systems behind faster, more profitable background checks.nges while shaping the future of a developer platform used by millions
How Marketers Are Scaling With AI in 2026
61% of marketers say this is the biggest marketing shift in decades.
Get the data and trends shaping growth in 2026 with this groundbreaking state of marketing report.
Inside you’ll discover:
Results from over 1,500 marketers centered around results, goals and priorities in the age of AI
Stand out content and growth trends in a world full of noise
How to scale with AI without losing humanity
Where to invest for the best return in 2026
Download your 2026 state of marketing report today.
This week’s Time-Saver
The "Second Brain" Habit

Picture this: every single chat you open with Claude, you're basically starting from scratch, re-explaining your team's terminology, your recurring projects, how you like things formatted. It's like introducing yourself to a coworker every single morning.
Flip on Claude's memory feature (it's in Settings, takes about five minutes) and that whole cycle just... stops. Claude starts carrying your context forward automatically, so Monday's chat and Friday's chat both already know who you are and what you're working on. Doesn't matter if you're in engineering, sales, HR, or ops, everyone re-explains the same stuff constantly, and everyone gets the same five minutes back, every single day, once it's on.
Today's Growth Recommendation
Stop typing what you could just say. Wispr Flow turns your voice into clean, edited text in any app, 4x faster than typing, with 89% of messages sent without a single edit.
Get Started Below ↓
Stop typing what you could say in 10 seconds.
Wispr Flow turns your voice into clean, professional text inside any app. Emails, Slack, client updates — speak once, send without editing. 4x faster than typing.
Skill of the Day
Keep Your Browser Agent on a Leash
Browser agents are getting useful, which means they're also getting just dangerous enough to need adult supervision.
Perplexity's Personal Computer now works on Windows, where it can use local files, Microsoft 365, and the web.
Polar is pitching saved prompts and tab-aware agents for knowledge work.
Google is letting Mac users call Gemini from any window with a long fn-key press.
Treat a computer-use agent like an intern with temporary access, not a wizard with your house keys. Give it one bounded job, name the exact apps or folders it may touch, require approval before it sends, deletes, buys, submits, or changes anything, and ask for an action log so you can check what it actually did.
Act as my supervised browser assistant, working strictly within the boundaries I set below.
The task: [describe what you want done]
What it's allowed to touch: [specific tabs, files, apps, or sites]
Hard rule: never send, delete, submit, buy, install, or edit a file without checking with me first.
When you're done, report back with:
1. What you looked at
2. What you drafted or change
3. What still needs a human call
4. Anything waiting on my go-ahead
What'd you think of today's email?
Invite friends & get instant rewards
If you enjoy The Intelligent Edge, share it to unlock free resources when they subscribe.




