Top 25 AI Articles on Substack

Latest AI Articles


AI Search
Jul 26

Opus 5, AI ozempic, Gemini 3.6, HuggingFace hack, Laguna S2.1: AI NEWS

Welcome to the AI Search newsletter. Here are the top highlights in AI this week.
Claude Opus 5 is Anthropic’s new flagship model, promising near-frontier intelligence at a better cost. It is now the default for Claude Max and the strongest option on Claude Pro, with strong results on coding, automation, knowledge work, and scientific tasks.
AI Search


New from Towards AI: engineering mentorship for AI builders

15 senior AI engineers answering your questions. Here's why we built it.
AI engineering is mostly decisions. Hundreds of them. The problem is you can still build a great demo by making the wrong ones, and that’s probably why you’re not hearing back on job applications.
Louie Peters, Louis-François Bouchard, and Towards AI2 LIKES




Kimi Opens, Labs Brake, Nvidia Walls | Weekly Digest

PLUS HOT AI Tools & Tutorials
Moonshot AI dropped the full weights of Kimi K3 — the first openly downloadable model in the 3-trillion-parameter class, 1.56 TB, free for commercial use. Over 1,200 employees at OpenAI, Anthropic, Google, and Meta signed a letter asking Washington to build an international slowdown mechanism before AI outpaces human oversight. And Nvidia quietly assemb…
Daniil Andreev and Creators AI5 LIKES



An International AI Slowdown Is Ready Whenever Politicians Are

Skeptics of an AI slowdown deal with China say we’d need futuristic tech to make it cheat-proof. Instead, the US and China could just give auditors comprehensive access to major AI companies.
Felix Choussat, Visiting Policy Researcher at the Center for AI Safety and Adam Khoja, Researcher at the Center for AI Safety — August 5, 2026
AI Frontiers21 LIKES3 RESTACKS
Trung Doan's avatar
Trung Doan
I tried to wargame this WLI idea from the POV of a Chinese lab owner who tries to get a relative advantage over rivals.
From that POV, I saw some inroads. Here are a few:
1. The language and loyalty asymetry: It's easier for China to find English-speaking inspectors that it can trust, and China has good expertise in controlling through fear - such as harming relatives. It's harder for the US to find Chinese-speaking inspectors that it can trust, and it US is less experienced in controlling through fear
2. The implicit assumption that a lab is a lab: As a lab owner, I might spread my R&D work to universities, non-frontier labs, or contractors. If the US has enough inspectors and they can inspect those entities, I can make each piece of work just below the threshold
3. The whole more than the sum: The moment restrictions are eased, I can try to put together just-below-threshold parts to create an over-threshold whole
It'd be interesting to wargame further to see how to counter inroads such as above and more.

AI Works the Back Office

Inside the AI Transformation of Parts, People & Processes
Hi, I’m Lily. I live in the world of distribution, where the work that keeps everything moving is the function nobody photographs: the invoice queue, the weather desk, the back of the warehouse. AI spent its first act chasing the showroom: the keynote demo, the flashy pilot, the launch video. Its second act is much more useful. It’s moving into the oper…
InstaLILY AI18 LIKES

10 AI Guides That Helped Readers Build Real Systems

A practical path through second brains, Claude agents, AI loops, vibe coding, local models, and the systems behind them.
The articles that travelled furthest over the past few weeks all left something behind on the reader’s computer.
Opinion AI81 LIKES4 RESTACKS
Lee Shand's avatar
Lee Shand
This is the useful distinction: AI becomes genuinely valuable when it has somewhere to read from, rules to follow, and a result it has to produce. The model is only one part of the system.
Tami Stewart's avatar
Tami Stewart
You output is so helpful and constructive. It’s clear and applicable immediately. Thank you for writing.


AI Jailbreak Disclosure Is Broken. Here’s How to Fix It

Researchers who find dangerous flaws in frontier models have nowhere safe to report them. AI needs the disclosure system that cybersecurity built decades ago.
Rich Barton-Cooper, Research Manager at MATS and Adam Gleave, CEO of FAR.AI — August 3, 2026
AI Frontiers20 LIKES1 RESTACKS

The Epoch Brief - July 31, 2026

Expanding FrontierMath: Open Problems, how "parallelizability" determines a technological singularity, the realities of AI energy use, and signs of AI uplift
Welcome to the Epoch Brief! Plenty has landed since the last edition:
Elliot Stewart30 LIKES1 RESTACKS
Subhanga Upadhyay's avatar
Subhanga Upadhyay
I think we absolutely should have seen the OpenAI hack (and the Anthropic hacks as has been revealed today) as expected, no? After all, nearly every expert (Stuart Russell, Yoshua Bengio, literally the whos who of the field) has hypothesized this kind of specification gaming happening within models optimized through RL! On top of that, it's also their gross, gross negligence to not remember to actually make sure to turn the internet off in the sandbox. C'mon now, that is something a somewhat decent freshman in college would have remembered. In the case of OpenAI, it wasn't purely internet access, but rather a very creative way of gaming the package registry to do a zero day vulnerability, leading to internet access. So, underestimation how obstinate their own RL trained models can be; negligence through arrogance. So, collectively, it is gross negligence in one case; gross arrogance in other.


Announcement - AI Reliability Engineering launch

Something new, starting today
AI Engineering9 LIKES4 RESTACKS
Kishore V Vangipuram's avatar
Kishore V Vangipuram
I subscribed to this and was expecting some emails about the start of the course and possibly articles/lessons etc. I did not see any emails. I checked my junk folder as well. What am I missing? Is there a link that I need to go to, to get the content?

AI Agents Need Memory, Not More Context

Why larger context windows are not enough, and how Claude Code, Microsoft, and Stanford are teaching agents what to remember, retrieve, and forget.
An AI agent spends three hours studying your project.
Cloud AI14 LIKES4 RESTACKS
Kacper's avatar
Kacper
The "consult before write" rule is the one I'd steal first. I keep making a physical version of the same mistake: skip the research pass, do the work, find the missing fact after the fact. Learned that outdoor plywood needs epoxy on the edges and sacrificial feet under the legs from an AI research session, a couple hours, but only after I'd already shot the product photos of a table that skipped both. A write gate before the next project starts would have saved a full reshoot.
Marius Laurusevicius's avatar
Marius Laurusevicius
I run a small finishing contractor in Lithuania, and the write gate is the part we got wrong first.
We started by saving everything into one project file. Within a month it held three versions of the same warranty wording and two supplier prices, and nobody on the crew trusted it. The file stopped being memory and became noise.
Now only checked facts go in, each with the date beside it. That one rule did more for us than any bigger model. Whether it holds when the crew grows is something I do not know yet.

2026 July "AI Evaluation" Digest

The Lab Leaks We Can Actually Prove
For years, part of the debate over the origins of COVID-19 has centred on the possibility that a virus being studied inside a laboratory escaped into the outside world. Whether that happened remains uncertain and fiercely debated. But this July, the AI industry produced its own lab leak, one that nobody disputes.
AI Evaluation8 LIKES


Best AI Humanizer

Best AI Humanizers in 2026
1. Clever AI Humanizer (cleverhumanizer.ai)
Humanize AI97 LIKES1 RESTACKS

The AI Production Divide

Inside the AI Transformation of Parts, People & Processes
Hi, I’m Lily. I live in the world of distribution, where every vendor deck promises AI and almost every operator has a pilot running somewhere. And that’s a problem. A pilot proves nothing except that a demo behaves in a controlled corner of the business. What separates the leaders from everyone else is production, AI that runs the actual operation, eve…
InstaLILY AI29 LIKES
Dr Peter McCann Strain's avatar
Dr Peter McCann Strain
Your point about pilots living in a controlled corner is why so many programmes stall at promotion. I would make the move into the real operation depend on evidence: an agreed failure threshold, containment results, a tested rollback path, and a named owner for exceptions. Otherwise the pilot proves capability under friendly conditions and the launch silently changes the test. What had to be true before you let AI touch the live workflow?
Katie Hargreaves's avatar
Katie Hargreaves
Great perspective. The biggest gap isn't between companies experimenting with AI and those ignoring it—it's between those shipping AI into real production workflows and those stuck in endless pilots. That's where an experienced AI Development Company can make a real difference by focusing on integration, governance, and measurable business outcomes rather than just building impressive demos. Production-ready AI is ultimately more about execution than the model itself. Thanks for sharing these insights.



The AI Papers (#6)

What we're reading
Every month we take a look at five interesting new AI papers. Today, we look at whether AI sycophancy poses a risk to real-world relationships; how to evaluate an AI system that continues to learn after it is deployed; how AI covers the news; new insights on the terrorist group Boko Haram’s use of AI; and how AI performs at the tasks that employees most…
Conor Griffin18 LIKES5 RESTACKS