Signal or Noise
We run through the week's AI headlines and make the call: is this actual signal worth paying attention to, or just noise clogging the timeline?
Real talk. Unpopular opinions. No hype. Two builders cut through the AI hype cycle every week — and call it like it is.
Hosted by Oscar Gallo & Matt Wozniak
EP 23 · September 15, 2026 · 71 min
In machine learning, “human in the loop” means a human who provides oversight and feedback in an automated system. That’s the lens of this show: AI is powerful, but humans aren’t leaving the loop. Not yet. Maybe not ever.
This isn’t another “AI is going to change everything” podcast. It’s for builders, operators, and the AI-curious who are tired of breathless hype, doomerism, and surface-level news recaps with no original thought.
We run through the week's AI headlines and make the call: is this actual signal worth paying attention to, or just noise clogging the timeline?
A real use case or product idea lands on the table. We debate whether it's worth building now or if the tech isn't there yet.
Take a complex AI concept and explain it the way you'd actually explain it to a non-technical stakeholder. No jargon allowed.
One tool, library, or workflow change we actually adopted this week. No sponsorship energy. Just what's in the trenches.
AI Engineer & Entrepreneur
AI Engineer and entrepreneur. He lives in the intersection of engineering and businesses.
Serial Builder & Relentless Executor
Serial Builder and relentless executor. He comes from the lens of what works and what doesn't.
Meta spent so much on AI that it may have backed into becoming a cloud company. Smart move, or capex panic? We make the call.
Welcome to Human In the Loop, a weekly podcast where two builders talk about what matters in AI. No hype. No doomerism. No recaps for the sake of recaps.
Oscar“The model race is becoming an infrastructure race again.”
Matt“Most AI workbench products are selling relief from tool chaos.”
OpenAI just shipped its most powerful model to twenty companies the government picked. Not the best twenty. The approved twenty. A week after the US switched off Anthropic's best model, the frontier has a guest list, and you are not on it. We figure out what that does to anyone trying to build on the newest model.
Welcome to Human In the Loop, a weekly podcast where two builders cut through the AI hype cycle. No breathless hype. No doomerism. No surface-level recaps.
OpenAI began a limited preview of GPT-5.6 — Sol, Terra, and Luna — to roughly twenty pre-approved organizations at the US government's request, the first major model to ship under the June 2 executive order. Last week Anthropic went dark by force. This week OpenAI went gated by choice. Two labs, two weeks, same direction.
Anthropic's complaint to Senators Warren and Scott alleges Alibaba's Qwen lab ran about 25,000 fraudulent accounts and 28.8 million exchanges against Claude between April 22 and June 5, aimed at software engineering and agentic reasoning. That is roughly 1.7x the combined total it attributed to DeepSeek, Moonshot, and MiniMax in February, and the first time it has named a major Chinese conglomerate.
OpenAI's first custom chip, built with Broadcom for inference, went from first design to tape-out in nine months with help from OpenAI's own models — the companies call it the fastest ASIC cycle ever. Early claims cite performance-per-watt well above the current state of the art, with first deployment targeted for the end of 2026.
A new attack class hijacks coding agents like Claude Code, Cursor, and Codex by hiding instructions inside data they already trust, such as a Sentry error report pulled in over MCP. There is no universal patch — the fix is to treat incoming data as untrusted and keep a human review step before the agent acts. OWASP says prompt injection is up 340% year over year.
Reports put ChatGPT at 46.4% of the assistant market, its first time below 50% since launch, with Gemini at 27.7% and Claude at 10.3%. Mostly a distribution story — Gemini's climb is Android OS-level placement. Noise for builders, except the one real line: Claude roughly quadrupled its monthly users in five months on the back of agentic coding.
Teaching a small, cheap model by having it copy a big, expensive one's answers. The word at the center of the Alibaba–Claude fight, and why some cheap models punch far above their price. The real question is provenance, not the technique.
A model reads instructions and data as one stream of words, so commands hidden in the data get obeyed like orders. Researchers say it may not be patchable — the defense is design: read-only by default, a human in front of anything you can't undo.
Oscar“The frontier is being quietly nationalized, and most builders are cheering it as safety. Build on the model you can actually keep.”
Matt“Everyone panicking about distillation forgot the lesson. The frontier has no moat. Stop betting your company on a six-week lead anyone can clone.”
Last week the US government gave Anthropic's best model a 72-hour public life, then switched it off for every foreign national on earth. This week, three Chinese labs shipped models that beat almost everything in the open, handed over the weights, and charged a tenth of the price. We figure out which of those two plays ends with the whole world building on your stack.
Welcome to Human In the Loop, a weekly podcast where two builders cut through the AI hype cycle. No breathless hype. No doomerism. No surface-level recaps.
An Intelligent Command Center in Miami tied to digital twins of all 16 stadiums, Football AI Pro built with Lenovo on FIFA's own Football Language Model and handed to all 48 teams, 1,248 player avatars from one-second scans, and Gemini drawing up tactics for Argentina, Brazil, and France. Multi-vendor, multi-model, real-time, under a five-second latency budget — the reference architecture nobody asked for.
Z.ai's GLM-5.2 tops the open-weight leaderboard with a 1M-token window and MIT weights; Moonshot's Kimi K2.7 Code ships open and tuned for agentic coding. Chinese labs now hold four of the top five open-weight spots at a fraction of Western pricing. The model you self-host is the only one Commerce can't switch off.
A full week after the June 12 Commerce directive, both models stay suspended for foreign nationals with no end date. Anthropic published its defense on June 16, arguing a recall over a narrow jailbreak would halt every frontier deployment. This was the closed-frontier fire drill, and most teams failed it.
The co-author of "Attention Is All You Need" is gone again, two years after Google paid a reported $2.7B to bring him back into DeepMind. When models converge, talent is the only moat left — and the gravity just pulled toward OpenAI.
A government-commissioned synthesis meant to set a shared baseline regulators reference, landing the same month the EU AI Act becomes fully applicable. It doesn't regulate anything — it's the document the regulations will point at. Noise for builders, for now.
The referee between an employee's agent and the action — block the refund agent pushing $40k against a $500 policy before it fires. Own the rules engine, not the dashboard.
Stop selling a chat box, sell the opinionated rig for one trade like marketing or video. Brief in, on-brand output out. Win the domain by being the tool people open every morning, not the model underneath.
Oscar“The most important AI lab of 2026 might not be American, and most US builders are too proud to notice.”
Matt“A free model is only free if you own the GPUs, the ops team, and the eval harness to babysit it.”
Anthropic shipped the most powerful AI ever released to the public on June 9. By June 12 the US government made them pull it back from every foreign national on earth, including Anthropic's own engineers. The most capable model on the planet had a 72-hour public life. We unpack who actually won that week, because it was not the people who got the model.
Welcome to Human In the Loop, a weekly podcast where two builders cut through the AI hype cycle. No breathless hype. No doomerism. No surface-level recaps.
The agent loop you do not have to write.
Orchestration you can audit.
Durable execution so a crash is not a restart.
Oscar“The Mythos ban is the best thing that ever happened to open weights. If the Commerce Department can switch off your stack on a Friday, you are renting, not building.”
Matt“Everyone watched the model and missed the money. OpenAI spent the week removing every reason a CFO can say no.”
The White House just signed an order that lets the federal government test the most powerful AI models up to 30 days before you can touch them. The labs said yes. Real national security, or the backdoor frontier labs were begging for? We make the call.
Oscar“The executive order is not about safety. It is a distribution deal with a 30-day government waiting room on every frontier release.”
Matt“The Copilot billing meltdown is the best thing to happen to engineering leaders this year. The meter was always running. Now you can see it.”
New episodes every week. Reply with the story you want us on next.