← All episodesEP 22

Ep 22: The post-model world: harnesses, provenance, and Quasar

September 8, 2026 · 80 min

EP 22September 8, 2026 · 80 minShip It or Skip It

Today's models can already read, write, code, browse, reason over documents, and use software. For most bounded digital tasks, the missing piece is no longer raw intelligence. It is the system around the model.

Oscar makes the case for the post-model world. A task-specific harness supplies the right context, divides the work, constrains tools, preserves state, retries failures, verifies the result, and asks a person when judgment matters. Model access is available to every competitor. The private traces, corrections, and evaluation cases generated by real work are not.

Matt makes the case for provenance. Multiverse Computing launched Quasar 438B as the highest-scoring European model on Artificial Analysis Intelligence Index. A day later, its own publication said Quasar is compressed and tuned from Z.ai's GLM-5.2. Useful compression and European deployment do not make the underlying lineage disappear.

Together, the opinions create one practical rule. Treat the model as replaceable when you design the product. Demand a clear record of the model and version when you govern it.

Watch on YouTubeListen on SpotifyListen on Apple
Signal or Noise

The week’s AI headlines, filtered.

  1. SIGNAL with a caveatQuasar 438B and GLM-5.2

    The headline says European leadership. The documentation says a compressed and tuned GLM-5.2 base. What does sovereignty need to mean for a buyer?

  2. SIGNALGPT-6 Astra

    OpenAI's first broadly deployed model at its Critical cyber threshold. Stronger cyber capability meets lower monitorability.

  3. SIGNALClaude Fable 5.1

    Anthropic cut cache-read pricing. That may alter the cost of agents that repeatedly load code and policy context.

  4. SIGNALGemini 3.8 Flash Cyber

    Google brings vulnerability discovery and automated patching into a faster tier for trusted defenders.

  5. SIGNAL pending independent testingMuse Spark 1.3

    Meta says its agent tracks long tasks better. The claim needs an independent test on work that matters to you.

Ship It or Skip It

Two product ideas for the agent economy.

The AI expense layer

Policy is checked before company money moves.

An agent kill switch

Find unapproved agents and stop unauthorized actions at runtime.

Unpopular opinions

Oscar and Matt make the case.

Oscar

Models do not matter anymore as a durable product moat. We already have enough intelligence for most bounded digital tasks. A good harness built for one specific task will beat a smarter general agent.

Matt

Model releases matter less than model provenance. Calling a compressed GLM-5.2 model Europe's leading model without foregrounding its lineage is sovereignty theater.

Sources referenced