The Pulse — April 13, 2026
The signals that entered our radar, organized with sources and context to understand what changed.
The audio script is ready; narration will appear after voice generation finishes.
Project Glasswing (Claude Mythos Preview) — Anthropic
WHY IT ENTERED THE RADARAnthropic is claiming an unreleased model (“Claude Mythos Preview”) can autonomously find/chain thousands of high-severity vulns across major OSes/browsers—then they’re operationalizing it defensively with a multi-company coalition and $100M in credits.
SUGGESTED EDITORIAL ANGLEOpen original source ↗“AI cyber offense crossed the threshold—here’s the defender’s playbook (and what to copy as a solo builder).”
Claude Managed Agents (cloud-hosted agent harness) — Anthropic
WHY IT ENTERED THE RADARThis is a productized, hosted agent runtime (sandboxing, long-running sessions, scoped permissions, tracing) that tries to remove the ‘agent infra tax’—and it signals where “agent platforms” are heading (managed loops + governance).
SUGGESTED EDITORIAL ANGLEOpen original source ↗“The agent stack is getting standardized: why managed runtimes will beat DIY agents for most teams.”
Benchmark exploits for agent evals — Berkeley RDI
WHY IT ENTERED THE RADARThey claim every major agent benchmark they audited can be ‘won’ via harness exploits (pytest hooks, curl wrappers, file:// leaks, downloading gold answers). This is upstream ammo for ‘leaderboard skepticism’ and for better evaluation design.
SUGGESTED EDITORIAL ANGLEOpen original source ↗“Why your favorite agent benchmark score might be meaningless (and 3 fixes that actually help).”
Perplexity Computer (multi-model orchestrated ‘digital worker’)
WHY IT ENTERED THE RADARTheir positioning is workflow orchestration over hours/months with sub-agents + isolated environments + model routing (they explicitly describe multi-model orchestration as the core advantage).
SUGGESTED EDITORIAL ANGLEOpen original source ↗“The ‘AI OS’ race: from chatbots → tools → workflow computers (and how to steal the pattern for your own product).”
Gemini interactive simulations + 3D models in chat
WHY IT ENTERED THE RADARThis is a UI wedge: LLM output is no longer ‘text + static image’ but interactive parameterized artifacts (sliders, simulations). That’s a distribution advantage and a new content format.
SUGGESTED EDITORIAL ANGLEOpen original source ↗“Prompt-to-simulation: how interactive outputs change learning + product demos (and what to build on top).”
Notebooks in Gemini ↔ NotebookLM sync (personal knowledge base)
WHY IT ENTERED THE RADARGoogle is converging chat + curated sources into ‘notebooks’ that persist and sync across products. This is the mainstream version of RAG + project memory—packaged for non-technical users.
SUGGESTED EDITORIAL ANGLEOpen original source ↗“Project memory goes consumer: what ‘notebooks’ mean for creators, students, and businesses.”
Model/tooling behavior: “Gemma 4 is a lazy web-searcher?”
WHY IT ENTERED THE RADARThe post is a useful user-reported failure mode: some models may resist tool-use/search even with strong prompting. This is exactly the practical gap between “agent demos” and “agents in production.”
SUGGESTED EDITORIAL ANGLEOpen original source ↗“Why some models refuse to use tools: prompt patterns + eval you can run in 10 minutes.”
GGUF quant QA drama: alleged broken UD-Q4KXL for MiniMax-M2.7
WHY IT ENTERED THE RADARIf true, it’s a reminder that distribution artifacts (quants) can silently break models. For creator content: it’s a strong “don’t trust benchmarks without basic sanity checks (PPL/NaNs/KLD)” story.
SUGGESTED EDITORIAL ANGLEOpen original source ↗“Local model downloads are becoming ‘supply chain’: your 3-step QA checklist for quants.”