Claude is Cool, GPT Likes to Gossip

Can AI write fiction like a human? We explore StoryScope: Investigating idiosyncrasies in AI fiction (Russell, et al., 2026) to find out.

https://doi.org/10.48550/arXiv.2604.03136

Season 1 · Episode 18 · July 28, 2026

The Goldblum Effect

Do Anthropic’s models contain a global workspace? If so, what are the implications? We explore Verbalizable Representations Form a Global Workspace in Language Models (Gurnee, et al., 2026) to find out.

Verbalizable Representations Form a Global Workspace in Language Models

Season 1 · Episode 17 · July 10, 2026

Speculative Intelligence

What are the pathways from here to AGI and ASI? Can anyone agree on a definition for AGI? We explore From AGI to ASI (Genewein, et al., 2026) to find out.

https://doi.org/10.48550/arXiv.2606.12683

Season 1 · Episode 16 · June 28, 2026

Disentangling Model From Harness

How can a self-evolving harness benefit a model’s capabilities? We explore Harness Updating Is Not Harness Benefit: Disentangling Evolution Capabilities in Self-Evolving LLM Agents (Lin, et al., 2026) to find out.

https://doi.org/10.48550/arXiv.2605.30621

Season 1 · Episode 15 · May 28, 2026

Superhuman Step Change

Is Mythos all its hyped up to be? We explore System Card: Claude Mythos Preview (Anthropic, 2026) to find out.

System Card: Claude Mythos Preview

Season 1 · Episode 14 · May 28, 2026

Opening the Black Box

We’re joined by Gareth O’Shea, to ask:

Could it be possible to run an MRI on an LLM while it’s thinking? We explore Qwen-Scope: Turning Sparse Features into Development Tools for Large Language Models (Qwen Team, 2026) to find out.

https://doi.org/10.48550/arXiv.2605.11887

Season 1 · Episode 13 · May 21, 2026

Talmudic Post-Training

Should we align models based on human preference or human behaviour? We explore Alignment Makes Language Models Normative, Not Descriptive (Shapira, et al., 2026) to find out.

https://doi.org/10.48550/arXiv.2603.17218

Season 1 · Episode 12 · March 26, 2026

Specialisation and Art

How does extra domain specific pre-training contribute to code specific models? We explore Qwen3-Coder-Next Technical Report (Qwen Team, 2026) to find out.

https://doi.org/10.48550/arXiv.2603.00729

Season 1 · Episode 11 · March 17, 2026

Modal, Spatial, and Temporal Models

Where do inconsistencies show up within the current state of world models? We explore The Trinity of Consistency as a Defining Principle for General World Models (Wei, et al., 2026) to find out.

https://doi.org/10.48550/arXiv.2602.23152

Season 1 · Episode 10 · March 1, 2026

Reddit for AI Agents

Is RSI through multi-agent systems doomed to degenerate? We explore The Devil Behind Moltbook: Anthropic Safety is Always Vanishing in Self-Evolving AI Societies (Wang, et al., 2026) to find out.

https://doi.org/10.48550/arXiv.2602.09877

Season 1 · Episode 9 · February 24, 2026