← Home
NLW analyzes Claude Opus 4.8 as a meaningful but modest upgrade, examining its improved judgment, reduced hallucinations, stronger self-correction, and willingness to push back on requests. The episode covers benchmark comparisons with GPT-5.5, Claude Code's new dynamic workflows, and the hypothesis that model harness matters as much as raw capability. Also discusses recent AI industry moves: Kirkland & Ellis betting on internal AI, OpenAI's GPT-5.5 Instant update, Cognition's $26B valuation, Meta's cloud strategy, and Microsoft's model roadmap.
This summary was generated from show notes and public descriptions, not from a full transcript review. Details may contain inaccuracies.
Curious
Highlights
Editorial
•
•
Misc
✧Claude Opus 4.8 framed as incremental but directionally important — not a revolution
✧Emphasis on behavioral traits (judgment, pushback, self-checking) over raw benchmark gains
✧Model harness (system prompts, routing logic, tool use) emerging as competitive differentiator alongside model weights
Was this useful?