We Tested GPT 5.6 Sol Early

July 9
1h 11m

View Transcript

Episode Description

We've spent six figures in tokens testing OpenAI's 5.6 Sol model to see whether if its better than Fable and what OpenAI have done to make it even better than 5.5. Also, we breakdown why we both moved our agents to Linux boxes, how to actually burn $65k on a single loop, and the Codex vs Claude Code subagent gap that's now bigger than the model gap itself.


Thanks to this episode's sponsors Clerk and General Translation:


Sources Available on our Substack:

See all episodes