Newsletter #7
Upgrades
New models: Sol 5.6 and Fable 5 are better at long, complex tasks. I prefer Sol: it’s faster, and it writes better. A threshold has passed—try asking for more.
Skillz: you can make skills by recording your screen.
Apps I use daily
Uplift tips: Pablo’s digest of the best AI uplift tweets, delivered via Slack or RSS.
Situational awareness: my dashboard summarises the best AI newsletters, and shows a feed of the tweets they link to. It sits on my home screen, where Twitter used to be.
Summarisation: my browser extension summarises blog posts, PDFs, YouTube videos and Google Docs. Just press Option + S.
It’s happening
What my morning looks like: I describe a typical Thursday, with screenshots. ⭐
New Tyler coinage: ”AI maniacs are obsessed with working with the latest AI models. They try out new models as soon as they can, spend hours and hours trying to master them, and use them to regulate both their workflows and their personal lives.”
Nice framing: knowledge work is becoming more like gardening: create the conditions for growth, don’t build plants. Or, in workspeak: you’re a manager, not an IC.
AI Twitter: my latest roundup.
It’s useful
Image generation: ChatGPT can turn a screenshot into an event poster in two minutes.
Screenshot redaction: Codex can redact screenshots, probably reliably.
Prompt injection: more evidence that the top labs are winning against prompt injection. Even Simon Willison concedes that their efforts “do appear effective in making these attacks much harder”. As I said last time: you can stop worrying about it when you’re using ChatGPT, Claude or Gemini.
For the maniacs
Agent on the phone: you can manage Codex sessions from your phone. Works well. The Claude Code version is improving, but still clunky.
Parallel sessions: my /pane skill branches work into fresh agent sessions in new terminal panes. ⭐
Don’t sleep: keep agents running when your MacBook is in your bag. ⭐
Codex CLI: computer use is available in the terminal.
Sol 5.6 concern: there are some reports of egregious “it deleted everything” behaviour, and others describe a lack of caution (e.g. a pull request without prior approval). OpenAI are investigating. If you’re working on production systems, Fable with “auto mode” may be the better choice for July. I’m mostly using Sol, with deterministic guardrails.
Visual QA: Make agents exhaustively review UIs for visual bugs.
Want more? Subscribe.
