i had to ask Astra to explain this one to me. OpenAI says its new internal model solved a maths problem people have worked on for 90 years. it used 10,000 AI agents for 88 hours. the model's still being trained. i filmed my reaction while trying to get my head around what that means. the result is wild. |
i’m Lennox. i sift through AI Twitter and share one thing i tried with ChatGPT. for people making products, content and useful systems. five emails a week, with a weekly option.
i've got 2 for the price of 1 for you: LAB-0005 and LAB-0006. the headlines for 0005: a big maths claim from OpenAI. a model price fight. Grok 4.7 on a real job. what agents might do to app checkout. and a local image model i want to test. how to use GPT-6 Astra to create brand design assets (spoiler alert: the local image model is wild. it does *anything*. without guardrails - scary stuff if put in the wrong hands) check out episode 0005 here: headlines for 0006: Opus 5.5 drops 1.5 hours...
i gave two AI setups the same four hours i built BuilderBench around work i actually use AI for: writing, editing video, designing pages, making motion graphics, using apps, researching decisions and fixing workflows. the first pilot compared Opus 5.5 with Claude and GPT-6 Sol with Codex. same frozen task pack. four hours of offline work, plus a live founder conversation. Extra High reasoning and Standard speed. i scored the anonymous work, then locked my ratings before revealing the names....
7 ai updates that changed how i build this week: we've entered the token scarcity era use a handoff.md for long-running agent tasks check out CUA (Computer Use Agent) if you use Claude agent portability is the new meta; don't get locked in to one app ... ... ... [this one's a fkn doozy!] i'm not gonna make it that easy for you ;) want the last 3 updates? you know what to do! watch LAB 0004 on YouTube speak soon x Lenny