
o1 Pro Mode – Full Analysis (plus o1 paper highlights)
Oh boy. o1 pro mode out on the same night as o1 full. I read the 49 page paper, ran my own tests, spent my fuel allowance on Pro Mode and will give you all the highlights. Suffice to say the story is ...
5 Dec 202416min

AI Breaks Its Silence: OpenAI’s ‘Next 12 Days’, Genie 2, and a Word of Caution
Calmest before the storm? Whatever analogy you want to use things had gotten quiet toward the end of 2024. But then tonight we got Genie 2, and a series of scheduled announcements from OpenAI. Sora is...
5 Dec 202415min

New Google Model Ranked ‘No. 1 LLM’, But There’s a Problem
A new and mysterious Gemini model appears at the top of the leaderboard, but is that the full story? I dig behind the headline to show you some anti-climactic results, give some context with leaks in ...
15 Nov 202415min

Leak: ‘GPT-5 exhibits diminishing returns’, Sam Altman: ‘lol’
The last few days have seen two narratives emerge. One, derived from yesterday’s OpenAI leak in TheInformation, that GPT-5/Orion is a disappointment, and less of a leap than GPT-3 to GPT-4. The second...
10 Nov 202415min

ChatGPT with Search, Altman Answers Anything and Simple Bench Out
The Google destroyer, the Perplexity crusher? Or just hype? ChatGPT with Search is here, and simultaneously Altman and co did an AMA on Reddit, covering GPT-5, Sora, SearchGPT and a lot more. Plus, th...
1 Nov 202415min

The New Claude 3.5 Sonnet: Better, Yes, But Not Just in the Way You Might Think
A new state of the art LLM (at least for creative writing and basic reasoning) but what lies behind the numbers that were put out? Is it for real, and are AI agents about to grab your mouse and shake ...
28 Okt 202422min



















