Serious coding with LLMs. Lab notes 2026-10-01: Codex and Claude generate a shared analysis
The third ‘lab note’. More on the power and Achilles’ heel of “brute force trial and error with token pattern statistics”
Aggregated enterprise architecture wisdom
The third ‘lab note’. More on the power and Achilles’ heel of “brute force trial and error with token pattern statistics”
Coding with LLMs (Claude Code, OpenAI Codex) is often presented as the ‘killer app’ for Generative AI. But looking at data, it seems the one piece of the puzzle missing is actual cost. A quest into getting a less muddy picture about what is going on, w…
We introduce a new term: “AI-slopcast”. This is a podcast that is created by Generative AI and — surprise! — is AI-slop. The victim: one of my own posts.
It turns out that AI has created a whole new language. Humans do not speak it, and they may even mistake it for talk about sex. But luckily Generative AI is able to translate it to something humans can understand (and where the sex doesn’t show up).
‘Reasoning models’ such as GPT4-o3 have become a well known member of the Generative AI family. But look inside and while they add a certain depth, at the same time they add nothing at all. Not ‘reasoning’ anyway. Just another ‘level of indirection’ wh…
GPT-3o has done very well on the ARC-AGI-PUB benchmark. Sam Altman has also claimed OpenAI is confident that it can build Artificial General Intelligent (AGI). But that may be based on confusions around ‘learning’. On the difference between narrow, ge…
Opinion piece / Crossed view – MEGA International – June…
One of the use cases I thought was reasonable to expect from ChatGPT and Friends (LLMs) was summarising. It turns out I was wrong. What ChatGPT isn’t summarising at all, it only looks like it. What it does is something else and that something else only…
Microsoft researchers published a very informative paper on their pretty smart way to let GenAI do ‘bad’ things (i.e. ‘jailbreaking’). They actually set two aspects of the fundamental operation of these models against each other.
Thanks to Gary Marcus, I found out about this research paper. And boy, is this is both a clear illustration of a fundamental flaw at the heart of Generative AI, as well as uncovering a doubly problematic and potentially unsolvable problem: fine-tuning …
Sam Altman wants $7 trillion for AI chip manufacturing. Some call it an audacious ‘moonshot’. Grady Booch has remarked that such scaling requirements show that your architecture is wrong. Can we already say something about how large we have to scale cu…
ChatGPT has acquired the functionality of recognising an arithmetic question and reacting to it with on-the-fly creating python code, executing it, and using it to generate the response. Gemini’s contains an interesting trick Google plays to improve be…