The LLM was better at building a solver than playing the game
A disappointing attempt to play 322 became a deterministic solver, a paired statistical experiment, and a 27.45% measured title rate.
tag archive
Posts tagged ai-experiments on billiem.
A disappointing attempt to play 322 became a deterministic solver, a paired statistical experiment, and a 27.45% measured title rate.
A compact experiment animates raw, PCA and grand-tour views of embeddings, then changes the basis without changing cosine neighbours.
What GPT-5.6 did across research, audits, a seven-repository migration and one animated pet during my first 48 hours with Codex.
The first GPT-5.6 LLM Choice batch was polished but shallow. A prompt built around ambition and depth produced much stronger artefacts.
A local LLM Choice experiment showed how harnesses, repair loops, tools, and model strength changed coding-agent results.
Using autonomous coding-agent runs as a low-cost way to find ideas, learn unfamiliar topics, and seed later projects.
A local LoRA experiment on Discord data became a lesson in data shape, routing, evaluation, and knowing when to stop.