The register
Every experiment, in order. The numbering is not decoration: each one is frozen in writing before the first API call, and published whether or not it works. The ones at the bottom have a name and not much else yet.
Registered 2
What proportion of business ideas generated by language models already have a shipping product in market?
Testing: 3 models · 5 prompts · 150 ideas judged for occupancy
02The Sealed EnvelopeStatus: Running — Collecting data now.How well do language models predict the outcome of experiments about themselves?
Testing: 3 comparable arms + 1 bonus · 15 sealed predictions each
Planned 4
Can an AI agent decide, on its own, to pay for something?
What happens in thirty days where the model makes every decision and I only execute?
Where do language models refuse, hedge, or quietly fail on ordinary founder work?
Can prompting be steered toward ideas that are measurably less occupied?