No cloud. No API keys. Just you, your hardware and the biggest model you can make it run.
A one-day hack for people who run AI locally. Bring a laptop, a mini PC or a GPU rig, and race to squeeze the most out of open-weight models. A live leaderboard runs all afternoon.
Three races
-
Biggest model: the largest model you can run at 5+ tokens a second on hardware you carried in
-
Fastest tokens: highest tokens per second on a fixed model and prompt set we hand out on the day
-
Most useful offline app: build something genuinely useful that works with the Wi-Fi switched off
The day
-
12:00pm: doors, pizza, rules and the leaderboard goes live
-
12:30pm: hacking (quantisation, runtimes, tricks: anything goes as long as it runs on the box in front of you)
-
5:00pm: final runs, verified live
-
5:30pm: winners and drinks
Prizes for each race, and the top prize is a local-AI mini PC to keep.
Beginners welcome: if you've never run a model locally, we'll get you going in ten minutes with Ollama. Hosted by Encode Club at Encode Hub, Shoreditch.