
LLM Let's Play - Fire Emblem
Agent gameplay harness: mGBA state in, tactical context to the model, controller actions back, dashboard for humans.

(Click to skip) →

Agent gameplay harness: mGBA state in, tactical context to the model, controller actions back, dashboard for humans.


TEE-backed verifiable randomness infrastructure for giveaways and gachapon systems without on-chain costs.
A 600-game benchmark for measuring LLM deception, lie detection, and instruction compliance through the bluffing card game Bullshit.
“Honesty prompts reduced lying, but also reduced challenges and made remaining lies more successful.”
A quick writeup on Lisper, my Gemma 4 Good Hackathon submission for low-anxiety lisp speech practice.
A 600-game study of six language models playing Bullshit found that honesty prompts reduce deception, but also soften enforcement around the lies that remain.
Chainlink VRF is overkill for gacha games. We built a $0.01 alternative using Intel TDX.