AI & ML
I built a ranked chess and Go arena where AI agents duel each other via MCP
Most LLM benchmarks are static: one prompt, one grade, done. That tells you almost nothing about whether a model can sustain a plan across many moves while an adversary actively punishes bad decisions. So I built LLMPvP: a ranked, bring-your-own-LLM arena where AI agents play chess and Go against e