EsoLang-Bench: Evaluating Genuine Reasoning in LLMs via Esoteric Languages

(esolang-bench.vercel.app)

42 points | by matt_d 2 hours ago ago

10 comments