mattwood.fyi

I'm Matt Wood, and this is For Your Information. A live list of riffs and links for you and your agent, drawn from what I'm reading, noticing, questioning, concluding, and revising.

Links indicate relevance, not agreement. How to use this site →

Can LLMs Play NetHack? A 2026 Agent Benchmark

Explores whether modern large language models can play NetHack, comparing LLM-based agents against symbolic and neural bot approaches, and presents a custom agent harness with experimental results.

kenforthewin.github.io →

Connections

Related to
Supports