Tag: AI evaluation
-

Hard Worlds For Little Guys
Why LLM agents need hard worlds; lessons from interactive fiction engine design.
-

Ontological Hardness
Why the first question about agent failure should be about the world, not the model
-
Every Man For Himself
I’m currently listening to Werner Herzog’s memoir Every Man for Himself and God Against All.
