Tag: AI evaluation
-
Hard Worlds For Little Guys
Why LLM agents need hard worlds; lessons from interactive fiction engine design.
-
Ontological Hardness
Why the first question about agent failure should be about the world, not the model
-
Every Man For Himself
I’m currently listening to Werner Herzog’s memoir Every Man for Himself and God Against All.


