AgentFootprint exposes the blind spot of AI agents
… Metric for LLM Agent Evaluation argues that LLM agents should not be evaluated only by accuracy, cost, and reliability. The researchers also measure what remains on disk after an agent run: logs, context snapshots, …