agent-eval-framework LLM Agent Evaluation Platform Goal Build platform to track efficacy of AI models Frontend to display and analyze results Highlight gaps in the design model's design