Skip to content

Latest commit

 

History

History
8 lines (6 loc) · 199 Bytes

File metadata and controls

8 lines (6 loc) · 199 Bytes

agent-eval-framework

LLM Agent Evaluation Platform

Goal

  • Build platform to track efficacy of AI models
  • Frontend to display and analyze results
  • Highlight gaps in the design model's design