Skip to content
#

ai-agent-evaluation

Here are 15 public repositories matching this topic...

applied-ai-field-guide

Frontend for Calibrate, a framework for evaluating AI agents: speech-to-text, text-to-speech, LLM evaluation, end-to-end simulations

  • Updated Sep 16, 2026
  • TypeScript

Agentic Workflow Evaluation: Text Summarization Agent. This project includes an AI agent evaluation workflow using a text summarization model with OpenAI API and Transformers library. It follows an iterative approach: generate summaries, analyze metrics, adjust parameters, and retest to refine AI agents for accuracy, readability, and performance.

  • Updated Feb 23, 2025
  • Python

Add this topic to your repo

To associate your repository with the ai-agent-evaluation topic, visit your repo's landing page and select "manage topics."

Learn more