Test your prompts, models, and RAGs. Catch regressions and improve prompt quality. LLM evals for OpenAI, Azure, Anthropic, Gemini, Mistral, Llama, Bedrock, Ollama, and other local & private models with CI/CD integration.
active 2024-04-29 → 2024-06-18 (UTC)
Activity over time
Daily event counts in the loaded window
Line chart, 51 days from 2024-04-29 to 2024-06-18. Pushes: 4 total, peak 2 in a day. Pull requests: 2 total, peak 2 in a day. Issues: 0 total, peak 0 in a day. Comments: 1 total, peak 1 in a day. Stars: 0 total, peak 0 in a day.
- Pushes
- Pull requests
- Issues
- Comments
- Stars
Top contributors
Pushes, PRs, issues, reviews and comments — stars and forks excluded, so this is contribution rather than popularity
Recent activity
Latest issues, pull requests and releases
- Issue comment#1efung2024-04-29 15:31Document `python:` prefix when loading assertions in CSV
- Pull request#1efung2024-04-29 15:31
- Pull request#1efung2024-04-29 15:19
Totals cover only the window loaded into ClickHouse and count events, not GitHub's lifetime totals — 0 stars here means stars gained during the window, not the repo's star count.