Evaluate is a library that makes evaluating and comparing models and reporting their performance easier and more standardized.
Features
- Implementations of dozens of popular metrics
- Examples available
- Documentation available
- Comparisons and measurements
- Metric cards
- Type checking
Categories
Machine LearningLicense
Apache License V2.0Follow Evaluate
Other Useful Business Software
Build Agents and Models on One Platform
Gemini Enterprise Agent Platform is Google Cloud's comprehensive platform for developers to build, scale, govern, and optimize agents and models. Choose from Google's most advanced models and third-party models like Anthropic's Claude Model Family.
Rate This Project
Login To Rate This Project
User Reviews
Be the first to post a review of Evaluate!