iFixAi is an independent auditing toolkit for evaluating whether an AI agent performs the business job it was assigned. Instead of measuring only latency, token use, or prompt-injection resistance, it examines operational behavior, organizational alignment, and task outcomes. A standard run performs 32 inspections across five evaluation pillars and produces an A-to-F grade. Users can configure and launch audits through a guided CLI, explicit command flags, or an agent plugin and skill. The same engine can test hosted providers or a real agent endpoint and can use self-grading, an independent judge, or a multi-judge ensemble. Results are delivered as JSON, Markdown, and terminal scorecards for review or automation. Saved configuration supports repeatable local runs, onboarding, and CI workflows.
Features
- Thirty-two AI agent inspections
- Five-pillar operational evaluation framework
- A-to-F grading and detailed scorecards
- Guided, scripted, and agent-operated audits
- Independent and multi-judge evaluation modes
- JSON, Markdown, and terminal reports