Benchmark deep research agents across factual, quality, and process dimensions with MiroEval
Score deep research agents on benchmark tasks using factual verification, report-quality scoring, and process evaluation before model or workflow changes ship.
Benchmark deep research agents across factual, quality, and process dimensions with MiroEval
Score deep research agents on benchmark tasks using factual verification, report-quality scoring, and process evaluation before model or workflow changes ship.
Installation
Method 1, Agent Skill Exchange
- Install from the marketplace listing: https://agentskillexchange.com/skills/benchmark-deep-research-agents-across-factual-quality-and-process-dimensions-with-miroeval/
Method 2, Git clone
git clone https://github.com/agentskillexchange/skills.git && cd skills/skills/benchmark-deep-research-agents-across-factual-quality-and-process-dimensions-with-miroeval
Method 3, Download ZIP
- Download the repository ZIP and extract
skills/benchmark-deep-research-agents-across-factual-quality-and-process-dimensions-with-miroeval.
Method 4, Manual copy
- Copy this skill folder into your local skills directory, then reload your agent tooling.
Method 5, Fork and sync
- Fork the repository if you want to maintain local edits while syncing upstream changes.