Score deep research agents on benchmark tasks using factual verification, report-quality scoring, and process evaluation before model or workflow changes ship.
Scanned 6/8/2026
Install to Claude Code
npx -y skills add agentskillexchange/skills --skill benchmark-deep-research-agents-across-factual-quality-and-process-dimensions-with-miroeval --agent claude-codeInstalls into .claude/skills of the current project.
Are you the author of Benchmark Deep Research Agents Across Factual Quality And Process Dimensions With Miroeval?
Add the live security badge to your README — it updates automatically with every re-scan.
[](https://www.skillsdirectory.com/skills/agentskillexchange-benchmark-deep-research-agents-across-factual-qual)More formats (shields.io, HTML) on the badges page.
---
name: "Benchmark deep research agents across factual, quality, and process dimensions with MiroEval"
slug: "benchmark-deep-research-agents-across-factual-quality-and-process-dimensions-with-miroeval"
description: "Score deep research agents on benchmark tasks using factual verification, report-quality scoring, and process evaluation before model or workflow changes ship."
github_stars: 34
verification: "listed"
source: "https://github.com/MiroMindAI/MiroEval"
author: "MiroMindAI"
publisher_type: "organization"
category: "Code Quality & Review"
framework: "Multi-Framework"
tool_ecosystem:
github_repo: "MiroMindAI/MiroEval"
github_stars: 34
---
# Benchmark deep research agents across factual, quality, and process dimensions with MiroEval
Score deep research agents on benchmark tasks using factual verification, report-quality scoring, and process evaluation before model or workflow changes ship.
## Prerequisites
Python, uv, model result JSON, required API keys for judge and retrieval services
## Installation
No source-backed install or usage instructions could be extracted automatically. Review the upstream project before running this skill in a sensitive workflow.
- Source: https://github.com/MiroMindAI/MiroEval
## Documentation
- https://github.com/MiroMindAI/MiroEval
## Source
- [Agent Skill Exchange](https://agentskillexchange.com/skills/benchmark-deep-research-agents-across-factual-quality-and-process-dimensions-with-miroeval/)
No comments yet. Be the first to comment!