Favicon of benchmark-qed-autoq

benchmark-qed-autoq Skill

AI Agent SkillPythonOpen source

Generate benchmark questions and assertions from input data using benchmark-qed. Use when: generating local, global, linked, or activity questions for RAG benchmarking, creating assertions for existing questions, computing assertion statistics, or running the autoq question generation pipeline. Also use when the user w Published by microsoft in benchmark-qed.

Decision snapshot

Is this a fit?

Best for

Data analysis, Includes SKILL.md, Reusable instructions

Works with

Compatibility not yet detected.

Access

API key required authentication, Credentials required

Setup

Copy skill directory

Project health

2 months ago · MIT license

Considerations

Access note: API key required.

What is benchmark-qed-autoq Skill?

Generate benchmark questions and assertions from input data using benchmark-qed. Use when: generating local, global, linked, or activity questions for RAG benchmarking, creating assertions for existing questions, computing assertion statistics, or running the autoq question generation pipeline. Also use when the user w Published by microsoft in benchmark-qed. This profile combines repository metadata with install, compatibility, and usage signals so developers can quickly decide whether it fits their agent workflow before opening the source repository.

Trust signal
95/100
Maintenance signal
90/100
Adoption signal
49/100

Automated repository signals based on public metadata such as recency, license, installation evidence, and adoption. These are not a security audit or endorsement. See how SkillIndex evaluates profiles.

Key capabilities

  • Includes SKILL.md support
  • Reusable instructions support
  • Data analysis
  • Data analysis use cases

Declared skill metadata

  • Source file: .apm/skills/benchmark-qed-autoq/SKILL.md

These fields retain source and confidence evidence from the indexed SKILL.md.

Compatibility and setup

Copy skill directory
  • Install or run with Copy skill directory
  • API key required

Requirements and access

API key required

Security and permissions

Review permissions before connecting any MCP server to an agent. Pay special attention to whether it can read local files, write data, call external services, or perform destructive actions.

Credentials requiredAPI key required authentication

When to use benchmark-qed-autoq Skill

  • Use it for data analysis.

Built with

PythonCopy skill directory

Editorial notes

Source

  • Creator: microsoft
  • Repository: microsoft/benchmark-qed
  • Skill file: .apm/skills/benchmark-qed-autoq/SKILL.md

What it does

Generate benchmark questions and assertions from input data using benchmark-qed. Use when: generating local, global, linked, or activity questions for RAG benchmarking, creating assertions for existing questions, computing assertion statistics, or running the autoq question generation pipeline. Also use when the user w

Skill instructions

Benchmark-QED Question Generation (autoq) Generate benchmark questions and assertions from input data for RAG evaluation. Prerequisites - A configured workspace with valid settings.yaml (use the /benchmark-qed-setup skill first) - A configured workspace with valid settings.yaml (use the benchmark-qed-setup skill to initialize and configure) - Input data (CSV or JSON) in the workspace input/ directory - Valid LLM API key in .env Run all commands with: bash uvx --from "git+https://github.com/microsoft/benchmark-qed" benchmark-qed <command Commands 1. Generate Questions (autoq) The main question generation pipeline. Generates benchmark questions from input data. bash uvx --from "git+https://github.com/microsoft/benchmark-qed" benchmark-qed autoq <settings.yaml <outputdir [OPTIONS] Options: | Option | Description | |--------|-------------| | --generation-types | Specific types to generate (repeatable). CLI default: all except datalinked, but this skill always includes datalinked | | --prin

Verified compatibility and discovery

Frequently asked questions

What is benchmark-qed-autoq?

benchmark-qed-autoq is a open-source AI agent skill with Copy skill directory. Generate benchmark questions and assertions from input data using benchmark-qed.

Who is benchmark-qed-autoq best for?

benchmark-qed-autoq is best for reusing agent instructions, scripts, and references, data analysis workflows.

How do I install benchmark-qed-autoq?

Install or run benchmark-qed-autoq using Copy skill directory. Check benchmark-qed-autoq for the latest setup command.

Is benchmark-qed-autoq actively maintained?

benchmark-qed-autoq may need a closer maintenance check before production use.

Share:

Stars
91
Forks
18
Last commit
2 months ago
Last verified
Aug 30, 2026
Metadata fetched
Aug 30, 2026
Repository age
1 year
License
MIT

Project health auto-fetched from the source repository.

Maintain this resource?

Review this source-backed profile, send a correction with evidence, or link to it from your documentation. Claims verify your relationship to the project; profile facts still require source evidence and editorial review.

Alternatives to benchmark-qed-autoq