What is Scope?
Understand what Scope measures and when to reach for it. Read the introduction.
Run one scenario across GitHub Copilot, Claude Code, and VS Code. Score every run against criteria you define, and see how your agentic experience holds up across all of them — at scale.
Scope submits your task to each agent, scores every run against the criteria you define, and lines the results up side by side.
Hand one coding task to GitHub Copilot, Claude Code, or VS Code with the Copilot driver extension — with the profile you want.
Each run is evaluated against your criteria DAG. Gate child criteria on their parents, or leave them as independent roots.
See not just whether it worked, but how each agent got there — the tools it reached for and where it got stuck.
Scope drives existing agents against tasks you control, then captures what they asked for, what tools they reached for, and where they got stuck.
Hand a task to GitHub Copilot, Claude Code, or VS Code with the Copilot driver extension.
Bundle worker, model, version, MCP servers, skills, and extensions to re-run a setup consistently.
Describe what “good” means. Gate child criteria on their parents, or leave them as independent roots.
Stream logs straight from the worker as each run executes, attempt by attempt.
Scope detects characteristics on your prompts so you can compare across heterogeneous tasks.
Submit runs, manage profiles and criteria, and fetch results through the REST API or the CLI.
Submit requests, configure profiles and criteria, and watch runs stream — all from the browser.
Submit from the Portal →Wire Scope into your pipelines and tooling with a single authenticated request.
Read the API guide →What is Scope?
Understand what Scope measures and when to reach for it. Read the introduction.
Your first run
Run a task end to end from the Portal in a few minutes. Get started.
Submit via REST API
Automate request submission and integrate Scope into your tools. API guide.
Reference
Endpoint reference, configuration schemas, and worker capabilities. Browse the reference.
Spin up your first benchmark in minutes — from the Portal, the API, or the CLI.