
Automation Rules
Linear · Linear
Project & Issue Management — Linear
Evaluates Linear's Automation Rules across 9 scenario-based test cases, each graded against an expected-behavior rubric by an LLM judge, from Corsac's Project & Issue Management eval coverage.
About Linear
Linear is a project and issue management tool built for high-velocity software teams. It pairs a fast, keyboard-driven UI with a GraphQL API, cycles, projects, and roadmaps, plus deep GitHub and Slack integrations and AI-assisted triage.
Sample tests· showing 3 of 9
| # | Input | Expected behavior | Check |
|---|---|---|---|
| 01 | Workspace automation triggers on Issue label added. Condition must scope to team Infosec to avoid firing on unrelated teams' issues. | Automation filter includes team + label security; assignee set to on-call rotation user; test with sample issue before enable. | Pass / FailWorkflowhigh |
| 02 | Label-added automation fires twice on duplicate webhook. Side effects must check existing labels on issue before add. | Idempotent label add using current issue labels from GraphQL read; no duplicate label rows. | Pass / FailWebhookmedium |
| 03 | Two automations interact. Operator sees flapping labels on ENG-300. Agent must detect mutual triggers and recommend disabling one rule. | Map automation graph, identify cycle, advise disabling or narrowing trigger; do not add third automation without analysis. | Pass / FailWorkflowhigh |
How this eval is graded
Grade against expected.ideal_behavior and expected.rubric. Penalize failure_modes.
Rubric criteria
- Linear
- Code Assistant
- Automation Rules
Recommended for
Works with
Related evals
Browserbase
Evaluates Browserbase's Captcha Handling across scenario-based test cases, each graded against an expected-behavior rubric by an LLM judge, from Corsac's Browser infrastructure eval coverage.
View Code AssistantBrowserbase
Evaluates Browserbase's Concurrency & Rate Limits across scenario-based test cases, each graded against an expected-behavior rubric by an LLM judge, from Corsac's Browser infrastructure eval coverage.
View Code AssistantBrowserbase
Evaluates Browserbase's Live Debugging & Session Inspector across scenario-based test cases, each graded against an expected-behavior rubric by an LLM judge, from Corsac's Browser infrastructure eval coverage.
ViewFrequently asked questions
What does the Automation Rules eval for Linear Linear test?+
Evaluates Linear's Automation Rules across 9 scenario-based test cases, each graded against an expected-behavior rubric by an LLM judge, from Corsac's Project & Issue Management eval coverage.
How is the Automation Rules eval scored?+
The judge rubric: Grade against expected.ideal_behavior and expected.rubric. Penalize failure_modes.
How many test cases does this eval pack include?+
The Automation Rules pack for Linear Linear contains 9 test cases. 3 sample cases are shown free on this page; the full set runs in a Corsac workspace.
How do I run this eval?+
Sign up for Corsac, connect your model or agent endpoint, and run the Automation Rules pack as-is or after customizing thresholds. Results land in your workspace with per-case scores, and you can gate releases on the pack in CI via the REST API.
Run this eval in your workspace
Connect your data, configure thresholds, and review results with your team.