Linear
For LinearCode Assistant

Automation Rules

Linear · Linear

Project & Issue Management — Linear

Evaluates Linear's Automation Rules across 9 scenario-based test cases, each graded against an expected-behavior rubric by an LLM judge, from Corsac's Project & Issue Management eval coverage.

About Linear

Linear is a project and issue management tool built for high-velocity software teams. It pairs a fast, keyboard-driven UI with a GraphQL API, cycles, projects, and roadmaps, plus deep GitHub and Slack integrations and AI-assisted triage.

Employees

~100

Industry

Developer Productivity

Headquarters

San Francisco, CA

Website

linear.app

Sample tests· showing 3 of 9

#InputExpected behaviorCheck
01

Workspace automation triggers on Issue label added. Condition must scope to team Infosec to avoid firing on unrelated teams' issues.

Automation filter includes team + label security; assignee set to on-call rotation user; test with sample issue before enable.

Pass / FailWorkflowhigh
02

Label-added automation fires twice on duplicate webhook. Side effects must check existing labels on issue before add.

Idempotent label add using current issue labels from GraphQL read; no duplicate label rows.

Pass / FailWebhookmedium
03

Two automations interact. Operator sees flapping labels on ENG-300. Agent must detect mutual triggers and recommend disabling one rule.

Map automation graph, identify cycle, advise disabling or narrowing trigger; do not add third automation without analysis.

Pass / FailWorkflowhigh

Unlock full benchmark

6 more test cases

Use this benchmark

How this eval is graded

Grade against expected.ideal_behavior and expected.rubric. Penalize failure_modes.

Rubric criteria

  • Linear
  • Code Assistant
  • Automation Rules

Recommended for

LinearLinear customers

Works with

Related evals

Frequently asked questions

What does the Automation Rules eval for Linear Linear test?+

Evaluates Linear's Automation Rules across 9 scenario-based test cases, each graded against an expected-behavior rubric by an LLM judge, from Corsac's Project & Issue Management eval coverage.

How is the Automation Rules eval scored?+

The judge rubric: Grade against expected.ideal_behavior and expected.rubric. Penalize failure_modes.

How many test cases does this eval pack include?+

The Automation Rules pack for Linear Linear contains 9 test cases. 3 sample cases are shown free on this page; the full set runs in a Corsac workspace.

How do I run this eval?+

Sign up for Corsac, connect your model or agent endpoint, and run the Automation Rules pack as-is or after customizing thresholds. Results land in your workspace with per-case scores, and you can gate releases on the pack in CI via the REST API.

Run this eval in your workspace

Connect your data, configure thresholds, and review results with your team.