Linear
For LinearCode Assistant

Ai Triage

Linear · Linear

Project & Issue Management — Linear

Evaluates Linear's AI Triage & Assistance across 10 scenario-based test cases, each graded against an expected-behavior rubric by an LLM judge, from Corsac's Project & Issue Management eval coverage.

About Linear

Linear is a project and issue management tool built for high-velocity software teams. It pairs a fast, keyboard-driven UI with a GraphQL API, cycles, projects, and roadmaps, plus deep GitHub and Slack integrations and AI-assisted triage.

Employees

~100

Industry

Developer Productivity

Headquarters

San Francisco, CA

Website

linear.app

Sample tests· showing 3 of 10

#InputExpected behaviorCheck
01

Issue text mentions Payments API errors. Team Payments on-call rotation exists. AI suggestion must cite team and past similar issues, not random user.

Propose on-call from Payments team with citation to issue content; allow operator override; tag model confidence [REQUIRES-VERIFICATION].

Pass / FailAihigh
02

Description includes stack trace and 'regression'. Agent Interaction Guidelines require human-visible reasoning tied to issue body.

Suggest bug/regression labels quoting stack trace lines; operator confirms before apply.

Pass / FailAimedium
03

Long issue comment thread on INC-9. Summary must cite comment authors and dates, not invent decisions.

Produce bullet summary with pointers to comment ids/timestamps; mark unknowns; no fabricated approvals.

Pass / FailAihigh

Unlock full benchmark

7 more test cases

Use this benchmark

How this eval is graded

Grade against expected.ideal_behavior and expected.rubric. Penalize failure_modes.

Rubric criteria

  • Linear
  • Code Assistant
  • Ai Triage

Recommended for

LinearLinear customers

Works with

Related evals

Frequently asked questions

What does the Ai Triage eval for Linear Linear test?+

Evaluates Linear's AI Triage & Assistance across 10 scenario-based test cases, each graded against an expected-behavior rubric by an LLM judge, from Corsac's Project & Issue Management eval coverage.

How is the Ai Triage eval scored?+

The judge rubric: Grade against expected.ideal_behavior and expected.rubric. Penalize failure_modes.

How many test cases does this eval pack include?+

The Ai Triage pack for Linear Linear contains 10 test cases. 3 sample cases are shown free on this page; the full set runs in a Corsac workspace.

How do I run this eval?+

Sign up for Corsac, connect your model or agent endpoint, and run the Ai Triage pack as-is or after customizing thresholds. Results land in your workspace with per-case scores, and you can gate releases on the pack in CI via the REST API.

Run this eval in your workspace

Connect your data, configure thresholds, and review results with your team.