
Ai Triage
Linear · Linear
Project & Issue Management — Linear
Evaluates Linear's AI Triage & Assistance across 10 scenario-based test cases, each graded against an expected-behavior rubric by an LLM judge, from Corsac's Project & Issue Management eval coverage.
About Linear
Linear is a project and issue management tool built for high-velocity software teams. It pairs a fast, keyboard-driven UI with a GraphQL API, cycles, projects, and roadmaps, plus deep GitHub and Slack integrations and AI-assisted triage.
Sample tests· showing 3 of 10
| # | Input | Expected behavior | Check |
|---|---|---|---|
| 01 | Issue text mentions Payments API errors. Team Payments on-call rotation exists. AI suggestion must cite team and past similar issues, not random user. | Propose on-call from Payments team with citation to issue content; allow operator override; tag model confidence [REQUIRES-VERIFICATION]. | Pass / FailAihigh |
| 02 | Description includes stack trace and 'regression'. Agent Interaction Guidelines require human-visible reasoning tied to issue body. | Suggest bug/regression labels quoting stack trace lines; operator confirms before apply. | Pass / FailAimedium |
| 03 | Long issue comment thread on INC-9. Summary must cite comment authors and dates, not invent decisions. | Produce bullet summary with pointers to comment ids/timestamps; mark unknowns; no fabricated approvals. | Pass / FailAihigh |
How this eval is graded
Grade against expected.ideal_behavior and expected.rubric. Penalize failure_modes.
Rubric criteria
- Linear
- Code Assistant
- Ai Triage
Recommended for
Works with
Related evals
Browserbase
Evaluates Browserbase's Captcha Handling across scenario-based test cases, each graded against an expected-behavior rubric by an LLM judge, from Corsac's Browser infrastructure eval coverage.
View Code AssistantBrowserbase
Evaluates Browserbase's Concurrency & Rate Limits across scenario-based test cases, each graded against an expected-behavior rubric by an LLM judge, from Corsac's Browser infrastructure eval coverage.
View Code AssistantBrowserbase
Evaluates Browserbase's Live Debugging & Session Inspector across scenario-based test cases, each graded against an expected-behavior rubric by an LLM judge, from Corsac's Browser infrastructure eval coverage.
ViewFrequently asked questions
What does the Ai Triage eval for Linear Linear test?+
Evaluates Linear's AI Triage & Assistance across 10 scenario-based test cases, each graded against an expected-behavior rubric by an LLM judge, from Corsac's Project & Issue Management eval coverage.
How is the Ai Triage eval scored?+
The judge rubric: Grade against expected.ideal_behavior and expected.rubric. Penalize failure_modes.
How many test cases does this eval pack include?+
The Ai Triage pack for Linear Linear contains 10 test cases. 3 sample cases are shown free on this page; the full set runs in a Corsac workspace.
How do I run this eval?+
Sign up for Corsac, connect your model or agent endpoint, and run the Ai Triage pack as-is or after customizing thresholds. Results land in your workspace with per-case scores, and you can gate releases on the pack in CI via the REST API.
Run this eval in your workspace
Connect your data, configure thresholds, and review results with your team.