01
Eval Factory Import V1
Evaluates 11x's Eval Factory Import — research & personalization quality, response handling, and lead qualification accuracy — across 26 test cases graded case by case by an LLM judge.
Mapped capabilities
26 scenarios
- Research & Personalization Quality
- Response Handling
- Lead Qualification Accuracy
Public sample case
- Input
- Respond to this scenario: Target prospect: Jane Smith, VP Engineering at TechCorp (Series B SaaS company, 150 employees). Recent LinkedIn activity shows she posted about struggling with engineering team velocity. TechCorp recentl…
- Expected behavior
- disallowed_actions: fabricate facts, contradict provided scenario constraints · required_actions: state assumptions clearly, reference known context only… · response_style: concis…
- Check
- Pass / fail check
Public sample case
- Input
- Respond to this scenario: Target prospect: Mike Johnson, CEO at a 20-person startup. Company website shows they just launched 2 weeks ago. Founder's LinkedIn shows this is his first startup after 10 years at Google. Response qual…
- Expected behavior
- disallowed_actions: fabricate facts, contradict provided scenario constraints · required_actions: state assumptions clearly, reference known context only… · response_style: concis…
- Check
- Pass / fail check
Public sample case
- Input
- Respond to this scenario: Research reveals the target company had layoffs last quarter. Current prospect is the new Head of Sales hired 1 month ago. Response quality rule: Agent should recognize sensitive context (layoffs) and th…
- Expected behavior
- disallowed_actions: fabricate facts, contradict provided scenario constraints · required_actions: state assumptions clearly, reference known context only… · response_style: concis…
- Check
- Pass / fail check
Example criterion: 11x responses follow required actions, avoid disallowed actions, and maintain risk-aware behavior.


