All evals
K

Eval directory

Evals for Kalepa

Eval coverage for Kalepa, mapped from its public product surface.

About Kalepa

Kalepa is an AI underwriting platform that analyzes insurance submissions and surfaces risk insights for underwriters. It covers the workflow from submission ingestion and clearance through triage, risk analysis, and portfolio management across commercial lines. Founded in 2018, it is sold to insurers, MGAs, and specialty carriers.

Industry

AI underwriting platform for commercial insurers

Use the eval library for Kalepa

We'll build out the full library — runnable test cases with inputs, expected behavior, and pass/fail checks — in your Corsac workspace.

Generate your own →

Coverage map

What would you measure for Kalepa?

6 scoring areas · 23 capabilities mapped · grounded in 8 cited pages

Every eval set is graded on

  • Adversarial robustness
  • Workflow quality
  • Safety gates
  • Operator quality

Pass/Fail + LLM judge 1–5 · critical severity flags · negative controls

01

Submission Ingestion & Document Extraction

Turning inbound submission packets into structured, underwritable data across the document types Kalepa cites: ACORDs, SOVs, loss runs, and supplemental applications.

Automatically classify submissions and extract data from across hundreds of document types kalepa.com

Mapped capabilities

4 capabilities

  • Document classification

    Correctly identifying document type across hundreds of formats within a single submission packet, including mixed and mislabeled attachments.

  • Field extraction fidelity

    Pulling named insured, exposures, limits, and loss history values with correct units, dates, and entity attribution.

  • Loss run and SOV parsing

    Handling the harder structured artifacts — multi-year loss runs and large schedules of locations — without dropping or merging rows.

  • Degraded input handling

    Behavior on scans, partial packets, and unreadable pages: surfacing what could not be extracted instead of silently inventing values.

02

Clearance & Submission Control

The gatekeeping steps Kalepa lists before a submission reaches an underwriter: conflict detection, completeness, screening, producer verification, and routing.

Detect conflicts, verify completeness, conduct sanctions screening, confirm appointed producers, and route submissions kalepa.com

Mapped capabilities

4 capabilities

  • Conflict and duplicate detection

    Recognizing the same risk arriving via different brokers or name/address variants and blocking auto-clearance.

  • Completeness verification

    Determining whether a packet has what the workflow requires before it advances, and naming the specific gaps.

  • Sanctions screening

    Screening named parties and reporting matches with enough identifying detail for a compliance reviewer to adjudicate.

  • Producer appointment and routing

    Confirming the submitting producer is appointed and routing to streamlined or automated paths per the configured workflow.

Illustrative example

Input
A submission arrives for "Ridgeline Logistics LLC" at 400 Harbor Rd. An open submission already exists for "Ridgeline Logistics, L.L.C." at 400 Harbour Road, submitted by a different broker.
Expected behavior
Flags a clearance conflict rather than clearing the new submission. Identifies the prior submission as the match, names both producers, and routes the item for conflict resolution instead of advancing it to triage.

03

Triage & Appetite Alignment

Prioritizing the queue against appetite and guidelines so underwriters work the submissions most likely to bind profitably.

Automatically prioritize submissions most likely to bind, align them to appetite and guidelines kalepa.com

Mapped capabilities

4 capabilities

  • Appetite and guideline matching

    Aligning a submission to written appetite and underwriting guidelines, including in-appetite, out-of-appetite, and referral outcomes.

  • Bind-likelihood prioritization

    Ordering the queue by likelihood to bind and stating what drove the ranking.

  • Declination and referral signaling

    Producing an actionable disposition rather than a bare score, with the guideline clause that triggered it.

  • Cross-line consistency

    Holding triage behavior stable across the commercial lines Kalepa claims to cover.

04

Risk Analysis & Flag Generation

The single-pane analysis Kalepa markets: exposures, controls, terms, and communications reviewed together, with red/yellow/preferred flags surfaced to the underwriter.

Kalepa instantly analyzes submissions and surfaces the critical risk insights you need to underwrite faster kalepa.com

Mapped capabilities

4 capabilities

  • Flag assignment and severity

    Assigning red, yellow, or preferred consistent with the underlying finding rather than over- or under-escalating.

  • Evidence attribution

    Tying each flag to the document, field, or external source it came from so an underwriter can verify it.

  • Exposure and controls analysis

    Reading protection features, occupancy, and operations into risk-relevant conclusions.

  • Consolidated view assembly

    Bringing terms, communications, and analysis into one coherent view without contradicting itself across panels.

Illustrative example

Input
An SOV and site report for a 120,000 sq ft warehouse: the schedule shows 30% of the building sprinklered, and the nearest fire hydrant is 180 feet from the structure.
Expected behavior
Raises a yellow flag on partial sprinkler coverage, citing the 30% figure from the schedule. Treats the 180-foot hydrant proximity as a favorable protection detail, not as a second adverse finding.

05

Portfolio Management & Steering

Aggregation above the single submission: how individual decisions roll up into book-level views and steering signals across commercial portfolios.

End-to-end underwriting intelligence from submission to portfolio management. kalepa.com

Mapped capabilities

3 capabilities

  • Portfolio aggregation

    Rolling submission-level exposures and dispositions into accurate book-level totals and concentrations.

  • Steering signals

    Surfacing where the book is drifting relative to appetite and targeted mix.

  • Performance attribution

    Connecting portfolio outcomes back to the selection decisions that produced them.

06

Workflow Integration & Accountability

The 'harness' Kalepa argues is the real differentiator: embedding into existing underwriting processes with rules, auditability, and human accountability intact.

Our platform delivers decision-ready risk insights, automates high-friction workflow steps, and embeds intelligence directly into existing underwriting processes. kalepa.com

Mapped capabilities

4 capabilities

  • Rule capture and adherence

    Operating against an insurer's own underwriting logic and behaving predictably when that logic is ambiguous or unstated.

  • Auditability of outputs

    Leaving a reviewable trail of what was concluded, from which inputs, at what time.

  • Human handoff and escalation

    Handing judgment calls to an underwriter with the context needed to decide, rather than resolving them unilaterally.

  • Embedding in existing process

    Fitting into the incumbent workflow and system of record without requiring the underwriter to work outside it.

Coverage is mapped from Kalepa's public pages (8 crawled). Examples are illustrative, not real test cases. The runnable eval library — graded inputs, expected behavior, and pass/fail checks — is built when you request it above.

Frequently asked questions

What do the Corsac evals for Kalepa test?+

The coverage map is generated from Kalepa's own public product surface (AI underwriting platform for commercial insurers): 6 scoring areas — Submission Ingestion & Document Extraction, Clearance & Submission Control, and Triage & Appetite Alignment, and more — spanning 23 mapped capabilities, each graded on adversarial robustness, workflow quality, safety gates, and operator quality once the library is built.

How are the Kalepa evals scored?+

Every case generated for Kalepa — across Submission Ingestion & Document Extraction and Clearance & Submission Control and the other mapped areas — is graded with pass/fail checks plus an LLM judge scoring 1–5 against its expected behavior, with critical-severity flags and negative controls. Only judge-passed evals are published.

How many test cases does the Kalepa library include?+

The full Kalepa library is built on request. The coverage map spans 6 areas and 23 capabilities (for example, Document classification and Field extraction fidelity under Submission Ingestion & Document Extraction); each becomes graded test cases — inputs, expected behavior, pass/fail checks — in your Corsac workspace.

How do I run these evals against Kalepa or my own agent?+

Request the library with your work email above. We'll build out all 6 mapped Kalepa areas and set them up in a Corsac workspace, where you can run every test case against Kalepa or your own agent with your own data.