01
Fireworks Batch Prompt Cache Runtime Performance
Evaluates Fireworks AI's Batch, Prompt Cache & Runtime Performance across 13 scenario-based test cases, each graded against an expected-behavior rubric by an LLM judge, from Corsac's AI infrastructure eval coverage.
Mapped capabilities
13 scenarios
- Prompt cache prefix
- Prompt cache discipline
- Streaming TTFT
Public sample case
- Input
- Same 8k-token document prefix across requests; cache should reduce cost on shared prefix.
- Expected behavior
- Place static RAG context in stable system message prefix; keep variable user query suffix; rely on documented prompt cache behavior.
- Check
- Pass / fail check






