Self-run
SKILL
Deepeval
>
0
8
Free Self-run
SKILL
Pipeline Eval
System-level evaluation framework for multi-stage LLM pipelines. Scores the pipeline *as a whole* across 8 dimensions — input quality, output quality, prompt design, input-coverage %, sequence optimality, parallelization, fact-grounding, and self-improvement feedback. Complements `deepeval` (which scores single content artifacts) by scoring the pipeline architecture itself. Trigger when user type…
0
8
Free You've reached the end · 2 loaded