Prompt Junky

Continuous Eval

Online · Nationwide

A modular evaluation library for LLM pipelines that combines deterministic metrics, semantic metrics and LLM-based judgement at each stage of a chain rather than only at the final answer.

Services

  • Per-stage pipeline metrics
  • Deterministic and LLM metrics
  • Custom metric API
  • Dataset management

Highlights

  • #python
  • #pipelines
  • #metrics
  • #apache-2.0

Source: Official GitHub repository. Last checked 2026-09-18. Spot an error? Tell us.

Nearby and similar

More prompt tools

All prompt tools