Skip to content
Courses/AI Router and content system/Quality scoring without hype

Quality scoring without hype

A router improves when you record which output was good, bad, or uncertain. Without scoring, you only have intuition.

Minimum score

Terminal
score:
  factuality: 1-5
  usefulness: 1-5
  format: 1-5
  safety: pass|fail
  needs_human: true|false
  reason: "invented citation in the second paragraph"

Who scores

  • Deterministic rules: valid JSON, complete fields, citations present.
  • Automatic evals: compare against expected cases.
  • LLM-as-judge: useful for triage, not as absolute truth.
  • Human: required for editorial, legal, financial, or reputational content.

Official source

Langfuse documents traces, evaluation, datasets, and scores for self-hostable LLM applications: https://langfuse.com/docs/observability/overview

Complete Aulafy mapSee how this lesson fits without leaving your path.

Complete Aulafy map

How all courses connect

This is not a checklist. Start with the foundation, choose an outcome, and go deeper only when your project needs more control.

  1. 1Understand
  2. 2Apply or build
  3. 3Operate with confidence
01

Choose an application

Turn the foundation into a visible outcome: a website, a business improvement, media, or an interactive experience.

Continue into the technical branch when you need to maintain code, data, or infrastructure.

02

Build with code

Prepare your environment, work with coding agents, and run models while keeping control of your projects.

This branch prepares you to design and operate reliable AI systems.

03

Take systems to production

Combine retrieval, agents, evaluation, security, deployment, and model adaptation when the problem requires it.

You do not need every course: choose the component your system needs and return as it grows.

View full catalogue