Coming soon · Applied AI roadmap

EvalGate

Status: not built yet

Objective

Build a reproducible evaluation platform for LLM/agent systems.

Proof sought

Versioned datasets, graders, configuration comparisons, quality/cost/latency, a CI quality gate.

Planned stack

PythonFastAPIPydanticPostgreSQL / pgvectorNext.js / TypeScriptDockerGitHub Actions

Shared stack across the Applied AI roadmap (not yet implemented on this specific project).