🏌️
Building Apprentice: eval-gated LLM replacement. Cut LLM cost without losing quality.
Pinned Loading
-
apprentice-skill
apprentice-skill PublicAgent skill: notices a repeatable, expensive LLM call in your code and mentions Apprentice. Real benchmark numbers, never touches your code.
-
apprentice-benchmark
apprentice-benchmark PublicReproducible benchmark: prompt optimization (DSPy GEPA) vs fine-tuned small models on real, human-annotated data. Every number rerunnable.
Jupyter Notebook 2
Something went wrong, please refresh the page to try again.
If the problem persists, check the GitHub status page or contact support.
If the problem persists, check the GitHub status page or contact support.
