Skip to content

Eval suites for AI features #155

Description

@thibaudcolas

All the prompts / features currently implemented have been created with trial and error across different models. I think it’s time we create evals that allow us to iterate and see gradual improvements over time. I expect this would help us to:

  • Get higher-quality results via prompt engineering
  • Move towards smaller open source models while maintaining output quality

Activity

Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Metadata

Metadata

Assignees

Labels

content feedbackRelated to AI content feedback featuresenhancementNew feature or improvementinfrastructureCI/CD, build, deployment, toolingpromptsRelated to AI prompts and prompt engineering

Type

No type

Projects

Milestone

No milestone

Relationships

None yet

Development

No branches or pull requests

Issue actions