Viteval

Next generation LLM evaluation framework powered by Vitest.

Documentation | Getting Started | Examples

Features

✅ Vitest-like API: easily define and run evals the same way you run tests.
✅ Local Dataset: quickly define and generate datasets that are stored locally or in a database.
✅ Scorer framework: easily define and use scorers to evaluate your LLM.
✅ CI ready: easily integrate with your CI pipeline and run evals the same way you run tests.
🧪 UI for viewing the results of Evals (alpha)
(SOON) Report Exports: custom reporters to help you visualize your evals, or upload them to a 3rd party
(SOON) Remote Datasets: easily define & use datasets that are stored in a database, S3 bucket, or a 3rd party

Installation

npm install viteval

Usage

You can use viteval to evaluate your LLM and add tests in your CI in one single file.

import { evaluate, scorers } from 'viteval';
import { generateText } from 'ai';

evaluate('Color detection', {
  data: async () => [
    { input: "What color is the sky?", expected: "Blue" },
    { input: "What color is grass?", expected: "Green" },
  ],
  task: async (input) => {
    const result = await generateText(input);
    return result.text;
  },
  scorers: [scorers.levenshtein],
  threshold: 0.8,
});

Now you can run the eval by running:

npx viteval

Name		Name	Last commit message	Last commit date
Latest commit History 134 Commits
.changeset		.changeset
.cursor/rules		.cursor/rules
.github		.github
.vscode		.vscode
docs		docs
examples		examples
packages		packages
tools		tools
.gitignore		.gitignore
.nukeignore		.nukeignore
AGENTS.md		AGENTS.md
CLAUDE.md		CLAUDE.md
CONTRIBUTING.md		CONTRIBUTING.md
LICENSE		LICENSE
README.md		README.md
biome.json		biome.json
jog.yaml		jog.yaml
nx.json		nx.json
package.json		package.json
pnpm-lock.yaml		pnpm-lock.yaml
pnpm-workspace.yaml		pnpm-workspace.yaml
tsconfig.base.json		tsconfig.base.json
vitest.workspace.ts		vitest.workspace.ts

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

Uh oh!

Repository files navigation

Viteval

Features

Installation

Usage

About

Uh oh!

Releases 92

Uh oh!

Contributors 3

Uh oh!

Languages

License

viteval/viteval

Folders and files

Latest commit

History

Repository files navigation

Viteval

Features

Installation

Usage

About

Resources

License

Contributing

Uh oh!

Stars

Watchers

Forks

Releases 92

Uh oh!

Contributors 3

Uh oh!

Languages