AI product

VI codebench

VI codebench is a coding benchmark that measures how well AI models turn natural-language prompts into full-stack web applications. It is associated with VALS, an AI evaluation company that develops continuously evolving assessments for measuring model performance before release and in practical coding workflows.

Mentioned in 1 video ↓

What VI codebench is used for

1 use taken from transcripts — each links to the moment in the video.

  • A coding benchmark that measures how well models can turn a natural-language prompt into a full-stack web application.

Videos mentioning VI codebench

1 in the library.