OPEN-SOURCE PROJECT

goodwatch-monorepo

A film-and-TV attribute-scoring experiment inside GoodWatch comparing Jev question designs and batch sizes.

Stars
38
Forks
2
License
MIT
Last commit
2026-09-20

What Jev does here

Asks whether predefined traits are present or how strongly they appear, recording scores, latency and Token usage.

Compares rating scales, input variants and batching over a frozen sample.

The source identifies a standalone scoring experiment, not proof of production recommendation use or superiority to vector recommendations; not rerun here.