raghu@dark-factory :~/kb/llm-jury $ cat

LLM jury

Aggregating verdicts from a panel of several diverse LLM evaluators instead of a single judge, to label data and score outputs more reliably and at lower cost.

grounded in: latest AI/tech trends theme 'LLMs as evaluators and juries': 'Using panels of LLMs to judge and label data is emerging as a practical building pattern.'

Connected concepts

LLM-as-judge, Evals & benchmarks, Multi-Agent Debate, Self-Consistency

Explore it live in the knowledge graph →