OpenRuna
Sign in

Vectara Hallucination Leaderboard

BENCHMARK

Leaderboard comparing LLM performance at producing hallucinations when summarizing short documents. Systematic evaluation of factual consistency across major models. Apache 2.0 licensed

View on GitHub

Overview

Vectara Hallucination Leaderboard: a free, copy-ready benchmark on OpenRuna. Leaderboard comparing LLM performance at producing hallucinations when summarizing short documents. Systematic evaluation of factual consistency across major models. Apac

What this benchmark does

Looking for a dependable benchmark? "Vectara Hallucination Leaderboard" gives you a tested starting point instead of a blank prompt box. Leaderboard comparing LLM performance at producing hallucinations when summarizing short documents. Systematic evaluation of factual consistency across major models. Apache 2.0 licensed It is catalogued next to similar resources on OpenRuna, so the rest of the toolkit you need is close by. Paste it straight into a chat, drop it into a system prompt, or store it as a reusable skill.

Use cases

  • Fork it as a baseline and layer in your own project context, constraints, and examples.
  • Keep it in a shared library as the canonical version of this benchmark for your organisation.
  • Use "Vectara Hallucination Leaderboard" when you need a repeatable benchmark for professional work without rewriting instructions every time.
  • Hand "Vectara Hallucination Leaderboard" to a new teammate so their benchmark output matches your team's quality bar from day one.

Example output

Ask the model to apply "Vectara Hallucination Leaderboard" to your scenario and it returns a structured answer — clear sections, actionable steps, and assumptions stated upfront — ready to paste into docs, tickets, or code comments. A short follow-up turn usually tightens the result to exactly what you need.

Tips by platform

Claude

With Claude, drop this benchmark into Project knowledge so every chat in the project inherits it. Ask Claude to restate the goal first, then run — it catches edge cases early.

ChatGPT

For ChatGPT, save this benchmark as a Custom Instruction or a saved prompt so it is one click away. Add your specifics in a follow-up rather than editing the original.

Cursor

Add this benchmark to your Cursor rules and invoke it from Agent mode for repeatable results. Link back to its OpenRuna page in the rule so the source stays discoverable.

Frequently asked questions

What is "Vectara Hallucination Leaderboard"?
It is a benchmark listed on OpenRuna — Leaderboard comparing LLM performance at producing hallucinations when summarizing short documents. Systematic evaluation of factual consistency across major models. Apache 2.0 licensed You can copy and adapt it for ChatGPT, Claude, Cursor, or any other AI assistant.
Is "Vectara Hallucination Leaderboard" free to use?
Most OpenRuna resources are open or CC0-licensed. Check the license shown on this page before commercial use; premium collections are clearly marked as such.
How do I get the best results from this benchmark?
Replace any placeholders, add your project context, and ask the model to confirm its assumptions first. Iterate over 2–3 follow-up turns rather than expecting a perfect first response.
Does "Vectara Hallucination Leaderboard" work with both Claude and ChatGPT?
Yes — it is model-agnostic text, so it runs on Claude, ChatGPT, Gemini, and Cursor. The tips on this page cover each of those assistants specifically.
Where can I find resources related to "Vectara Hallucination Leaderboard"?
Scroll to the Related resources section on this page, or open the matching category hub on OpenRuna to find connected prompts, tools, agents, and datasets in the same topic area.

Related resources