Team Ai
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01OdiaGenAI /RAG_Evaluation_Datasettabular1K<n<10K0 likes206 downloads3y agoHugging Face02tarekmasryo /rag-qa-logs-corpus-data 🧠📚 RAG QA Logs & Corpus (Synthetic) 🧪 Multi-table synthetic RAG telemetry for quality, hallucinations, latency, and cost A production-style, privacy-safe synthetic dataset that mimics telemetry exported from a real RAG system — from corpus → chunks → retrieval events → eval runs. ✅ Fully synthetic (no real users / orgs / PII). ⚡ Quick facts Total rows: 103,255 across 6 linked tables Labels (in eval_runs): is_correct, hallucination_flag, faithfulness_label… See the full description on the dataset page: https://huggingface.co/datasets/tarekmasryo/rag-qa-logs-corpus-data.tabularquestion-answering100K<n<1M2 likes93 downloads8mo agoHugging Face03dev7halo /kor-rag-opentesttabularn<1K0 likes91 downloads2y agoHugging Face04mtntasci /turkish-legal-rag Turkish Legal RAG Corpus — Türk Hukuku için Açık RAG Datasetı Tek cümle: 25 önemli Türk kanununun (mevzuat.gov.tr kaynaklı, madde bazlı temiz chunk'lar) + 290 manuel doğrulanmış soru-cevap altın benchmark'ının olduğu açık kaynak Türkçe hukuk RAG datasetı. 🇹🇷 Türkçe Özet — Bu dataset, Türkçe hukuk uygulamaları için sıfırdan üretilmiş açık ve denetlenebilir bir RAG corpus'udur. mevzuat.gov.tr üzerinden alınan 25 ana kanunun madde madde temizlenmiş, chunk'lanmış sürümünü (6.350… See the full description on the dataset page: https://huggingface.co/datasets/mtntasci/turkish-legal-rag.tabulartext-retrieval1K<n<10K2 likes74 downloads5mo agoHugging Face05eneSadi /turkuaz-rag Turkuaz-RAG: A Novel Turkish Multi-Context Retrieval Benchmark Turkuaz-RAG is the first benchmark specifically created for evaluating multi-context retrieval tasks in Turkish. It addresses a major gap in low-resource language research by providing multi-context questions, answers, and corresponding contexts. Description of Benchmark Languages: Turkish Size: ~2,500 triplets (question, contexts, answer) Context Sources: Turkish news articles from MLSUM Question Types:… See the full description on the dataset page: https://huggingface.co/datasets/eneSadi/turkuaz-rag.tabularquestion-answering1K<n<10K3 likes67 downloads11mo agoHugging Face06vkshdev /rag-hallucination-benchmark RAG Hallucination Benchmark Context Retrieval-Augmented Generation (RAG) is the industry standard for reducing LLM hallucinations, but detecting when a RAG system fails is a massive challenge. Most existing benchmarks focus only on massive Deep Learning models and lack tabular features. This dataset provides a clean, engineered setup to train models (from XGBoost to RoBERTa) to detect hallucinations, predict context faithfulness, and measure answer relevance.… See the full description on the dataset page: https://huggingface.co/datasets/vkshdev/rag-hallucination-benchmark.tabulartext-classification10K<n<100K0 likes62 downloads1mo agoHugging Face07MahdiAbdoZahra /RAG_vs_FineTuning_Comparison_Persian_V1tabularn<1K1 likes60 downloads20d agoHugging Face08Ragab-Adel /privacy-preserving-real-world-human-motion-sample Privacy-Preserving Real-World Human Motion Sample A market-validation sample of anonymous 2D skeleton/pose observations derived from a real-world indoor CCTV stream. Why this sample exists We are validating demand for continuously collected, privacy-oriented real-world human-motion data before expanding to multi-camera releases. Current public sample 750 public observations derived pose/skeleton data anonymous track identifiers no raw RGB video no… See the full description on the dataset page: https://huggingface.co/datasets/Ragab-Adel/privacy-preserving-real-world-human-motion-sample.tabularothern<1K1 likes58 downloads2mo agoHugging Face09omaressam1111 /multi-tafseer-quran-rag Quran Tafseer RAG Dataset A structured Arabic dataset of Quranic tafseer collected from eight classical and modern tafseer books.The dataset contains verse-aligned tafseer passages designed for Retrieval-Augmented Generation (RAG) systems and Arabic NLP research. Each record links a Quran verse with its corresponding tafseer explanation from one of the tafseer books and includes rich metadata such as surah information, tafseer source, and embedding-ready text. The dataset was… See the full description on the dataset page: https://huggingface.co/datasets/omaressam1111/multi-tafseer-quran-rag.tabularquestion-answering10K<n<100K7 likes43 downloads5mo agoHugging Face10vein05 /ragscale-interaction-matrix ragscale Interaction Matrix Reader answers under raw and compressed RAG evidence: 176,864 rows, one per benchmark item, reader model, and evidence policy, across LongMemEval, HotpotQA, MuSiQue, and NQ-Open. This is the interaction matrix released with the paper Compression Is Not Evaluation-Neutral: Fixed RAG Compression Can Distort Reader Comparisons (Sugam Panthi and Rabab Abdelfattah, arXiv:2606.21807). The paper gives 8 to 20 reader models the same stored compressed text and… See the full description on the dataset page: https://huggingface.co/datasets/vein05/ragscale-interaction-matrix.tabularquestion-answering100K<n<1M0 likes37 downloads5d agoHugging Face11dev-jonathanb /cs50-educational-rag CS50 Pedagogical RAG Dataset 📜 Dataset Description This repository contains the data artifacts for the undergraduate thesis, which explores the use of a pedagogical chatbot with Retrieval-Augmented Generation (RAG) for Harvard's CS50: Introduction to Computer Science course. The project involved several stages of data processing, from raw content collection to the generation and curation of a high-quality evaluation dataset. To ensure full transparency and… See the full description on the dataset page: https://huggingface.co/datasets/dev-jonathanb/cs50-educational-rag.tabularquestion-answeringn<1K0 likes31 downloads1y agoHugging Face12AdamLucek /legal-rag-positives-synthetic Synthetic QnA Chunk Pairs from Legal Documents This dataset contains excerpts from legal cases' court opinions that mention artificial intelligence, along with corresponding question-answer pairs derived from the content. The data was sourced from CourtListener's public API and processed to create a structured dataset suitable for question-answering tasks. Specifically including cases: Senetas Corporation, Ltd. v. DeepRadiology Corporation Electronic Privacy Information Center v.… See the full description on the dataset page: https://huggingface.co/datasets/AdamLucek/legal-rag-positives-synthetic.tabular1K<n<10K4 likes29 downloads2y agoHugging Face13SakataConsul /rag-eval-ja-repro RAG Eval JA Repro Current version / 現行版: v1.1 2026-07-12 更新(v1.1): rag_evaluation_master.csv、採用PDF manifest、PDF checksumを更新し、6月30日公開時のローカル精度検証を同じ4条件で再実行しました。旧版の記述は取り消し線で残し、v1.1の値を併記します。 TL;DR (EN): A derived reproducibility dataset for allganize/RAG-Evaluation-Dataset-JA. It adds (1) derived *_new answer/question columns (with per-item rationale), and (2) a Wayback-pinned + SHA-256 corpus manifest so anyone can fetch byte-identical source PDFs. The original CSV is not modified;… See the full description on the dataset page: https://huggingface.co/datasets/SakataConsul/rag-eval-ja-repro.tabularquestion-answeringn<1K0 likes26 downloads3mo agoHugging Face14MarcoFurrer /swiss-building-law-rag-bench Swiss Cantonal Building Law RAG Benchmark Evaluation benchmark for Retrieval-Augmented Generation (RAG) systems on Swiss cantonal building law documents. Created as part of a bachelor thesis on systematic RAG pipeline optimisation for German legal text. Dataset contents File Entries Language Description data/german/golden_dataset.jsonl 318 DE German Q&A pairs grounded to article-level passages data/multilingual/golden_dataset.jsonl 270 DE/FR/IT… See the full description on the dataset page: https://huggingface.co/datasets/MarcoFurrer/swiss-building-law-rag-bench.tabularquestion-answeringn<1K1 likes20 downloads4mo agoHugging Face15raghad23 /eou_AudioTextaudio100K<n<1M0 likes16 downloads10mo agoHugging Face16ofai /RagabilityCorpusCurrent version: Dataset_v0.4.tsv (converted to ragability format: v0d4.hjson) Ragability Corpus In the following, we introduce WikiContradict (the empirical basis for the Ragability Corpus), describe the Ragability Corpus, and finally explain how the dataset can be extended and how a new one can be created. Empirical basis WikiContradict is a benchmark for evaluating LLMs on real-world knowledge conflicts from Wikipedia (see the Hou et. al. 2025 and the dataset for more… See the full description on the dataset page: https://huggingface.co/datasets/ofai/RagabilityCorpus.tabularn<1K0 likes14 downloads7mo agoHugging Face17Egertekin /turkish-hospital-medical-rag-advanced Turkish Hospital Medical Articles - Advanced RAG System & Vector Database Bu proje, 14 farklı hastane grubuna ait geniş ölçekli Türkçe tıbbi makaleler üzerinde çalışan, ileri düzey teknik parametrelerle optimize edilmiş bir Retrieval-Augmented Generation (RAG) sistemi ve Vektör Veritabanı uygulamasıdır. Proje kapsamında ham veriler Hugging Face üzerinden tüm hastane split'leriyle çekilmiş, gelişmiş chunking stratejileriyle parçalanmış, magibu/embeddingmagibu-200m modeliyle… See the full description on the dataset page: https://huggingface.co/datasets/Egertekin/turkish-hospital-medical-rag-advanced.tabular10K<n<100K0 likes14 downloads2mo agoHugging Face18SMARTICT /Pubmed-RAG-TR-LLM-EvalLLM-as-a-judge evalaution results using "claude-haiku-4-5-20251001" for SMARTICT/Pubmed-RAG-TR-LLM dataset. tabular1K<n<10K0 likes13 downloads8mo agoHugging Face19jason1966 /algozee_rag-based-hallucination-reduction-in-llms RAG-Based Hallucination Reduction in LLMs Introduction to Large Language Models and Hallucination Problem Dataset Info Source: Kaggle Original Size: 0.17 MB Kaggle Downloads: 43 Files: 1 Files llm_rag_dataset_6k.csv.csv Mirrored from Kaggle tabular1K<n<10K0 likes13 downloads6mo agoHugging Face20seanlee10 /legal-rag-qatabular10K<n<100K0 likes13 downloads5mo agoHugging Face21Valentinaewelu /ragbench-5dtabularn<1K0 likes9 downloads7mo agoHugging Face22joelkoch /rag_evalSome datasets for evaluating RAG systems, created by following this huggingface cookbook. tabularn<1K0 likes8 downloads2y agoHugging Face23PandaBambooVane /ragas_evaluationV1tabularn<1K0 likes8 downloads2y agoHugging Face24raguv /healthcare_datasettabularn<1K0 likes7 downloads2y agoHugging Face25PandaBambooVane /RAG12000-LLaMA3.1-8B-gguf_AR-RAG_v2tabularn<1K0 likes6 downloads2y agoHugging Face26RaginiPranay /superkart-sales-datasettabular10K<n<100K0 likes6 downloads6mo agoHugging Face27ragzhf2026 /GL-Raghu-engine-predictive-maintainencetabular10K<n<100K0 likes6 downloads4mo agoHugging Face28raghavneon /test_123tabularn<1K0 likes5 downloads3y agoHugging Face29mayur456 /court_data_RAG_unsuptabularn<1K0 likes5 downloads3y agoHugging Face30raghurayar /Machine-Failure-Predictiontabular10K<n<100K0 likes5 downloads10mo agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.