Amnesty QA Dataset A grounded question answering dataset for evaluating RAG (Retrieval Augmented Generation) systems, created from reports collected from Amnesty International. This dataset is designed for testing and evaluating RAG pipelines with real world human rights content. Dataset Structure Each sample contains: user input : The question to be answered reference : Ground truth answer for evaluation response : Generated answer from the system retrieved contexts : List of relevant context passages retrieved for answering the question Example Usage Available Languages The dataset is available in three languages (all use the v3 schema): English (recommended): english v3 Hindi : hindi v3 Malayalam : malayalam v3 Dataset Splits Only the eval split is available for this dataset, containing 20 carefully curated question answer pairs. Legacy Versions ⚠️ Note : Versions v1 and v2 are deprecated and maintained only for backwards compatibility. Please use v3 for all new projects. Legacy version schemas (click to expand) v1 (deprecated): question , ground truths (list), answer , contexts v2 (deprecated): question , ground truth (string), answer , contexts Citation If you use this dataset…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy