Datasets

Dataset

Bharat_NanoQuoraRetrieval_or

by carlfeynman

Bharat-NanoBEIR: Indian Language Information Retrieval Dataset Overview This dataset is part of the Bharat-NanoBEIR collection, which provides information retrieval datasets for Indian languages. It is derived from the NanoBEIR project, which offers smaller…

CC-BY-4.0text-retrieval 30 01K < n < 10K rows

published 23 Jan 2025

Preview

First 5 rows of corpus / train, from the Hugging Face dataset viewer.

_idtext
58ଅନଲାଇନରେ ଟଙ୍କା ମାଗିବାର ସର୍ବୋତ୍ତମ ଉପାୟ କ ଣ?
127ମୁଁ କାହିଁକି ସବୁବେଳେ ଅବସାଦଗ୍ରସ୍ତ ହୋଇପଡ଼େ?
238ଏମିତି କିଛି ଅଦ୍ଭୁତ ଟେକ୍ନୋଲୋଜି କ ଣ ଅଛି ଯାହା ବିଷୟରେ ଅଧିକାଂଶ ଲୋକ ଜାଣନ୍ତି ନାହିଁ?
3311000 ତଳେ ଗଭୀର ବାସ୍ ସହିତ କେଉଁଟି ସର୍ବୋତ୍ତମ ଇୟରଫୋନ୍?
407ଲୋକମାନେ ହିଲାରୀ କ୍ଲିଣ୍ଟନଙ୍କୁ କାହିଁକି ଘୃଣା କରନ୍ତି?

Cite this

CC-BY-4.0textCite

Tags

task_categories:text-retrievaltask_ids:document-retrievalmultilinguality:monolingualsource_datasets:NanoQuoraRetrievallanguage:orlicense:cc-by-4.0size_categories:1K<n<10Kformat:parquetmodality:textlibrary:datasetslibrary:pandaslibrary:mlcroissantlibrary:polarsregion:ustext-retrieval