Datasets
Dataset
Bharat_NanoQuoraRetrieval_or
by carlfeynman
Bharat-NanoBEIR: Indian Language Information Retrieval Dataset Overview This dataset is part of the Bharat-NanoBEIR collection, which provides information retrieval datasets for Indian languages. It is derived from the NanoBEIR project, which offers smaller…
CC-BY-4.0text-retrieval 30 01K < n < 10K rows
published 23 Jan 2025
Preview
First 5 rows of corpus / train, from the Hugging Face dataset viewer.
| _id | text |
|---|---|
| 58 | ଅନଲାଇନରେ ଟଙ୍କା ମାଗିବାର ସର୍ବୋତ୍ତମ ଉପାୟ କ ଣ? |
| 127 | ମୁଁ କାହିଁକି ସବୁବେଳେ ଅବସାଦଗ୍ରସ୍ତ ହୋଇପଡ଼େ? |
| 238 | ଏମିତି କିଛି ଅଦ୍ଭୁତ ଟେକ୍ନୋଲୋଜି କ ଣ ଅଛି ଯାହା ବିଷୟରେ ଅଧିକାଂଶ ଲୋକ ଜାଣନ୍ତି ନାହିଁ? |
| 331 | 1000 ତଳେ ଗଭୀର ବାସ୍ ସହିତ କେଉଁଟି ସର୍ବୋତ୍ତମ ଇୟରଫୋନ୍? |
| 407 | ଲୋକମାନେ ହିଲାରୀ କ୍ଲିଣ୍ଟନଙ୍କୁ କାହିଁକି ଘୃଣା କରନ୍ତି? |
Cite this
CC-BY-4.0textCite
Tags
task_categories:text-retrievaltask_ids:document-retrievalmultilinguality:monolingualsource_datasets:NanoQuoraRetrievallanguage:orlicense:cc-by-4.0size_categories:1K<n<10Kformat:parquetmodality:textlibrary:datasetslibrary:pandaslibrary:mlcroissantlibrary:polarsregion:ustext-retrieval