Datasets
Dataset
roots_indic-or_pib
by bigscience-data
ROOTS Subset: rootsindic-orpib pib Dataset uid: pib Description Sentence aligned parallel corpus between 11 Indian Languages, crawled and extracted from the press information bureau website. Homepage Licensing Creative Commons Attribution-ShareAlike 4.
CC-BY-SA-4.0other 12 010K < n < 100K rows
published 18 May 2022
Preview
The dataset does not exist, or is not accessible without authentication (private or gated). Please check the spelling of the dataset name or retry with authentication.
Cite this
CC-BY-SA-4.0textCite
Tags
language:orlicense:cc-by-sa-4.0size_categories:10K<n<100Kformat:parquetmodality:textlibrary:datasetslibrary:pandaslibrary:mlcroissantlibrary:polarsregion:us