Datasets

Dataset

roots_indic-or_pib

by bigscience-data

ROOTS Subset: rootsindic-orpib pib Dataset uid: pib Description Sentence aligned parallel corpus between 11 Indian Languages, crawled and extracted from the press information bureau website. Homepage Licensing Creative Commons Attribution-ShareAlike 4.

CC-BY-SA-4.0other 12 010K < n < 100K rows

published 18 May 2022

Preview

The dataset does not exist, or is not accessible without authentication (private or gated). Please check the spelling of the dataset name or retry with authentication.

Cite this

CC-BY-SA-4.0textCite

Tags

language:orlicense:cc-by-sa-4.0size_categories:10K<n<100Kformat:parquetmodality:textlibrary:datasetslibrary:pandaslibrary:mlcroissantlibrary:polarsregion:us