DistAya
's Collections
Datasets
updated
shayekh/perplexity__aya_dataset__train
Updated
•
43
Viewer
•
Updated
•
540k
•
72
•
1
argilla/magpie-ultra-v0.1
Viewer
•
Updated
•
50k
•
932
•
222
Magpie-Align/Magpie-Qwen2-Pro-1M-v0.1
Viewer
•
Updated
•
1M
•
288
•
14
HuggingFaceTB/smollm-corpus
Viewer
•
Updated
•
237M
•
11.6k
•
312
Viewer
•
Updated
•
100k
•
16.1k
•
179
BanglaLLM/bangla-alpaca-orca
Viewer
•
Updated
•
172k
•
93
•
3
AhmadMustafa/Urdu-Instruct-News-Article-Generation
Viewer
•
Updated
•
112k
•
254
•
4
AhmadMustafa/Urdu-Instruct-News-Headline-Generation
Viewer
•
Updated
•
112k
•
173
AhmadMustafa/Urdu-Instruct-News-Category-Classification
Viewer
•
Updated
•
112k
•
382
Viewer
•
Updated
•
10k
•
427
•
40
akbargherbal/six_millions_instruction_dataset_for_arabic_llm_ft
Viewer
•
Updated
•
6.37M
•
164
•
1
CohereForAI/aya_collection_language_split
Viewer
•
Updated
•
514M
•
32.4k
•
95
Viewer
•
Updated
•
63k
•
638
•
34
Viewer
•
Updated
•
20.4M
•
5.23k
•
602
convaiinnovations/Nadi_Indic466k_Instruct
Viewer
•
Updated
•
466k
•
124
•
2
ai4bharat/indic-instruct-data-v0.1
Viewer
•
Updated
•
404k
•
514
•
24
Viewer
•
Updated
•
9.97k
•
75
•
2
HAERAE-HUB/qarv-instruct-ko
Viewer
•
Updated
•
10.2k
•
77
•
20
MarkrAI/KoCommercial-Dataset
Viewer
•
Updated
•
175k
•
1.37k
•
143