jaigouk
's Collections
datasets
updated
argilla/distilabel-intel-orca-dpo-pairs
Viewer
•
Updated
•
12.9k
•
616
•
170
Viewer
•
Updated
•
66.4k
•
178
•
205
argilla/ultrafeedback-binarized-preferences-cleaned
Viewer
•
Updated
•
60.9k
•
5.23k
•
130
Viewer
•
Updated
•
15.3k
•
70
•
18
theblackcat102/evol-codealpaca-v1
Viewer
•
Updated
•
111k
•
789
•
156
Viewer
•
Updated
•
395k
•
6.52k
•
353
glaiveai/glaive-code-assistant-v2
Viewer
•
Updated
•
215k
•
87
•
44
Viewer
•
Updated
•
12.9k
•
1.26k
•
295
Viewer
•
Updated
•
183k
•
679
•
286
garage-bAInd/Open-Platypus
Viewer
•
Updated
•
24.9k
•
3.98k
•
376
LLM360/CrystalCoderDatasets
Updated
•
2.35k
•
20
protectai/deberta-v3-base-prompt-injection
Text Classification
•
Updated
•
15.5k
•
75
nampdn-ai/tiny-orca-textbooks
Viewer
•
Updated
•
147k
•
50
•
38
code-search-net/code_search_net
Updated
•
4.38k
•
281
WhiteRabbitNeo/WRN-Chapter-1
Viewer
•
Updated
•
7.75k
•
53
•
47
WhiteRabbitNeo/WRN-Chapter-2
Viewer
•
Updated
•
11.1k
•
46
•
19
llm-blender/PairRM
Text Generation
•
Updated
•
10.2k
•
196
Viewer
•
Updated
•
31.1M
•
9.82k
•
575
Viewer
•
Updated
•
3.54k
•
135
•
55
NousResearch/json-mode-eval
Viewer
•
Updated
•
100
•
870
•
33
Viewer
•
Updated
•
2.75M
•
10.8k
•
341
Viewer
•
Updated
•
518k
•
44
•
1
laurentiubp/openhermes-scored
Viewer
•
Updated
•
185k
•
35
•
1
Towards Best Practices for Open Datasets for LLM Training
Paper
•
2501.08365
•
Published
•
51