Data, embedding, and index of MassiveDS by "Scaling Retrieval-Based Language Models with a Trillion-Token Datastore"
Rulin Shao
rulins
AI & ML interests
None yet
Recent Activity
updated
a dataset
about 9 hours ago
rulins/gpqa_preprocessed
published
a dataset
about 9 hours ago
rulins/gpqa_preprocessed
updated
a dataset
1 day ago
rulins/gpqa_searched_results_from_massiveds_non_cc
Organizations
Collections
1
models
4
datasets
13
rulins/gpqa_preprocessed
Viewer
•
Updated
•
1.19k
rulins/gpqa_searched_results_from_massiveds_non_cc
Viewer
•
Updated
•
198
•
2
rulins/MassiveDS-1.4T
Updated
•
2.47k
•
10
rulins/reasonir_bright_gpt4_reasoning_query_scores
Preview
•
Updated
•
11
rulins/DeepSeek-R1-Distill-Qwen-32B_NUMINA_train_amc_aime_merged_thoughts
Viewer
•
Updated
•
3.64k
•
46
•
1
rulins/DeepSeek-R1-Distill-Qwen-32B_NUMINA_train_amc_aime
Viewer
•
Updated
•
3.64k
•
799
•
2
rulins/pes2o_v3
Viewer
•
Updated
•
150M
•
182
rulins/MasssiveDS-1.4T-raw-data
Viewer
•
Updated
•
514M
•
299
rulins/MassiveDS-1.4T-raw-data
Viewer
•
Updated
•
514M
•
572
•
6
rulins/mmlu_searched_results_from_massiveds
Viewer
•
Updated
•
33.5k
•
329