RM
Models
All models matching “RM”Datasets
All datasets matching “RM”RMISC
RMISC: A Large-scale Real-world Multivariate Corpus for Time Series Foundation Models
This dataset card describes the main branch of RMISC. RMISC is a large-scale, real-world multivariate time-series corpus for pretraining and benchmarking time series foundation models (TSFMs). The complete corpus contains around 200 sub-datasets, 2 million original time-series files, 16 billion timesteps, and 142 billion time points across energy, finance, environment, industry, traffic, and… See the full description on the dataset page: https://huggingface.co/datasets/nju-zhangsq/RMISC.RMBenchbabilong
BABILong (100 samples) : a long-context needle-in-a-haystack benchmark for LLMs
Preprint is on arXiv and code for LLM evaluation is available on GitHub.
BABILong Leaderboard with top-performing long-context models.
bAbI + Books = BABILong
BABILong is a novel generative benchmark for evaluating the performance of NLP models in
processing arbitrarily long documents with distributed facts.
It contains 11 configs, corresponding to different sequence lengths in tokens:… See the full description on the dataset page: https://huggingface.co/datasets/RMT-team/babilong.babilong-1k-samples
BABILong (1000 samples) : a long-context needle-in-a-haystack benchmark for LLMs
Preprint is on arXiv and code for LLM evaluation is available on GitHub.
BABILong Leaderboard with top-performing long-context models.
bAbI + Books = BABILong
BABILong is a novel generative benchmark for evaluating the performance of NLP models in
processing arbitrarily long documents with distributed facts.
It contains 9 configs, corresponding to different sequence lengths in tokens: 0k… See the full description on the dataset page: https://huggingface.co/datasets/RMT-team/babilong-1k-samples.RoG-cwq
Dataset Card for "RoG-cwq"
More Information needed
RoG-webqsp
Dataset Card for "RoG-webqsp"
More Information needed
