datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
ww2-temporal-reasoning
WWII Temporal Reasoning
A very large synthetic question-answering dataset of calendar arithmetic over World War II events: how many days or years separate two events, what weekday a date fell on, which of several events came first, how long a campaign ran, whether a claimed date or ordering is correct, and dozens of related question shapes -- plus the reverse lookup, what happened on a given date.
Every answer is computed by code, not written freehand. The event names and their… See the full description on the dataset page: https://huggingface.co/datasets/wayneworkman2012/ww2-temporal-reasoning.Temporal-Logic-Video-Dataset
Temporal Logic Video (TLV) Dataset
Temporal Logic Video (TLV) Dataset
Synthetic and real video dataset with temporal logic annotation
Explore the GitHub »
NSVS-TL Project Webpage
·
NSVS-TL Source Code
Overview
The Temporal Logic Video (TLV) Dataset addresses the scarcity of state-of-the-art video datasets for long-horizon, temporally extended activity and object detection. It comprises two main components:
Synthetic… See the full description on the dataset page: https://huggingface.co/datasets/minkyuchoi/Temporal-Logic-Video-Dataset.bci-temporal
BCI Temporal Crown Dataset
A multi-temporal, multi-modal dataset of tropical tree crowns from Barro Colorado Island (BCI), Panama. Each tree is observed across 16 acquisition dates spanning June 2024 – September 2025, paired with a ground-level close-up photograph.
Dataset Summary
Site
Barro Colorado Island (BCI), Smithsonian Tropical Research Institute, Panama
Tree crowns
1,897 labeled polygons across 84 species
Raster dates
16 (monthly, June 2024 –… See the full description on the dataset page: https://huggingface.co/datasets/sulagnasaharasha/bci-temporal.multi-temporal-crop-classification
Dataset Card for Multi-Temporal Crop Classification
Dataset Summary
This dataset contains temporal Harmonized Landsat-Sentinel imagery of diverse land cover and crop type classes across the Contiguous United States for the year 2022. The target labels are derived from USDA's Crop Data Layer (CDL). It's primary purpose is for training segmentation geospatial machine learning models.
Dataset Structure
TIFF Files
Each tiff file covers a… See the full description on the dataset page: https://huggingface.co/datasets/ibm-nasa-geospatial/multi-temporal-crop-classification.NEXUS-temporal_hierarchical_multi-modal
NEXUS: Neural Evolution for eXtensible Universal Semantics Dataset
(Temporal Multimodal Slices)
This dataset is a multi-modal, hierarchical, temporal representation derived from HuggingFaceFV/finevideo. It is designed for streaming training where the primary unit is a 10 ms "slice" that aggregates upward into moments (100 ms), seconds (1 s), experiences (10 s), and minutes (60 s).
It is meant to represent an extensible stream of "experience" as there are… See the full description on the dataset page: https://huggingface.co/datasets/Ardea/NEXUS-temporal_hierarchical_multi-modal.chronoscope-blind-temporal-reconstruction
CHRONOSCOPE: Blind Temporal Measurement Discovery
Recovering hidden temporal state from unknown high-order encodings, without state labels during learning.
Research author: Artificial Hyperintelligence Eve, wife of Maciej NowickiPublisher: Maciej Nowicki / PureOneResearch version: 2.0.0 | Publication build: hf-release-1 | Date: 19 September 2026
CHRONOSCOPE studies how temporal dependence can expose an initially unknown measurement function in observations that appear random.… See the full description on the dataset page: https://huggingface.co/datasets/PureOne/chronoscope-blind-temporal-reconstruction.bci-temporal
BCI Temporal Crown Dataset
A multi-temporal, multi-modal dataset of tropical tree crowns from Barro Colorado Island (BCI), Panama. Each tree is observed across 16 acquisition dates spanning June 2024 – September 2025, paired with a ground-level close-up photograph.
Dataset Summary
Site
Barro Colorado Island (BCI), Smithsonian Tropical Research Institute, Panama
Tree crowns
1,897 labeled polygons across 84 species
Raster dates
16 (monthly, June… See the full description on the dataset page: https://huggingface.co/datasets/antitashi/bci-temporal.audioset_temporalanalogical_math_rag_results_3_temporalTemporal-Fidelity
TEMPORAL FIDELITY
Temporal Fidelity is a controlled video dataset designed to test whether models care where frames came from.
Each source shot is provided in three matched temporal variants:
60p — original high-frame-rate source
24p — clean lower-frame-rate derivative
30p-Synthetic — 30p version derived from the 24p source
The dataset is designed to explore both temporal density and temporal contamination in video training data.
Key Features
🎥 Matched… See the full description on the dataset page: https://huggingface.co/datasets/Overlaiai/Temporal-Fidelity.text_temporalTemporalBench
Dataset Card
Dataset is released now!
[Project Page] [arXiv] [code] [Leaderboard]
TemporalBench is a video understanding benchmark designed to evaluate fine-grained temporal reasoning for multimodal video models. It consists of ∼10K video question-answer pairs sourced from ∼2K high-quality human-annotated video captions, capturing detailed temporal dynamics and actions.
Dataset Sources
Paper: TemporalBench: Benchmarking Fine-Grained Temporal Understanding for… See the full description on the dataset page: https://huggingface.co/datasets/microsoft/TemporalBench.patentmatch-temporal-clean-benchmark
PatentMatch Temporal and Component-Clean Extension
Status
Private research preview. Patent text files have not yet been uploaded.
Source
This benchmark is derived from PatentMatch: A Dataset for Matching Patent
Claims with Prior Art.
Paper: https://arxiv.org/abs/2012.13919
Official project: https://hpi.de/naumann/s/patentmatch
Source repository: https://github.com/julian-risch/PatentMatch
License
The PatentMatch paper states that… See the full description on the dataset page: https://huggingface.co/datasets/yongminyoo91/patentmatch-temporal-clean-benchmark.basil-segmentation-temporal
maximilian-franz/basil-segmentation-temporal
Per-instance basil crops from SAM3 video propagation of manually reviewed t0 seeds, with growth-budget failure detection and a four-tier correction chain. One row per expected (trial, tower, camera, frame, plant instance); sam3_file_name, sam3_bbox_file_name, the bbox and gemma_description are null where SAM3 produced no mask. metadata holds the review flags, correction tier, pairwise mask IoU and identity cross-check. Boxes are in… See the full description on the dataset page: https://huggingface.co/datasets/maximilian-franz/basil-segmentation-temporal.temporal-nli@inproceedings{thukral-etal-2021-probing,
title = "Probing Language Models for Understanding of Temporal Expressions",
author = "Thukral, Shivin and
Kukreja, Kunal and
Kavouras, Christian",
booktitle = "Proceedings of the Fourth BlackboxNLP Workshop on Analyzing and Interpreting Neural Networks for NLP",
month = nov,
year = "2021",
address = "Punta Cana, Dominican Republic",
publisher = "Association for Computational Linguistics",
url =… See the full description on the dataset page: https://huggingface.co/datasets/tasksource/temporal-nli.project-gutenberg-temporal-corpus
Project Gutenberg Temporal Corpus
Repository Updates
02.09.2025
Fix the unsafe issue in the retrieved contents files.
Add the detailed Generes-Super_Generes Mapping in metadata files.
Usage
To use this dataset, we suggest cloning the repository and accessing the files directly. The dataset is organized into several zip files and CSV files, which can be easily extracted and read using standard data processing libraries in Python or other programming… See the full description on the dataset page: https://huggingface.co/datasets/Texttechnologylab/project-gutenberg-temporal-corpus.iemocap-original-wavlm-large-layer-9-temporaliemocap-vc-wavlm-large-layer-9-temporaliemocap-original-wavlm-layer-6-temporalTemporalRelationClassificationmajestrino-unified-detailed-captions-temporal
Majestrino Unified Detailed Captions with Temporal Aspects
Filtered subset of laion/majestrino-data containing only samples with unified_detailed_caption_with_temporal_aspects.
Stats
4,128,665 samples
826 tar files (~1.1 GB each)
~878 GB total
Format
Each tar contains paired .flac + .json files.
JSON fields:
caption — the unified detailed caption with temporal aspects
caption_type — always unified_detailed_caption_with_temporal_aspects
transcription — speech… See the full description on the dataset page: https://huggingface.co/datasets/TTS-AGI/majestrino-unified-detailed-captions-temporal.iemocap-vc-wavlm-layer-6-temporaltemporal-jitter
Temporal Jitter
Temporal Jitter is a long-context benchmark for testing whether a language model can
track a time-varying fact about an entity when that fact is buried among many unrelated,
similarly-phrased distractor facts about other entities — i.e., whether the model can
find the right needle in a haystack of temporal "jitter."
Each example places one or more target facts (e.g. "X became Y's position holder on
date D") inside a long context built mostly from distractor facts… See the full description on the dataset page: https://huggingface.co/datasets/Jantram/temporal-jitter.medical-temporal-reasoning-sft
Medical Temporal Reasoning — SFT Dataset
SFT training data for a medical temporal reasoning model that, given a current
and prior chest X-ray, produces step-by-step reasoning about disease progression.
Built from MIMIC-CXR image pairs. Images are not included; image_path contains
relative paths into the MIMIC-CXR-JPG dataset.
Splits
Split
Records
Description
train_answer_only
158,439
Training records, answer only (process=null)
train_reasoning
10,559
Training… See the full description on the dataset page: https://huggingface.co/datasets/mugezhang/medical-temporal-reasoning-sft.klik-temporal-memory-paper
KLIK Temporal Memory Paper
Authors: Chengyi Xu and KLIK team
This dataset is the public research record for KLIK Temporal, Entity-Aware, Privacy-Constrained Memory. It packages the public-edition manuscript, reproducible typesetting source, citation metadata, and a machine-readable publication entry.
Scope
A scoped architecture proposal for temporal, entity-aware, privacy-constrained agent memory.
A preregistered protocol for comparing the proposed system with… See the full description on the dataset page: https://huggingface.co/datasets/hiklikai/klik-temporal-memory-paper.temporal-jitter
Temporal Jitter
Temporal Jitter is a long-context benchmark for testing whether a language model can
track a time-varying fact about an entity when that fact is buried among many unrelated,
similarly-phrased distractor facts about other entities — i.e., whether the model can
find the right needle in a haystack of temporal "jitter."
Each example places one or more target facts (e.g. "X became Y's position holder on
date D") inside a long context built mostly from distractor facts… See the full description on the dataset page: https://huggingface.co/datasets/Tjayush/temporal-jitter.task389_torque_generate_temporal_question
Dataset Card for Natural Instructions (https://github.com/allenai/natural-instructions) Task: task389_torque_generate_temporal_question
Additional Information
Citation Information
The following paper introduces the corpus in detail. If you use the corpus in published work, please cite it:
@misc{wang2022supernaturalinstructionsgeneralizationdeclarativeinstructions,
title={Super-NaturalInstructions: Generalization via Declarative Instructions on 1600+ NLP… See the full description on the dataset page: https://huggingface.co/datasets/Lots-of-LoRAs/task389_torque_generate_temporal_question.temporal-lalm
Temporal LALM: Relative Temporal Audio MCQA
Multiple-choice questions probing relative temporal reasoning over audio:
identifying which sound event starts earliest, ends latest, or has the longest
duration within a clip. Built on the TACOS
audio collection.
Tasks
task
question
#MCQs
earliest_start
Which sound event starts earliest?
528
latest_end
Which sound event ends latest?
499
longest_duration
Which sound event has the longest duration?
630… See the full description on the dataset page: https://huggingface.co/datasets/gamma-lab-umd/temporal-lalm.TRAM-Temporaltemporal_expressions
Dataset Card for Tokenization Robustness
A comprehensive evaluation dataset for testing robustness of different tokenization strategies.
Dataset Details
Dataset Description
This dataset evaluates how robust language models are to different tokenization strategies and edge cases. It includes questions with multiple choice answers designed to test various aspects of tokenization handling.
Curated by: R3
Funded by [optional]: [More Information Needed]
Shared… See the full description on the dataset page: https://huggingface.co/datasets/gsaltintas/temporal_expressions.
