datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
music-tempo-and-key-reference-data
BPM Interval and Camelot Reference Data
This public dataset contains two reference assets maintained by j022315051:
bpm-interval-reference.csv: BPM values with milliseconds per beat and the calculation used.
camelot.js: major and minor pitch-class mappings to Camelot codes.
These reference assets support the browser-based music tools published at Tap BPM Now.
Scope and boundaries
The release contains reference tables and mapping code only. It does not contain… See the full description on the dataset page: https://huggingface.co/datasets/j022315051/music-tempo-and-key-reference-data.paymind-reference-data
PayMind Reference Dataset
Synthetic/reference payment-routing data for PayMind, an open-source payment route intelligence engine.
This dataset is designed to demonstrate PayMind's training, evaluation, and routing workflow across route selection, transaction reliability, and expected settlement time.
Important: This dataset contains synthetic/reference data only. It does not contain real customers, real transactions, payment credentials, personally identifiable information, or… See the full description on the dataset page: https://huggingface.co/datasets/navk8690/paymind-reference-data.developer-reference-datasets
Developer Reference Datasets
Open, reproducible lookup tables that web and app developers reach for constantly — computed from first principles, not scraped, so every value is exact and re-runnable. CC BY 4.0.
Quick answers (straight from the data)
What is 16:9 in pixels? 1920×1080, 1280×720, 3840×2160. 9:16 (Stories, Reels, TikTok) is those flipped. → aspect-ratios, resolutions
What contrast ratio does WCAG require? 4.5:1 for normal text (AA), 3:1 for large… See the full description on the dataset page: https://huggingface.co/datasets/cleanorlabs/developer-reference-datasets.reference-free-rl-summarization-data
Reference-free RL Summarization Experimental Data
This repository contains experimental data splits, metadata, and processed subsets used for a study on verifier-composable penalty-shaped reinforcement learning for reference-free summarization.
Configs
vnexpress: Vietnamese VnExpress train/validation/test split used in the study. Unless explicit redistribution permission is available, this config releases metadata and split information only.
cnn_dailymail_subset:… See the full description on the dataset page: https://huggingface.co/datasets/phuongntc/reference-free-rl-summarization-data.reference-datasetThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"robot_type": "so_follower",
"total_episodes": 3,
"total_frames": 597,
"total_tasks": 1,
"chunks_size": 1000,
"data_files_size_in_mb": 100,
"video_files_size_in_mb": 200,
"fps": 30,
"splits": {
"train": "0:3"
},
"data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/rhecker/reference-dataset.paymind-reference-data-v2
PayMind Synthetic Payment Dataset — V4
Synthetic payment-routing data for developing, training and benchmarking PayMind.
This dataset provides the V4 synthetic training environment for PayMind, an open-source payment intelligence connector.
It is designed for three predictive responsibilities:
Engine
Objective
Candidate Generator
Learn which payment routes fit a transaction
Reliability Engine
Estimate transaction success probability
Settlement Intelligence… See the full description on the dataset page: https://huggingface.co/datasets/navk8690/paymind-reference-data-v2.
