Team Ai
15 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01THULab /tutorial-ball-2 tutorial-ball-2 (LeRobot) — TsFile This dataset is a lossless conversion to the Apache TsFile format of the HuggingFace LeRobot dataset notmahi/tutorial-ball-2: a low-dimensional robot tutorial trajectory dataset (no video). Original dataset Source dataset: notmahi/tutorial-ball-2 Format: early LeRobot format (meta_data/ + safetensors) Content: purely numeric low-dimensional state/action trajectories — 314,074 frames / 751 episodes / 30 fps. No images or video… See the full description on the dataset page: https://huggingface.co/datasets/THULab/tutorial-ball-2.tabulartime-series-forecastingn<1K0 likes116 downloads2mo agoHugging Face02miyuki2026 /tutorialstext100K<n<1M0 likes68 downloads8mo agoHugging Face03code-rag-bench /online-tutorialsThe online tutorials retrieval source for code-rag-bench, consisting tutorials pages collected from GeeksforGeeks, W3Schools, tutorialspoint, and Towards Data Science. text10K<n<100K1 likes43 downloads2y agoHugging Face04neurodeskorg /qsiprep-tutorial-data QSIPrep derivatives for the QSIRecon tutorial (NeurodeskEDU) Input data for the NeurodeskEDU notebook examples/diffusion_imaging/qsirecon.ipynb. This is a mirror of OSF project q7v8c. It was copied on 2026-10-07 because OSF rate-limits (HTTP 429) the downloads made by the notebooks' automated test runs. Only the QSIPrep output archive the notebook uses is mirrored, at the same path as on OSF. manifest.json lists each file's size and SHA-256 hash. The hashes match the ones OSF… See the full description on the dataset page: https://huggingface.co/datasets/neurodeskorg/qsiprep-tutorial-data.textn<1K0 likes25 downloads5h agoHugging Face05iapp /dpo_thai_tutorial Thai DPO Tutorial Dataset (ชุดข้อมูล DPO ภาษาไทย) ข้อมูลสำหรับการเรียนรู้ Direct Preference Optimization (DPO) ภาษาไทย Description Dataset นี้สร้างขึ้นเพื่อการศึกษาและสาธิตเทคนิค DPO สำหรับการจัดแนว LLM 100 ตัวอย่าง จากข้อมูลจริง คำถามภาษาไทยหลากหลายหัวข้อ (การเงิน, เศรษฐกิจ, ความรู้ทั่วไป) Chosen: คำตอบภาษาไทยที่มีคุณภาพ มีการคิดวิเคราะห์ Rejected: คำตอบภาษาอังกฤษหรือคำตอบที่ไม่เหมาะสม Data Format { "instruction": "คำถามหรือคำสั่งภาษาไทย", "input":… See the full description on the dataset page: https://huggingface.co/datasets/iapp/dpo_thai_tutorial.textn<1K0 likes24 downloads8mo agoHugging Face06neurodeskorg /mrtrix-tutorial-data MRtrix diffusion MRI tutorial data Intermediate results used by the three-part NeurodeskEDU MRtrix series (examples/diffusion_imaging/MRtrix_1.ipynb to MRtrix_3.ipynb), so that each part can start without re-running the previous one. This is a mirror of the OSF project y2dq4. It was copied on 2026-10-07 because OSF rate-limits (HTTP 429) the downloads made by the notebooks' automated test runs. Folder Contents Downloaded by preprocessed/ Part 1 outputs (denoised… See the full description on the dataset page: https://huggingface.co/datasets/neurodeskorg/mrtrix-tutorial-data.textn<1K0 likes24 downloads10h agoHugging Face07GoldenGrapeGentleman1 /battle-game-grpo-tutorial turn-based battle game GRPO tutorial dataset Pre-built GRPO records for the ROCm AI Developer Hub tutorial. Split File Records demo data/demo.jsonl 64 train data/train.jsonl 2048 validate data/validate.jsonl 32 Use via tutorial notebook Step 12 (load_grpo_tutorial_records) or regenerate with prepare_grpo_tutorial_data.py. Companion scripts: https://github.com/GoldenGrapeGentleman/battle game-showdown-agent-scripts textreinforcement-learning1K<n<10K1 likes22 downloads2mo agoHugging Face08marcduda /langchain_tutorialtext1K<n<10K0 likes20 downloads3y agoHugging Face09GoldenGrapeGentleman1 /pokemon-showdown-grpo-tutorial Pokémon Showdown GRPO tutorial dataset Pre-built GRPO records for the ROCm AI Developer Hub tutorial. Split File Records demo data/demo.jsonl 64 train data/train.jsonl 2048 validate data/validate.jsonl 32 Use via tutorial notebook Step 12 (load_grpo_tutorial_records) or regenerate with prepare_grpo_tutorial_data.py. Companion scripts: https://github.com/GoldenGrapeGentleman/pokemon-showdown-agent-scripts textreinforcement-learning1K<n<10K0 likes19 downloads2mo agoHugging Face10rezashamji /biology-tutorial-feb19textn<1K0 likes13 downloads8mo agoHugging Face11marin-dna /zoonomia-v1-v3_ccre_non_promoter-tutorial Tiny Zoonomia non-promoter cCRE tutorial sample This dataset contains 256 unchanged rows from the public marin-dna/zoonomia-v1-v3_ccre_non_promoter cross-mammal enhancer dataset. It exists only to keep MarinDNA's local CPU training tutorial fast; it is not an independent biological dataset or a representative benchmark. Contents data/train.jsonl contains the first 256 records from source shard data/train/shard_0000.jsonl.zst at immutable source revision… See the full description on the dataset page: https://huggingface.co/datasets/marin-dna/zoonomia-v1-v3_ccre_non_promoter-tutorial.tabularn<1K0 likes9 downloads2mo agoHugging Face12rezashamji /cardiology-tutorialtextn<1K0 likes7 downloads8mo agoHugging Face13shzhang /tutorial_datasets_github_issuestabular1K<n<10K0 likes6 downloads4y agoHugging Face14camcalderon777 /ec3014-tutorialstextn<1K0 likes5 downloads4mo agoHugging Face15ChavyvAkvar /example-tutorialtextn<1K0 likes2 downloads3y agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.