Team Ai
20 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01dvislobokov /go-ml-complation Go full-line completion dataset (go/types) Caret-based full-line completion samples extracted from permissively licensed Go repositories with the goflc builder (Go go/parser + go/types). Each sample is an exact editor position: left_context ends at the caret, target_text is the rest of the physical line (no newline, trailing whitespace excluded), right_context follows it. Original source is reconstructable from offsets (byte offsets as used by Go tooling, plus UTF-16 offsets for… See the full description on the dataset page: https://huggingface.co/datasets/dvislobokov/go-ml-complation.tabulartext-generation100M<n<1B0 likes486 downloads5h agoHugging Face02dmariaa70 /GO-MO GO-MO: A large-scale graph-augmented traffic dataset for data-driven spatio-temporal traffic analysis This is the official dataset repository for the GO-MO traffic dataset. The GO-MO dataset is a traffic dataset extracted from the publicly available Open Data Portal of the City Council of Madrid (Spain). GO-MO comprises more than 1.5 billion records of three traffic-related metrics together with spatio-temporal data and metadata, spanning a ten-year period (2015-2024).… See the full description on the dataset page: https://huggingface.co/datasets/dmariaa70/GO-MO.tabulartime-series-forecasting1B<n<10B0 likes167 downloads3mo agoHugging Face03mencosk /gomodel-go-expert-v4 GoModel Go Expert v4 Dataset Description A high-quality dataset for fine-tuning Qwen2.5-Coder-7B to be an expert Go software engineer with tool-calling capabilities. This is version 4, substantially rebuilt from v3 with: Structured messages format (not pre-rendered ChatML text) Go AST-extracted code from real repositories using go/parser Go 1.26 feature coverage (February 2026 release) Senior/staff-level engineering content (architecture, distributed systems, API… See the full description on the dataset page: https://huggingface.co/datasets/mencosk/gomodel-go-expert-v4.tabulartext-generation10K<n<100K0 likes70 downloads2mo agoHugging Face04BubuDavid /Selena-Gomez-With-Lyrics-And-Spotify-Audio-Featurestabularn<1K0 likes66 downloads3y agoHugging Face05mencosk /gomodel-go-expert-v5tabular10K<n<100K0 likes57 downloads2mo agoHugging Face06double-blind-anonymous /go-mo-dataset GO-MO, a massive Graph agumented Open urban MObility dataset This is the official dataset repository for the GO-MO traffic dataset. The GO-MO dataset is a traffic dataset extracted from the publicly available Open Data Portal of the City Council of Madrid (Spain). GO-MO comprises more than 1.5 billion records of three traffic-related metrics together with spatio-temporal data and metadata, spanning a ten-year period (2015-2024). Additionally, the GO-MO dataset introduces two graph… See the full description on the dataset page: https://huggingface.co/datasets/double-blind-anonymous/go-mo-dataset.tabulartime-series-forecasting1B<n<10B0 likes48 downloads9mo agoHugging Face07mencosk /gomodel-go-expert-v6tabular10K<n<100K0 likes39 downloads2mo agoHugging Face08Gomly /spotify_audio_features Spotify Tracks & Audio Features Dataset Overview This dataset contains a comprehensive collection of Spotify tracks, combining rich audio feature analysis with track metadata. It is formatted as a high-performance Parquet dataset (ZStandard compressed), optimized for large-scale tabular analysis, machine learning, and recommender system research. Data Source The raw data for this dataset was originally gathered and hosted by Anna's Archive. Original Blog Post:… See the full description on the dataset page: https://huggingface.co/datasets/Gomly/spotify_audio_features.tabulartabular-regression100M<n<1B0 likes26 downloads7mo agoHugging Face09mencosk /gomodel-go-expert-v7tabular10K<n<100K0 likes24 downloads2mo agoHugging Face10gomipapa /record-test1This dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0", "robot_type": "so101_follower", "total_episodes": 4, "total_frames": 2990, "total_tasks": 1, "chunks_size": 1000, "data_files_size_in_mb": 100, "video_files_size_in_mb": 200, "fps": 30, "splits": { "train": "0:4" }, "data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/gomipapa/record-test1.tabularrobotics1K<n<10K0 likes19 downloads9mo agoHugging Face11vietnguyen28 /lekiwi_gomaxtabular10K<n<100K0 likes18 downloads10mo agoHugging Face12mencosk /gomodel-go-expert-v3 GoModel Go Expert v3 Dataset description GoModel Go Expert v3 is an English instruction and completion dataset for training Go coding assistants. It combines curated production Go, code-specific synthetic tasks, and agentic tool trajectories. Every JSONL record contains a full Qwen2.5-compatible ChatML conversation in its text field. Key changes from v2 Tool calls now use Qwen2.5's native <tool_call> tags instead of bare JSON. Tool definitions use… See the full description on the dataset page: https://huggingface.co/datasets/mencosk/gomodel-go-expert-v3.tabulartext-generation1K<n<10K0 likes17 downloads2mo agoHugging Face13vietnguyen28 /lekiwi_gomax2tabular10K<n<100K0 likes15 downloads10mo agoHugging Face14electricsheepafrica /africa-senegal-production-de-gombo-dbc0117c Production De Gombo | Africa (DHORT) 1 rows - 1 Africa country/area - detected - source table - Engineered by Electric Sheep Africa TL;DR This dataset contains 1 rows from DHORT, covering Production De Gombo. It is published as ML-ready Parquet with consistent Hugging Face metadata, source provenance, and analysis-friendly loading examples. What This Dataset Measures Official statistics datasets help analysts inspect public data as published by… See the full description on the dataset page: https://huggingface.co/datasets/electricsheepafrica/africa-senegal-production-de-gombo-dbc0117c.tabulartabular-classificationn<1K0 likes13 downloads2mo agoHugging Face15gomi5353 /NYCU_Cup_Stackingtabular10K<n<100K0 likes10 downloads3mo agoHugging Face16Naomiihao /dataset_v3_synth_top50custom_selena_gomez_middle_2tabular10K<n<100K0 likes8 downloads5mo agoHugging Face17Naomiihao /dataset_v3_synth_top50custom_selena_gomez_left_2tabular10K<n<100K0 likes8 downloads5mo agoHugging Face18mencosk /gomodel-go-expert-v2 GoModel Go Expert v2 Dataset description GoModel Go Expert v2 is an English instruction and completion dataset for training Go coding assistants. It combines curated production Go with synthetic instruction tasks and agentic tool trajectories. Version 2 is a new dataset and does not replace the earlier GoModel repositories. Each JSONL record is already serialized as a full Qwen-compatible ChatML conversation in its text field. Data sources Source… See the full description on the dataset page: https://huggingface.co/datasets/mencosk/gomodel-go-expert-v2.tabulartext-generation1K<n<10K0 likes8 downloads2mo agoHugging Face19electricsheepafrica /africa-senegal-rendement-gombo-ab251afb Rendement Gombo | Africa (DHORT) 1 rows - 1 Africa country/area - detected - source table - Engineered by Electric Sheep Africa TL;DR This dataset contains 1 rows from DHORT, covering Rendement Gombo. It is published as ML-ready Parquet with consistent Hugging Face metadata, source provenance, and analysis-friendly loading examples. What This Dataset Measures Official statistics datasets help analysts inspect public data as published by… See the full description on the dataset page: https://huggingface.co/datasets/electricsheepafrica/africa-senegal-rendement-gombo-ab251afb.tabulartabular-classificationn<1K0 likes6 downloads2mo agoHugging Face20Naomiihao /dataset_v3_synth_top50custom_selena_gomez_right_2tabular10K<n<100K0 likes4 downloads5mo agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.