datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
hnm-fashion-recommendations-data
Dataset Rekomendasi Fashion H&M
Dataset ini berisi data transaksi, atribut pelanggan, dan metadata produk yang telah dianonimkan dari H&M Group. Kumpulan data komprehensif ini memungkinkan pemodelan perilaku pembelian pelanggan secara mendalam.
Wawasan yang dihasilkan dapat dimanfaatkan untuk berbagai tujuan bisnis yang strategis, mulai dari meningkatkan personalisasi pengalaman berbelanja, mengoptimalkan manajemen inventaris untuk efisiensi produksi, hingga mendukung inisiatif… See the full description on the dataset page: https://huggingface.co/datasets/einrafh/hnm-fashion-recommendations-data.aksahaha_crop-recommendation
crop recommendation
Crop Growth Recommendations: Optimal Conditions for Higher Yields
Dataset Info
Source: Kaggle
Original Size: 0.06 MB
Kaggle Downloads: 4,065
Files: 1
Files
Crop_recommendation.csv
Mirrored from Kaggle
Crop-recommendationCrop-Recommendation-Parameters
🌱 Crop Recommendation Dataset
A machine learning dataset for crop recommendation based on soil properties and environmental conditions. The dataset contains measurements of essential soil nutrients and climatic parameters, along with the crop label that is suitable for those conditions.
This dataset can be used for machine learning classification, agricultural analytics, decision-support systems, and smart farming applications.
📌 Dataset Overview
Property… See the full description on the dataset page: https://huggingface.co/datasets/Samarth-27/Crop-Recommendation-Parameters.HM-Personalized-Fashion-Recommendationschatgpt-software-recommendations
ChatGPT Software Recommendations: 21 Categories, 2,100 Answers
This dataset records which software brands ChatGPT names, recommends and picks when buyers ask about 21 software categories, and which websites it cites. It covers 2,100 ChatGPT answers (100 buyer questions in each of 21 categories), coded brand by brand, with Google's organic top 10 for the same questions as the control.
It is published by High Salience, an AI search and SEO agency. Every file here is also published… See the full description on the dataset page: https://huggingface.co/datasets/highsalience/chatgpt-software-recommendations.Movie-recommendation-datasynthesized-cloud-optimization-recommendations
Synthesized Cloud-Optimization Recommendations
18 scenarios that pair cloud telemetry with a hand-crafted optimization
recommendation. Use them to train models or to evaluate AI agents.
Summary
Each scenario has multi-tier telemetry, a Terraform file describing the
deployed infrastructure, and a gold-standard recommendation.
The dataset is built around a simple input-output mapping. The input is
telemetry plus the infrastructure. The output is an optimization… See the full description on the dataset page: https://huggingface.co/datasets/ameau01/synthesized-cloud-optimization-recommendations.myket-android-application-recommendation-dataset
Myket Android Application Install Dataset
This dataset contains information on application install interactions of users in the Myket android application market. The dataset was created for the purpose of evaluating interaction prediction models, requiring user and item identifiers along with timestamps of the interactions.
Data Creation
The dataset was initially generated by the Myket data team, and later cleaned and subsampled by Erfan Loghmani a master student at… See the full description on the dataset page: https://huggingface.co/datasets/erfanloghmani/myket-android-application-recommendation-dataset.JobCCC-Conversational-Job-Recommendation-Bangladesh
JobCCC: A Conversational Code-Mixed Corpus for Job Recommendation in Bangladesh
Dataset Creators
Authors: Md. Arman Hossain, Mubashir Jawad, Fariha Khandakar Moon, and Sonia Binte Siraj
Supervisor: Dr. Nafis Sadeq
Institution: Department of Computer Science & Engineering, East West University
Dataset Summary
JobCCC (Conversational Code-Mixed Corpus) is a multi-turn conversational benchmark and job recommendation dataset tailored for the… See the full description on the dataset page: https://huggingface.co/datasets/Armans33115/JobCCC-Conversational-Job-Recommendation-Bangladesh.E-Commerce-Product-RecommendationF1-driver-car-setup-coupling-optimization-recommendations-v0.1What this dataset tests
Whether a system can propose setup adjustmentsthat increase driver-car coupling resonance.
Focus
Setup candidatespredicted resonance gainstability trade-offcondition sensitivitypersonalized setup profile
Required outputs
setup adjustment candidates
predicted resonance gain
stability trade-off index
track condition sensitivity
personalized setup profile
All indices0 to 1
Higher gainmeans larger coupling improvement.
Constraints
Setup recommendations only.Do… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/F1-driver-car-setup-coupling-optimization-recommendations-v0.1.fitness-recommendation-datasetRec-Gaze-Click-Cursor-Eye-Tracking-Movie-Recommendation-Dataset-for-Carousel-Interfaces
RecGaze Dataset
This is the HuggingFace RecGaze dataset from the paper: 'RecGaze: The First Eye Tracking and User Interaction Dataset for Carousel Interfaces'.
Link to open-acess paper: SIGIR 2025
Dataset Description
The RecGaze dataset is the first comprehensive feedback dataset on carousels that includes eye tracking results, clicks, cursor movements, and selection explanations. The dataset comprises of interactions from 3 movie selection tasks with 40… See the full description on the dataset page: https://huggingface.co/datasets/santideleon/Rec-Gaze-Click-Cursor-Eye-Tracking-Movie-Recommendation-Dataset-for-Carousel-Interfaces.training_data_maximum_weight_recommendationfitness-recommendation-datasetai-recommendation-index
AI Recommendation Index
An open quarterly dataset tracking which software brands AI assistants mention and recommend first across fixed buying prompts.
Q2 2026
144 complete responses
6 AI models
24 fixed prompts
4 software categories
English and Finnish
Each record includes the complete prompt and response, collection timestamp, model identifiers, normalized brand mentions, first explicit recommendation, and returned citations.
Use cases
AI… See the full description on the dataset page: https://huggingface.co/datasets/nikoalho/ai-recommendation-index.clinical-quad-pk-sampling-sparse-data-model-misspecification-dose-recommendation-error-v0.1Clinical Quad PK Sampling Sparse Data Model Misspecification Dose Recommendation Error v0.1
Each row is a site monthly snapshot.
Core quad
PK sampling densitySparse dataModel misspecificationDose recommendation error
Target
label_decision_error_risk_next_90d
Files
data/train.csvdata/tester.csvscorer.py
Evaluation
Run model on data/tester.csvReturn predictions row alignedScore with scorer.py
License
MIT
This dataset identifies a measurable coupling pattern associated with systemic instability.… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/clinical-quad-pk-sampling-sparse-data-model-misspecification-dose-recommendation-error-v0.1.LGUplus_recommendation_candidates
LGUplus_recommendation_candidates
LG U+ 페르소나 기반 TV 추천 파이프라인의 1차 후보 목록(A∪B, combo당 최대 20개)과
후보 id를 프로그램 상세로 조인하기 위한 프로그램 테이블입니다.
최종 top5 선정(교사 LLM)은 이 후보에서 5개를 고르는 다음 단계이며, 여기엔 후보까지만 포함합니다.
구성
하루를 12개 2시간 블록으로 나눔. (persona_id, date, block) 마다 후보 2종:
cand_stat (후보A · 시청통계): 기반주에 그 페르소나가 가장 많이 본 채널 top10, 각 채널에서 그 블록에 가장 먼저 시작하는 프로그램 1개 (≤10).
cand_persona (후보B · 주간페르소나): 그 블록 프로그램을 주간페르소나 프로필과 ko-sroberta 임베딩 유사도 + 시청장르 비례배분으로 뽑은 10개.
두 후보는 겹칠 수 있으며 dedup 후 A∪B ≤… See the full description on the dataset page: https://huggingface.co/datasets/ENERZAiKR/LGUplus_recommendation_candidates.emotion-aware-music-recommendation-datasetsteam_recommendation_systemlegal-client-advice-fact-law-recommendation-coherence-v0.1What this dataset does
You receive
facts summary
legal test
application
risk range language
recommendation
deadlines and next steps
You decide
coherent
or
incoherent
Daily use
safe-to-send check
missing fact flag
wrong test flag
deadline and next step check
amazon-reviews-recommendationsystem_recommendation_dataset
Dataset Card for Dataset Name
This dataset card aims to be a base template for new datasets. It has been generated using this raw template.
Dataset Details
Dataset Description
Curated by: [More Information Needed]
Funded by [optional]: [More Information Needed]
Shared by [optional]: [More Information Needed]
Language(s) (NLP): [More Information Needed]
License: [More Information Needed]
Dataset Sources [optional]
Repository: [More… See the full description on the dataset page: https://huggingface.co/datasets/kolonam/system_recommendation_dataset.restaurant-recommendation-jpproduct-recommendation-dataeval_data_maximum_weight_recommendationRecommendation361asia-owid-flu-vaccines-older-people-recommendation
Flu Vaccines Older People Recommendation | Asia (Our World in Data)
🌏 207 observations · 42 Asia countries · 2014–2020 · Repackaged by Electric Sheep Asia
TL;DR
This dataset contains 207 observations of Flu Vaccines Older People Recommendation data across 42 Asia countries, spanning 2014–2020.
About the source
Source: Our World in Data
Publisher: Our World in Data
License: cc-by-4.0
Topic: Flu Vaccines Older People Recommendation… See the full description on the dataset page: https://huggingface.co/datasets/electricsheepasia/asia-owid-flu-vaccines-older-people-recommendation.recommendation_songs
