Team Ai
22 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01renjiepi /easy_5000_data_processingtext10K<n<100K0 likes115 downloads9mo agoHugging Face02renjiepi /medium_5000_data_processingtext1K<n<10K0 likes84 downloads9mo agoHugging Face03laion /nemotron-terminal-data_processing nemotron-terminal-data_processing Per-source partition of nvidia/Nemotron-Terminal-Corpus, filtered to source == "data_processing". The difficulty column preserves the original easy / medium / mixed split (na for the dataset_adapters/* files, which did not carry a difficulty label). Partitioning scheme: adapters_{code,math,swe} — rows from dataset_adapters/{code,math,swe}.parquet {skill} (e.g. debugging, security, …) — rows from synthetic_tasks/skill_based/{easy,medium… See the full description on the dataset page: https://huggingface.co/datasets/laion/nemotron-terminal-data_processing.textquestion-answering1K<n<10K0 likes70 downloads6mo agoHugging Face04renjiepi /medium_5000-data_processing_n100k1text1K<n<10K0 likes61 downloads9mo agoHugging Face05renjiepi /medium_5000_data_processing_fixedtext1K<n<10K0 likes60 downloads9mo agoHugging Face06renjiepi /easy_5000-data_processing_n100k1text1K<n<10K0 likes51 downloads9mo agoHugging Face07renjiepi /mixed_1000_data_processingtextn<1K0 likes44 downloads9mo agoHugging Face08renjiepi /easy_5000_data_processing_fixedtext1K<n<10K0 likes39 downloads9mo agoHugging Face09renjiepi /mixed_1000-data_processing_1000_n30k1textn<1K0 likes37 downloads9mo agoHugging Face10DCAgent2 /terminal_bench_2_nemotron_terminal_data_processing__Qwen3_8B_20260413_170737textn<1K0 likes37 downloads6mo agoHugging Face11renjiepi /mixed_1000_data_processing_fixedtextn<1K0 likes28 downloads9mo agoHugging Face12Nexdata-kr /1.5-million-Korean-Test-Questions-Structured-Analysis-Processing-Data-Sample Description 한국어 시험 문제 구조화 분석·가공 데이터로, 약 150만 개의 시험 문제를 포함하고 있습니다. 문제 유형, 문제, 정답, 해설 등의 정보를 포함하며, 과목은 [초등학교] 국어, 수학, 영어, 사회, 과학; [중학교] 국어, 영어, 수학, 과학, 사회; [고등학교] 국어, 영어, 수학, 물리, 화학, 생물, 역사, 지리로 구성되어 있습니다. 문제 유형에는 객관식, 빈칸 채우기, 참·거짓 문제, 단답형 문제 등이 포함됩니다. 본 데이터셋은 대규모 교과 지식 강화 및 학습 데이터 구축 등의 작업에 활용할 수 있습니다. 자세한 내용은 아래 링크를 참고해 주세요: https://ko.nexdata.ai/datasets/llm/1634?source=Hf.kr Specifications Data content 한국어 K12 시험 문제 Amount 약 150만 개의… See the full description on the dataset page: https://huggingface.co/datasets/Nexdata-kr/1.5-million-Korean-Test-Questions-Structured-Analysis-Processing-Data-Sample.textn<1K0 likes27 downloads1mo agoHugging Face13Nexdata-AI /1.5-million-Korean-Test-Questions-Structured-Analysis-Processing-Data-Sample Description Korean Test Questions Structured Analysis Processing Data, around 1.5 million questions, contains question types, questions, answers, explanations, etc..For subjects, include [Primary School] Korean, Mathematics, English, Social Studies, Science; [Middle School] Korean, English, Mathematics, Science, Social Studies; [High School] Korean, English, Mathematics, Physics, Chemistry, Biology, History, Geography; question Types indlude single-choice question, fill-in… See the full description on the dataset page: https://huggingface.co/datasets/Nexdata-AI/1.5-million-Korean-Test-Questions-Structured-Analysis-Processing-Data-Sample.textn<1K0 likes23 downloads2mo agoHugging Face14DCAgent2 /swebench_verified_random_100_folders_nemotron_terminal_data_processing__Qwen3_87dd0272etextn<1K0 likes17 downloads6mo agoHugging Face15nahiar /sentiment_data_train_id_en_sentiment_30k_post-processingtext10K<n<100K0 likes14 downloads8mo agoHugging Face16DCAgent2 /dev_set_v2_nemotron_terminal_data_processing__Qwen3_8B_20260413_175806textn<1K0 likes11 downloads6mo agoHugging Face17nahiar /sentiment_30k_data_train_sentimen_id_post-processingtext10K<n<100K0 likes8 downloads8mo agoHugging Face18Nexdata-AI /10-million-English-Test-Questions-Text-Parsing-And-Processing-Data-Sample Description 10 Million - English Test Questions Text Parsing And Processing Data, Each question contains title, answer, parse, subject, grade, question type; The educational stages cover primary, middle, high school, and university; Subjects cover mathmatics, biology, accounting, etc.The data are questions text under the Anglo-American system, which can be used to enhance the subject knowledge of large models For more details, please refer to the link:… See the full description on the dataset page: https://huggingface.co/datasets/Nexdata-AI/10-million-English-Test-Questions-Text-Parsing-And-Processing-Data-Sample.textn<1K0 likes6 downloads2mo agoHugging Face19OsakanaTeishoku /structured_data_with_cot_dataset_512_v2_dpo_before_processingtext1K<n<10K0 likes5 downloads9mo agoHugging Face20nahiar /sentiment_data_train_sentimen_id_post-processingtext10K<n<100K0 likes5 downloads8mo agoHugging Face21nahiar /spam_post-processing-data-train-final-25k-id-entext10K<n<100K0 likes4 downloads8mo agoHugging Face22infra777 /my-data-processingtext1K<n<10K0 likes4 downloads3mo agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.