Team Ai
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01ashikshaffi08 /code_qa_10Ktext1K<n<10K2 likes66 downloads2y agoHugging Face02joshuasundance /python-code-instructions-85k-mypo-qaqc joshuasundance/python-code-instructions-85k-mypo QA/QC artifact This dataset repo is a QA/QC derivative generated by myponline. What is included Root-level train.parquet / validation.parquet / test.parquet with full QA/QC annotations. filtered_basic/ with rows that pass structural QA/QC checks. filtered_strict/ with rows whose chosen side passes structural QA/QC plus standalone ruff and mypy --strict. summary.json with aggregate counts and provenance.… See the full description on the dataset page: https://huggingface.co/datasets/joshuasundance/python-code-instructions-85k-mypo-qaqc.tabular10K<n<100K1 likes38 downloads6mo agoHugging Face03vm2825 /CodeQA-datasettext10K<n<100K1 likes34 downloads1y agoHugging Face04AI4Manufacturing /w_cad2-codeqagated Dataset Card Dataset Description [Placeholder] Dataset Structure [Placeholder] Uses [Placeholder] Limitations [Placeholder] License [Placeholder] Citation [Placeholder] image100K<n<1M0 likes27 downloads7d agoHugging Face05vm2825 /small_repos_multi_file_chatgpt_5_qas_part5_code_qa-datasettext10K<n<100K0 likes26 downloads1y agoHugging Face06lissadesu /codeqa_reduced Dataset Card for "codeqa_final" More Information needed tabular10K<n<100K0 likes23 downloads3y agoHugging Face07lissadesu /code_qa_updatedtabular10K<n<100K0 likes22 downloads3y agoHugging Face08ostapeno /opc-annealing-corpus-synth-qa-code_python_js_tstext1M<n<10M0 likes22 downloads2y agoHugging Face09reasoning-degeneration-dev /rlm-codeqa-gpt-5-nano-20260226-001320 rlm-codeqa-gpt-5-nano-20260226-001320 RLM evaluation results for codeqa using gpt-5-nano. Metrics accuracy: 0.3333333333333333 correct: 1 total: 3 Configuration Parameter Value Backend openai Max iterations 30 Seed 42 Num examples 3 Configs Config Description results Per-example evaluation results rlm_call_traces Per-iteration RLM call traces for the visualizer tabularn<1K0 likes22 downloads8mo agoHugging Face10spatel-learn /codeqa-vagen-trajectoriesimagen<1K0 likes22 downloads7mo agoHugging Face11lissadesu /codeqa_v2 Dataset Card for "codeqa_v2" More Information needed tabular10K<n<100K0 likes18 downloads3y agoHugging Face12smamooler /codeqa-gt50-javatext1K<n<10K0 likes18 downloads6mo agoHugging Face13vm2825 /single_file_code_qa_1k_repos-datasettext10K<n<100K0 likes16 downloads1y agoHugging Face14Tapos-Minmoy /code-qa-6ktext1K<n<10K0 likes14 downloads1y agoHugging Face15reasoning-degeneration-dev /rlm-codeqa-gpt-5-nano-20260225-224658 rlm-codeqa-gpt-5-nano-20260225-224658 RLM evaluation results for codeqa using gpt-5-nano. Metrics accuracy: 0.3333333333333333 correct: 1 total: 3 Configuration Parameter Value Backend openai Max iterations 30 Seed 42 Num examples 3 Traces Per-iteration RLM call traces are available in a companion dataset: reasoning-degeneration-dev/rlm-codeqa-gpt-5-nano-20260225-224658__rlm_call_traces textn<1K0 likes14 downloads8mo agoHugging Face16SousiOmine /codeqa-agent-distill-260703-jatextn<1K0 likes14 downloads3mo agoHugging Face17vm2825 /small_repos_multi_file_chatgpt_5_qas_code_qa_1k-datasettext1K<n<10K0 likes13 downloads1y agoHugging Face18SousiOmine /codeqa-agent-distill-260705-jatextn<1K0 likes13 downloads3mo agoHugging Face19vm2825 /small_repos_multi_file_chatgpt_5_qas_part4_code_qa-datasettext1K<n<10K0 likes11 downloads1y agoHugging Face20reasoning-degeneration-dev /rlm-codeqa-gpt-5-nano-20260225-224658__rlm_call_traces rlm-codeqa-gpt-5-nano-20260225-224658__rlm_call_traces Per-iteration RLM call traces for codeqa using gpt-5-nano. These traces power the agg_visualizer and contain one row per RLM iteration. Parent dataset Results: reasoning-degeneration-dev/rlm-codeqa-gpt-5-nano-20260225-224658 Schema Column Description example_idx Index of the evaluation example rlm_iter RLM iteration number within this example prompt Serialised prompt for this iteration… See the full description on the dataset page: https://huggingface.co/datasets/reasoning-degeneration-dev/rlm-codeqa-gpt-5-nano-20260225-224658__rlm_call_traces.tabularn<1K0 likes11 downloads8mo agoHugging Face21vm2825 /all_1000_multifile_generated_qa_pairs_code_qa-datasettext1K<n<10K0 likes10 downloads1y agoHugging Face22vm2825 /small_repos_multi_file_chatgpt_5_qas_part2_3_code_qa-datasettext10K<n<100K0 likes10 downloads1y agoHugging Face23vm2825 /filtered_repos2_clean_files_picked_gemini_flash_all_code_qa-datasettext10K<n<100K0 likes10 downloads1y agoHugging Face24SousiOmine /codeqa-agent-distill-260709-jatextn<1K0 likes10 downloads3mo agoHugging Face25SousiOmine /codeqa-agent-distill-260704-jatextn<1K0 likes10 downloads3mo agoHugging Face26qadeesanoor /code-switching-codesaviours-si26-qadeesanoorDataset Summary This dataset contains Roman Urdu–English code-switched sentences, the way mixed-language text actually gets written in everyday Pakistani texting, tweeting, and casual conversation (e.g. "Aaj ka din bohot busy tha, had 3 meetings back to back"). Each sentence is broken down word-by-word, and every word is tagged with a language label. No existing Roman Urdu NLP resource handles this kind of within-sentence language mixing well — this dataset is a step toward building tools… See the full description on the dataset page: https://huggingface.co/datasets/qadeesanoor/code-switching-codesaviours-si26-qadeesanoor.text1K<n<10K0 likes10 downloads2mo agoHugging Face27smamooler /codeqa-gt50-combinedtext10K<n<100K0 likes9 downloads6mo agoHugging Face28SunnyLin /Math-VR-QA-codeimagen<1K0 likes9 downloads6mo agoHugging Face29lissadesu /codeqa_v3 Dataset Card for "codeqa_v3" More Information needed tabular10K<n<100K0 likes8 downloads3y agoHugging Face30beimnet777 /Family-code-of-Ethiopia-QA-TESTtextn<1K0 likes8 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.