Team Ai
10 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01open-r1 /verifiable-coding-problems-python Dataset Card for Verifiable Coding Problems Python 10k This dataset contains all Python problems from PrimeIntellect's verifiable-coding-problems dataset. We have formatted the verification_info and metadata columns to be proper dictionaries, but otherwise the data is the same. Please see their dataset for more details. text10K<n<100K12 likes2.8k downloads2y agoHugging Face02open-r1 /verifiable-coding-problems-python_decontaminated-testedtext10K<n<100K0 likes844 downloads2y agoHugging Face03open-r1 /verifiable-coding-problems-python_decontaminated-tested-shuffledtext10K<n<100K2 likes602 downloads2y agoHugging Face04open-r1 /verifiable-coding-problems-python_decontaminatedtext10K<n<100K5 likes446 downloads2y agoHugging Face05open-athena /nemotron-gym-competitive-coding-qwen3.5-122b-131k-opencode-traces Agent trace dataset Decoding the literal token IDs The prompt_token_ids / completion_token_ids / logprobs columns are the verbatim tokens the serving engine emitted, stored PER AGENT STEP as a list-of-lists (one inner list per turn). To turn them back into text you MUST use the exact tokenizer the model was served with — a generic same-family tokenizer will decode word tokens to garbage. Served model / tokenizer source: Qwen/Qwen3.5-122B-A10B-FP8 from transformers… See the full description on the dataset page: https://huggingface.co/datasets/open-athena/nemotron-gym-competitive-coding-qwen3.5-122b-131k-opencode-traces.text10K<n<100K0 likes208 downloads2mo agoHugging Face06suzhentxt /open-r1-truncated-coding-pythontext10K<n<100K0 likes61 downloads1y agoHugging Face07open-athena /nemotron-gym-competitive-coding-minimax-m27-131k-tracestext1K<n<10K0 likes54 downloads4mo agoHugging Face08open-athena /nemotron-gym-competitive-coding-qwen3.5-122b-32k-tracestext10K<n<100K0 likes52 downloads3mo agoHugging Face09open-llm-leaderboard-old /details_uukuguy__speechless-coding-7b-16k-tora Dataset Card for Evaluation run of uukuguy/speechless-coding-7b-16k-tora Dataset Summary Dataset automatically created during the evaluation run of model uukuguy/speechless-coding-7b-16k-tora on the Open LLM Leaderboard. The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_uukuguy__speechless-coding-7b-16k-tora.0 likes22 downloads3y agoHugging Face10open-llm-leaderboard-old /details_speechlessai__speechless-coding-7b-16k-tora Dataset Card for Evaluation run of speechlessai/speechless-coding-7b-16k-tora Dataset Summary Dataset automatically created during the evaluation run of model speechlessai/speechless-coding-7b-16k-tora on the Open LLM Leaderboard. The dataset is composed of 1 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_speechlessai__speechless-coding-7b-16k-tora.0 likes16 downloads3y agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.