Team Ai
18 results

open-coding

open-r1 /verifiable-coding-problems-python Dataset Card for Verifiable Coding Problems Python 10k This dataset contains all Python problems from PrimeIntellect's verifiable-coding-problems dataset. We have formatted the verification_info and metadata columns to be proper dictionaries, but otherwise the data is the same. Please see their dataset for more details. text10K<n<100K12 likes2.8k downloads2y agoHugging Faceopen-r1 /verifiable-coding-problems-python_decontaminated-testedtext10K<n<100K0 likes844 downloads2y agoHugging Faceopen-r1 /verifiable-coding-problems-python_decontaminated-tested-shuffledtext10K<n<100K2 likes602 downloads2y agoHugging Faceopen-r1 /verifiable-coding-problems-python_decontaminatedtext10K<n<100K5 likes446 downloads2y agoHugging Faceopen-athena /nemotron-gym-competitive-coding-qwen3.5-122b-131k-opencode-traces Agent trace dataset Decoding the literal token IDs The prompt_token_ids / completion_token_ids / logprobs columns are the verbatim tokens the serving engine emitted, stored PER AGENT STEP as a list-of-lists (one inner list per turn). To turn them back into text you MUST use the exact tokenizer the model was served with — a generic same-family tokenizer will decode word tokens to garbage. Served model / tokenizer source: Qwen/Qwen3.5-122B-A10B-FP8 from transformers… See the full description on the dataset page: https://huggingface.co/datasets/open-athena/nemotron-gym-competitive-coding-qwen3.5-122b-131k-opencode-traces.text10K<n<100K0 likes208 downloads2mo agoHugging Facesuzhentxt /open-r1-truncated-coding-pythontext10K<n<100K0 likes61 downloads1y agoHugging Face