datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
local-code-arena-starcoder2_15b
Local Code Arena Telemetry: MBPP Benchmark on StarCoder2 15B (Base)
This repository hosts the raw evaluation metrics, execution telemetry logs, and structural syntax outputs captured from running the Mostly Basic Python Problems (MBPP) benchmark against the flagship StarCoder2 15B base foundational model.
This specific partition documents the final limits of scaling raw, unaligned foundational weights inside conversational evaluation loops, establishing an absolute baseline for… See the full description on the dataset page: https://huggingface.co/datasets/ShahzebKhoso/local-code-arena-starcoder2_15b.local-code-arena-starcoder2_3b
Local Code Arena Telemetry: MBPP Benchmark on StarCoder2 3B (Base)
This repository hosts the raw evaluation metrics, execution telemetry logs, and structural syntax outputs captured from running the Mostly Basic Python Problems (MBPP) benchmark against the next-generation StarCoder2 3B base foundational model.
This specific partition documents the behavioral dynamics of modern raw foundational weights inside automated conversational pipelines, highlighting the persistent… See the full description on the dataset page: https://huggingface.co/datasets/ShahzebKhoso/local-code-arena-starcoder2_3b.local-code-arena-starcoder2_7b
Local Code Arena Telemetry: MBPP Benchmark on StarCoder2 7B (Base)
This repository hosts the raw evaluation metrics, execution telemetry logs, and structural syntax outputs captured from running the Mostly Basic Python Problems (MBPP) benchmark against the next-generation StarCoder2 7B base foundational model.
This specific partition documents the behavioral dynamics of modern, mid-tier raw foundational weights inside automated conversational evaluation workflows, defining the… See the full description on the dataset page: https://huggingface.co/datasets/ShahzebKhoso/local-code-arena-starcoder2_7b.
