data-analysis
sentiment_analysis_generic_datasetSentiment-Analysis-of-Banking-Dataset-GGUFkorean_sentiment_analysis_dataset3distilbert-base-uncased-sentiment-analysis-movie-reviewsgitlab-mr-analysis-default-modelv5_balanced_dataset_fine-tuning-java-indo-sentiment-analysist-3-classDataMind-Analysis-Qwen2.5-7Bkorean_sentiment_analysis_dataset3_best
data_analysis
Dataset Card for "livebench/data_analysis"
LiveBench is a benchmark for LLMs designed with test set contamination and objective evaluation in mind. It has the following properties:
LiveBench is designed to limit potential contamination by releasing new questions monthly, as well as having questions based on recently-released datasets, arXiv papers, news articles, and IMDb movie synopses.
Each question has verifiable, objective ground-truth answers, allowing hard questions to be… See the full description on the dataset page: https://huggingface.co/datasets/livebench/data_analysis.routing_analysis-code-data
routing_analysis — code and data
This dataset repository stores the non-checkpoint portion of the
routing_analysis filesystem snapshot as individual files under their
original relative paths. Files are uploaded directly; they are not packed into
split tar archives.
The selection is enumerated from the filesystem and does not consult
.gitignore. Ignored files and .gitignore files themselves are therefore
included whenever they belong to the code/data selection.
Training… See the full description on the dataset page: https://huggingface.co/datasets/lylybig/routing_analysis-code-data.DataMind-Analysis-SFT-DataThis repository contains the data presented in Why Do Open-Source LLMs Struggle with Data Analysis? A Systematic Empirical Study
Code: https://github.com/zjunlp/DataMind
religious-artwork-analysis-data
Data
Download from Kaggle (needs an API token from https://www.kaggle.com/settings):
pip install kaggle
python data/download.py
Expected layout after download:
data/artwork_metadata.csv 3,997 rows — filename, religion (1,000 each of
buddhism / christianity / hinduism; 997 islam),
sub_religion, artist, title, year, place,
source, source_id, source_url, image_url
data/images/ the… See the full description on the dataset page: https://huggingface.co/datasets/cvikl/religious-artwork-analysis-data.sentiment_analysis_data
Dataset Card for "sentiment_analysis_data"
More Information needed
routing_analysis-finetuning-data
routing_analysis finetuning data
A mirror of routing_analysis/finetuning/data/: the JSONL training splits used
in the multilingual MoE language-expansion experiments, plus the generator
scripts, runners, manifests and logs that produced them.
Layout
Path
Contents
Size
splits/
Document-count tiers {lang}_{tier}.jsonl, {lang}_manifest.json, splits_summary.json
~36 GB
token_budget_splits/
Token-budget tiers {lang}_{1Btok,1p5Btok,2Btok}.jsonl and… See the full description on the dataset page: https://huggingface.co/datasets/lylybig8/routing_analysis-finetuning-data.
