datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
MMSD2.0
MMSD2.0: Towards a Reliable Multi-modal Sarcasm Detection System
This is a copy of the dataset uploaded on Hugging Face for easy access. The original data comes from this work, which is an improvement upon a previous study.
Usage
from typing import TypedDict, cast
import pytorch_lightning as pl
from datasets import Dataset, load_dataset
from torch import Tensor
from torch.utils.data import DataLoader
from transformers import CLIPProcessor
class… See the full description on the dataset page: https://huggingface.co/datasets/coderchen01/MMSD2.0.ds-coder-instruct-v1
Dataset Card for DS Coder Instruct Dataset
DS Coder is a dataset for instruction fine tuning of language models. It is a specialized dataset focusing only on
data science (eg. plotting, data wrangling, machine learnig models, deep learning, and numerical computations). The dataset contains code examples both in R and Python.
The goal of this dataset is to enable creation of small-scale, specialized language model assistants for data science projects.
Dataset Details… See the full description on the dataset page: https://huggingface.co/datasets/ed001/ds-coder-instruct-v1.qwen3-coder-gb10-vs-rtx5090-benchmark
NVIDIA GB10 vs. GeForce RTX 5090 - Local LLM Inference Benchmark
Model: Qwen3-Coder-30B-A3B-InstructFormat: GGUF, Q4_K_M, 18.63 GBRuntime: LM Studio / llama.cppAuthor: Efehan A.Benchmark date: 5 August 2026
This repository contains a decode-focused local inference benchmark comparing an NVIDIA GB10 system with a Windows workstation containing two GeForce RTX 5090 GPUs. Telemetry shows that the inference workload was carried primarily by a single RTX 5090 (GPU 0), while GPU 1… See the full description on the dataset page: https://huggingface.co/datasets/mreltera/qwen3-coder-gb10-vs-rtx5090-benchmark.cocoindian-traditional-artificial-jewellery
Traditional and Handmade Indian Jewellery Dataset
This dataset contains a comprehensive collection of traditional and handmade Indian jewelry, sourced from various e-commerce platforms and manufacturer websites. It provides a rich set of attributes for each jewelry piece, making it a valuable resource for various data analysis, machine learning, and market research tasks.
Dataset Overview
This dataset is designed to provide detailed information about Indian jewelry… See the full description on the dataset page: https://huggingface.co/datasets/Coder-Dragon/indian-traditional-artificial-jewellery.face_shapeqwen3-coder-480brebus-dataset
|🔄 🚍| Re-Bus: A Large and Diverse Multimodal Benchmark for evaluating the ability of Vision-Language Models to understand Rebus Puzzles
Understanding Rebus Puzzles requires a variety of skills such as image recognition, cognitive skills, commonsense reasoning, and multi-step reasoning, making this a challenging task for current Vision-Language Models. In this paper, we present Re-Bus, a large and diverse benchmark of 1,333 English Rebus Puzzles containing different artistic… See the full description on the dataset page: https://huggingface.co/datasets/codergautam/rebus-dataset.Cityscapes_KAISTv2This is the result obtained by inference on the Cityscapes dataset using the KAIST weights provided by the official F-ViTA documentation (resolution 512).
retirement-invitation-mamtaCityscapes_M3FDThis repository contains the inference results on the CityScapes dataset. The model weights used for this inference were obtained by training on the M3FD dataset, utilizing the official training configuration provided by F-ViTA.
C3BEnglish | 简体中文
C³B: Comics Cross-Cultural Benchmark
Culture In a Frame: C³B as a Comic-Based Benchmark for Multimodal Cultural Awareness
ICLR 2026
About C³B
C³B (Comics Cross-Cultural Benchmark) is a multicultural, multitask, and multilingual benchmark for evaluating cultural awareness capabilities of Multimodal Large Language Models (MLLMs).
Progressive task difficulty: From basic visual recognition, to higher-level cultural conflict understanding, to cultural content… See the full description on the dataset page: https://huggingface.co/datasets/Coder109/C3B.DocMSU
DocMSU: A Comprehensive Benchmark for Document-level Multimodal Sarcasm Understanding
This is a replication of the DocMSU dataset for easier access.
Reference
Du, H., Nan, G., Zhang, S., Xie, B., Xu, J., Fan, H., Cui, Q., Tao, X. and Jiang, X., 2024, March. DocMSU: A Comprehensive Benchmark for Document-Level Multimodal Sarcasm Understanding. In Proceedings of the AAAI Conference on Artificial Intelligence (Vol. 38, No. 16, pp. 17933-17941).
qwen3-coder-30bSupervised-Fog-Removal-DatasetSupervised Fog Removal Dataset
Overview
This dataset contains 80,000 paired images designed for supervised image dehazing / fog removal tasks.
Each sample consists of:
a clean image (ground truth)
a synthetically fogged version of that image
The fog is generated using a physics-inspired atmospheric scattering model combined with depth estimation, allowing the fog to behave realistically with respect to scene geometry.
Unlike simple uniform haze overlays, this dataset simulates depth-aware fog… See the full description on the dataset page: https://huggingface.co/datasets/Aeye-coder/Supervised-Fog-Removal-Dataset.lexi-coder-v4.3imdb-clean-2010plus-singleface
IMDB-Clean 2010+ SingleFace Subset
Dataset Description
This dataset is a filtered subset of IMDB-Clean for age estimation and resolution impact studies.It contains 1,000 face images selected under strict criteria:
Source: IMDB-WIKI / IMDB-Clean
Year filter: only photos taken after 2010
Face detection: exactly one face per image (using InsightFace)
Image quality: resolution ≥ 1024 px
A metadata CSV (subset_metadata.csv) is included with age, bounding box, and… See the full description on the dataset page: https://huggingface.co/datasets/codershiyar/imdb-clean-2010plus-singleface.School-Supplies-1
School Supplies Dataset
This dataset contains a comprehensive collection of school supply products. It is a valuable resource for a variety of tasks, including market analysis, price comparison, and inventory management. Some example of products are eraser, pencil box, notebook, geometry etc.
Data Description
The dataset is structured as a single table (e.g., a CSV file) with the following columns:
Column Name
Description
Data Type
Product Name
The name of… See the full description on the dataset page: https://huggingface.co/datasets/Coder-Dragon/School-Supplies-1.CAD-experiment-manifests-seed42-vision-qwen3vl32b-v1-strict-coder-v1CAD-experiment-manifests-seed42-vision-qwen3vl4b-coderprompts-v1CAD-experiment-manifests-seed202-vision-qwen3vl4b-coderprompts-v1CAD-experiment-manifests-seed42-vision-qwen3vl32b-coderprompts-v1CAD-experiment-manifests-seed202-vision-qwen3vl32b-coderprompts-v1CAD-experiment-manifests-seed303-vision-qwen3vl32b-coderprompts-v1vision_coder_llmimage_generationAerial-Visual-Groundingcode_RongyunCAD-experiment-manifests-seed303-vision-qwen3vl4b-coderprompts-v1agentic_dataset_qwen3_coder_web
