datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
trade_vision_dataset
TradeVision: Hierarchical Physical Business & Multimodal Retail Provenance Dataset
This dataset is continuously seeded from OpenStreetMap, matched to Google Place IDs, harvested for temporal store photos, and enriched with zero-shot computer vision using Hugging Face Hub native pipelines.
Dataset Structure
The dataset is partitioned into three relational subsets loadable via Hugging Face datasets:
from datasets import load_dataset
# 1. Load Canonical Businesses… See the full description on the dataset page: https://huggingface.co/datasets/drksci/trade_vision_dataset.VisionEncoder-Eval-ReproDataA Strong Baseline for Evaluating Vision Encodersin Multimodal Large Language Models
Yilin Yang1,* ·
Jun-Tao Tang2,* ·
Kengyi Wang3 ·
Siyuan Su3 ·
Gaoyong Luo4 ·
Mingda Chen1,†
1School of Artificial Intelligence, Shanghai Jiao Tong University
2Nanjing University ·
3Fudan University ·
4Independent Researcher
*Equal contribution. †Corresponding author.… See the full description on the dataset page: https://huggingface.co/datasets/336labs/VisionEncoder-Eval-ReproData.
