datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
multimodal-example
Multimodal Example Dataset
Small example dataset for testing multimodal (vision-language) fine-tuning with ms-swift.
Structure
├── train.jsonl # 10 training samples
├── test.jsonl # 2 validation samples
├── images/ # All referenced images (400x300 JPEG)
│ ├── dog_portrait.jpg
│ ├── forest_river.jpg
│ ├── laptop_desk.jpg
│ ├── mountain_lake.jpg
│ ├── ocean_rocks.jpg
│ ├── coffee_cup.jpg
│ ├── bookshelf.jpg
│ ├──… See the full description on the dataset page: https://huggingface.co/datasets/f13rnd/multimodal-example.videosearch-r1-demo-examples
VideoSearch-R1 demo examples
23 hand-picked qualitative examples from VideoSearch-R1: Iterative Video Retrieval and Reasoning via Soft Query Refinement (ECCV 2026, arXiv:2607.00446).
Each example has the video files the model saw, the text query, the ground-truth moment, and the model's full multi-turn log (reasoning, search, verification, grounding, final answer).
Model outputs are verbatim eval logs of the released checkpoints VideoSearchR1/didemo-stage2 and… See the full description on the dataset page: https://huggingface.co/datasets/happy8825/videosearch-r1-demo-examples.Boat_unity_exampleintent2edge-examples
Intent2Edge Toy Examples
Product: Intent2Edge, a compiler that turns a natural-language edge-vision request into a trained ONNX candidate with a recomputable proof bundle. The pipeline runs a typed sequence: prompt, dataset config, training, calibration with conformal abstention, ONNX export, an optional compression step with an accuracy-floor gate, and a hard-case capture store.
What this is (and isn't)
This is not a benchmark dataset. It holds the small… See the full description on the dataset page: https://huggingface.co/datasets/Dhi-Technologies/intent2edge-examples.Boat_unity_exampleBoat_unity_example
