datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
multimodal-peer-collaboration-samples
Multimodal Peer Collaboration Samples - Embodied Map Task with Two Camera Angles
Two non-experts collaborate to build working circuits under asymmetric information: the instructor has the manual, the student has the components, and synchronized audio and dual-camera video capture how shared understanding emerges.
▶ Watch the interactions · See Expert Instruction samples · Discuss the full collection
Sister collection: Expert Instruction, a teacher and a student in… See the full description on the dataset page: https://huggingface.co/datasets/fluid-concepts/multimodal-peer-collaboration-samples.multimodal-ai-taxonomy
Multimodal AI Taxonomy
A comprehensive, structured taxonomy for mapping multimodal AI model capabilities across input and output modalities.
Dataset Description
This dataset provides a systematic categorization of multimodal AI capabilities, enabling users to:
Navigate the complex landscape of multimodal AI models
Filter models by specific input/output modality combinations
Understand the nuanced differences between similar models (e.g., image-to-video with/without audio… See the full description on the dataset page: https://huggingface.co/datasets/danielrosehill/multimodal-ai-taxonomy.nekoneko-industry-multimodal-retrieval-benchmark
Nekoneko Industry Multimodal Retrieval Benchmark v0.1
English description: A fully synthetic Japanese benchmark for evaluating retrieval across office documents, spreadsheets, presentations, text-to-speech audio, and silent videos.
公開版: v0.1。架空の企業情報を使う小規模な検索ベンチマークです。オリジナルの資料・質問・注釈はCC BY 4.0、評価スクリプトはMITで提供します。合成音声の作成方法と権利確認の範囲は SOURCE_NOTICE.md を参照してください。
概要… See the full description on the dataset page: https://huggingface.co/datasets/MakiAi/nekoneko-industry-multimodal-retrieval-benchmark.
