Team Ai
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01VLR-CVC /ComicsPAP Comics: Pick-A-Panel Updated val and test on 25/02/2025 This is the dataset for the ICDAR 2025 Competition on Comics Understanding in the Era of Foundational Models. Please, check out our 🚀 arxiv paper 🚀 for more information 😊 The competition is hosted in the Robust Reading Competition website and the leaderboard is available here. The dataset contains five subtask or skills: Sequence Filling Given a sequence of comic panels, a missing panel, and a set of option panels, the… See the full description on the dataset page: https://huggingface.co/datasets/VLR-CVC/ComicsPAP.image10K<n<100K15 likes1.9k downloads1y agoHugging Face02ansabgillani /ComicsPAP Comics: Pick-A-Panel Updated val and test on 25/02/2025 This is the dataset for the ICDAR 2025 Competition on Comics Understanding in the Era of Foundational Models. Please, check out our 🚀 arxiv paper 🚀 for more information 😊 The competition is hosted in the Robust Reading Competition website and the leaderboard is available here. The dataset contains five subtask or skills: Sequence Filling Given a sequence of comic panels, a missing panel, and a set of option panels, the… See the full description on the dataset page: https://huggingface.co/datasets/ansabgillani/ComicsPAP.image100K<n<1M0 likes1.6k downloads10mo agoHugging Face03crimsonmythos /comics-tools FLUX workflows — c0sm1c_m1a (Mia) Generation workflows for the FLUX.1-dev LoRA line. SDXL-era workflows remain in workflows/ root (historical). mia-flux-v1-test.json — LoRA smoke test Minimal single-sampler graph, stock nodes only (no custom packs needed). UI-format JSON: drag-drop onto the ComfyUI canvas or Import. Graph: CheckpointLoaderSimple → LoraLoader → CLIPTextEncode (pos) + CLIPTextEncode (neg, inert) + EmptyLatentImage → KSampler → VAEDecode → SaveImage… See the full description on the dataset page: https://huggingface.co/datasets/crimsonmythos/comics-tools.0 likes250 downloads20d agoHugging Face04purdueinformatics /comics-audio-video Comics Audio Video Data Notes Dataset summary Preparation notes and schema examples for Comics tasks using Audio Video data. Full source material is intentionally not bundled, so provenance and licensing remain explicit. Included material dataset.py — loading, cleaning, and split preparation code. dataset_infos.json — schema and split metadata. metadata_sample.jsonl — small, human-readable records for checking the schema. README.md — data card and… See the full description on the dataset page: https://huggingface.co/datasets/purdueinformatics/comics-audio-video.0 likes131 downloads26d agoHugging Face05EarthnDusk /Star_Marvel_comics Dataset Card for Star Villain Marvel Comics LoRa Data set for Duskfallcrew/Star_Marvel_comics_LoRa Trained with: https://colab.research.google.com/github/Linaqruf/kohya-trainer/blob/main/kohya-LoRA-dreambooth.ipynb Where else can i find this ? Both safetensors files are at: https://civitai.com/models/14831/star-ryan-ripley The outputs aren't comic format. I haven't tested it in WEB UI yet, the scripted outputs largely rely on the actual model.… See the full description on the dataset page: https://huggingface.co/datasets/EarthnDusk/Star_Marvel_comics.text-to-image1K<n<10K1 likes94 downloads4y agoHugging Face06josantos6 /comics-audio-video-benchmark Comics Audio Video Data Notes Dataset summary This repository contains a preparation pipeline and a small metadata sample for Comics work with Audio Video inputs. It does not claim to be a complete benchmark release; the loader documents how source data is normalized and validated. Included material load_data.py — loading, cleaning, and split preparation code. dataset_infos.json — schema and split metadata. metadata_sample.jsonl — small… See the full description on the dataset page: https://huggingface.co/datasets/josantos6/comics-audio-video-benchmark.0 likes93 downloads23d agoHugging Face07aleksawtf /comics_dataset_lineart_1024image1K<n<10K7 likes86 downloads2y agoHugging Face08EarthnDusk /ComicsPonyXLimagen<1K0 likes63 downloads3y agoHugging Face09watanabekelvin /dl-comics Comics Image Depth Data Notes Dataset summary This repository contains a preparation pipeline and a small metadata sample for Comics work with Image Depth inputs. It does not claim to be a complete benchmark release; the loader documents how source data is normalized and validated. Included material clean.py — loading, cleaning, and split preparation code. dataset_infos.json — schema and split metadata. metadata_sample.jsonl — small, human-readable… See the full description on the dataset page: https://huggingface.co/datasets/watanabekelvin/dl-comics.0 likes61 downloads11d agoHugging Face10wwojcikantoni /comics-collection Comics Image Depth Data Notes Dataset summary This data card accompanies a lightweight Comics loader for Image Depth metadata. It is meant for pipeline inspection, source adaptation, and reproducible split preparation. Included material dataloader.py — loading, cleaning, and split preparation code. dataset_infos.json — schema and split metadata. metadata_sample.jsonl — small, human-readable records for checking the schema. README.md — data card and… See the full description on the dataset page: https://huggingface.co/datasets/wwojcikantoni/comics-collection.0 likes53 downloads26d agoHugging Face11bghira /comicstrips-gpt4o-blip3 Comic Strips Dataset Details Dataset Description This dataset contains indie comics from Reddit, then captioned with GPT4o and BLIP3. Currently, only the GPT4o captions are available in this repository. The BLIP3 captions will be uploaded soon. Roughly 1400 images were captioned at a cost of ~$11 using GPT4o (25 May 2024 version). Curated by: @pseudoterminalx Funded by @pseudoterminalx License: MIT Dataset Sources Unlike other free-to-use… See the full description on the dataset page: https://huggingface.co/datasets/bghira/comicstrips-gpt4o-blip3.image1K<n<10K11 likes49 downloads2y agoHugging Face12rraoswati /nlp-comics Comics Image Audio Data Notes Dataset summary Preparation notes and schema examples for Comics tasks using Image Audio data. Full source material is intentionally not bundled, so provenance and licensing remain explicit. Included material dataloader.py — loading, cleaning, and split preparation code. dataset_infos.json — schema and split metadata. metadata_sample.jsonl — small, human-readable records for checking the schema. README.md — data card… See the full description on the dataset page: https://huggingface.co/datasets/rraoswati/nlp-comics.0 likes47 downloads27d agoHugging Face13yazeedalha /paper-comics-2023 Comics Pointcloud Text Data Notes Dataset summary This repository contains a preparation pipeline and a small metadata sample for Comics work with Pointcloud Text inputs. It does not claim to be a complete benchmark release; the loader documents how source data is normalized and validated. Included material loader.py — loading, cleaning, and split preparation code. dataset_infos.json — schema and split metadata. metadata_sample.jsonl — small… See the full description on the dataset page: https://huggingface.co/datasets/yazeedalha/paper-comics-2023.0 likes47 downloads24d agoHugging Face14amritastatistics04 /comics-sensor-fusion-mini Comics Sensor Fusion Data Notes Dataset summary This repository contains a preparation pipeline and a small metadata sample for Comics work with Sensor Fusion inputs. It does not claim to be a complete benchmark release; the loader documents how source data is normalized and validated. Included material prepare.py — loading, cleaning, and split preparation code. dataset_infos.json — schema and split metadata. metadata_sample.jsonl — small… See the full description on the dataset page: https://huggingface.co/datasets/amritastatistics04/comics-sensor-fusion-mini.0 likes46 downloads21d agoHugging Face15LevPopov /postdoc-comics-2024 Comics Text Tabular Data Notes Dataset summary This data card accompanies a lightweight Comics loader for Text Tabular metadata. It is meant for pipeline inspection, source adaptation, and reproducible split preparation. Included material dataloader.py — loading, cleaning, and split preparation code. dataset_infos.json — schema and split metadata. metadata_sample.jsonl — small, human-readable records for checking the schema. README.md — data card… See the full description on the dataset page: https://huggingface.co/datasets/LevPopov/postdoc-comics-2024.0 likes45 downloads22d agoHugging Face16aleksawtf /comics_dataset_512_inv_manga_correctedimage1K<n<10K1 likes43 downloads2y agoHugging Face17CaiThuocLa25 /comic-setup0 likes43 downloads2mo agoHugging Face18madisonjones0910 /comics-pointcloud-text-v2-2024 Comics Pointcloud Text Data Notes Dataset summary Preparation notes and schema examples for Comics tasks using Pointcloud Text data. Full source material is intentionally not bundled, so provenance and licensing remain explicit. Included material dataset.py — loading, cleaning, and split preparation code. dataset_infos.json — schema and split metadata. metadata_sample.jsonl — small, human-readable records for checking the schema. README.md — data… See the full description on the dataset page: https://huggingface.co/datasets/madisonjones0910/comics-pointcloud-text-v2-2024.0 likes43 downloads26d agoHugging Face19mwilliamsdale /homework-comics Comics Video Text Data Notes Dataset summary This data card accompanies a lightweight Comics loader for Video Text metadata. It is meant for pipeline inspection, source adaptation, and reproducible split preparation. Included material loader.py — loading, cleaning, and split preparation code. dataset_infos.json — schema and split metadata. metadata_sample.jsonl — small, human-readable records for checking the schema. README.md — data card and usage… See the full description on the dataset page: https://huggingface.co/datasets/mwilliamsdale/homework-comics.0 likes43 downloads11d agoHugging Face20brunoymartins /comics-collection Comics Audio Text Data Notes Dataset summary This repository contains a preparation pipeline and a small metadata sample for Comics work with Audio Text inputs. It does not claim to be a complete benchmark release; the loader documents how source data is normalized and validated. Included material preprocess.py — loading, cleaning, and split preparation code. dataset_infos.json — schema and split metadata. metadata_sample.jsonl — small… See the full description on the dataset page: https://huggingface.co/datasets/brunoymartins/comics-collection.0 likes42 downloads28d agoHugging Face21joshuathomaswood /comics-video-text Comics Video Text Data Notes Dataset summary This data card accompanies a lightweight Comics loader for Video Text metadata. It is meant for pipeline inspection, source adaptation, and reproducible split preparation. Included material dataset.py — loading, cleaning, and split preparation code. dataset_infos.json — schema and split metadata. metadata_sample.jsonl — small, human-readable records for checking the schema. README.md — data card and… See the full description on the dataset page: https://huggingface.co/datasets/joshuathomaswood/comics-video-text.0 likes41 downloads21d agoHugging Face22edwardswilliam /comics-audio-video-clean Comics Audio Video Data Notes Dataset summary Preparation notes and schema examples for Comics tasks using Audio Video data. Full source material is intentionally not bundled, so provenance and licensing remain explicit. Included material dataset.py — loading, cleaning, and split preparation code. dataset_infos.json — schema and split metadata. metadata_sample.jsonl — small, human-readable records for checking the schema. README.md — data card and… See the full description on the dataset page: https://huggingface.co/datasets/edwardswilliam/comics-audio-video-clean.0 likes41 downloads13d agoHugging Face23briangonzalezora /comics-image-depth Comics Image Depth Data Notes Dataset summary This data card accompanies a lightweight Comics loader for Image Depth metadata. It is meant for pipeline inspection, source adaptation, and reproducible split preparation. Included material prepare.py — loading, cleaning, and split preparation code. dataset_infos.json — schema and split metadata. metadata_sample.jsonl — small, human-readable records for checking the schema. README.md — data card and… See the full description on the dataset page: https://huggingface.co/datasets/briangonzalezora/comics-image-depth.0 likes40 downloads21d agoHugging Face24filipgrabowski /comics-image-text Comics Image Text Data Notes Dataset summary This data card accompanies a lightweight Comics loader for Image Text metadata. It is meant for pipeline inspection, source adaptation, and reproducible split preparation. Included material prepare.py — loading, cleaning, and split preparation code. dataset_infos.json — schema and split metadata. metadata_sample.jsonl — small, human-readable records for checking the schema. README.md — data card and… See the full description on the dataset page: https://huggingface.co/datasets/filipgrabowski/comics-image-text.0 likes39 downloads25d agoHugging Face25jonas-neumann /comics-sensor-fusion Comics Sensor Fusion Data Notes Dataset summary Preparation notes and schema examples for Comics tasks using Sensor Fusion data. Full source material is intentionally not bundled, so provenance and licensing remain explicit. Included material build_dataset.py — loading, cleaning, and split preparation code. dataset_infos.json — schema and split metadata. metadata_sample.jsonl — small, human-readable records for checking the schema. README.md — data… See the full description on the dataset page: https://huggingface.co/datasets/jonas-neumann/comics-sensor-fusion.0 likes39 downloads17d agoHugging Face26darrenhuasaw /comics-audio-video-benchmark Comics Audio Video Data Notes Dataset summary This data card accompanies a lightweight Comics loader for Audio Video metadata. It is meant for pipeline inspection, source adaptation, and reproducible split preparation. Included material loader.py — loading, cleaning, and split preparation code. dataset_infos.json — schema and split metadata. metadata_sample.jsonl — small, human-readable records for checking the schema. README.md — data card and… See the full description on the dataset page: https://huggingface.co/datasets/darrenhuasaw/comics-audio-video-benchmark.0 likes39 downloads6d agoHugging Face27aleksawtf /comics_orig_lineartimage1K<n<10K1 likes38 downloads2y agoHugging Face28Felixschmidt /comics-corpus Comics Image Audio Data Notes Dataset summary This data card accompanies a lightweight Comics loader for Image Audio metadata. It is meant for pipeline inspection, source adaptation, and reproducible split preparation. Included material build_dataset.py — loading, cleaning, and split preparation code. dataset_infos.json — schema and split metadata. metadata_sample.jsonl — small, human-readable records for checking the schema. README.md — data card… See the full description on the dataset page: https://huggingface.co/datasets/Felixschmidt/comics-corpus.0 likes38 downloads20d agoHugging Face29Arjunchopra /comics-corpus Comics Multimodal3 Data Notes Dataset summary Preparation notes and schema examples for Comics tasks using Multimodal3 data. Full source material is intentionally not bundled, so provenance and licensing remain explicit. Included material build_dataset.py — loading, cleaning, and split preparation code. dataset_infos.json — schema and split metadata. metadata_sample.jsonl — small, human-readable records for checking the schema. README.md — data… See the full description on the dataset page: https://huggingface.co/datasets/Arjunchopra/comics-corpus.0 likes37 downloads11d agoHugging Face30ggarniersebastien1987 /comics-image-text Comics Image Text Data Notes Dataset summary Preparation notes and schema examples for Comics tasks using Image Text data. Full source material is intentionally not bundled, so provenance and licensing remain explicit. Included material preprocess.py — loading, cleaning, and split preparation code. dataset_infos.json — schema and split metadata. metadata_sample.jsonl — small, human-readable records for checking the schema. README.md — data card and… See the full description on the dataset page: https://huggingface.co/datasets/ggarniersebastien1987/comics-image-text.0 likes36 downloads13d agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.