bloom
Datasets
All datasets matching “bloom”figma-slide-benchmark
Figma Slide Editing Benchmark
Benchmark accompanying our EMNLP 2026 Industry Track (Main) accepted paper "ACE: A
Self-Correcting Agentic Canvas Editor for Multi-Slide Presentation
Automation".
📄 Paper: https://arxiv.org/pdf/2608.24103
💻 Code: https://github.com/BloomBerry/agentic-canvas-editor
Overview
Each benchmark item is a slide-editing task defined as a pair of Figma Slides
documents:
*_TestA — the input deck the agent starts from.
*_GroundTruthA — the… See the full description on the dataset page: https://huggingface.co/datasets/BloomBerry/figma-slide-benchmark.misalignment-indicators-bloom-rolloutsphytoplankton-microscopy
Bloombio Phytoplankton Microscopy Dataset
The Bloombio Phytoplankton Microscopy Dataset is a curated, citable marine science dataset consisting of 43 phytoplankton species across 10,433 high-resolution light microscopy images with bounding box annotations.
Developed as part of the Bloombio Marine Intelligence Platform — a platform that equips scientists with an autonomous AI agent, compressing sampling setup, multi-modal species identification, and environmental risk assessment… See the full description on the dataset page: https://huggingface.co/datasets/bloombio/phytoplankton-microscopy.vocab-bloom-hub-en
Vocab Bloom Hub — English
A structured English lexical dataset with translations into Russian, Spanish, French, German, Portuguese, Chinese and Arabic, maintained by the Vocab Bloom Hub project — documentation, the API reference and a playground at vocab-bloom-hub.com.
Every entry carries IPA transcription, a CEFR level, one or more sense-level definitions with usage examples, synonym and antonym links per sense, translations per sense in seven languages, and inflected forms —… See the full description on the dataset page: https://huggingface.co/datasets/Fristail27/vocab-bloom-hub-en.details_bigscience__bloom-7b1
Dataset Card for Evaluation run of bigscience/bloom-7b1
Dataset Summary
Dataset automatically created during the evaluation run of model bigscience/bloom-7b1 on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 10 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_bigscience__bloom-7b1.bloomberg_financial_news_120k
