Copilot
Datasets
All datasets matching “Copilot”OS-Atlas-data
GUI Grounding Pre-training Data for OS-ATLAS
This document describes the acquisition of the pre-training data used by OS-ATLAS OS-ATLAS: A Foundation Action Model for Generalist GUI Agents.
[🏠Homepage] [💻Code] [🚀Quick Start] [📝Paper] [🤗Models] [🤗ScreenSpot-v2]
Notes: In GUI grounding data, the position of the target element is recorded in the bbox key, represented by [left, top, right, bottom].
Each value is a [0, 1] decimal number indicating the ratio of the… See the full description on the dataset page: https://huggingface.co/datasets/OS-Copilot/OS-Atlas-data.ScreenSpot-v2python-text-copilot-training-instruct-ai-research-2024-02-03
Python Copilot Instructions on How to Code using Alpaca and Yaml
Training and test datasets for building coding multimodal models that understand how to use the open source GitHub projects for the Agora Open Source AI Research Lab:
Agora GitHub Organization
Agora Hugging Face
This dataset is the 2024-02-03 update for the matlok python copilot datasets. Please refer to the Multimodal Python Copilot Training Overview for more details on how to use this dataset.
Details… See the full description on the dataset page: https://huggingface.co/datasets/matlok/python-text-copilot-training-instruct-ai-research-2024-02-03.python-image-copilot-training-using-import-knowledge-graphs
Python Copilot Image Training using Import Knowledge Graphs
This dataset is a subset of the matlok python copilot datasets. Please refer to the Multimodal Python Copilot Training Overview for more details on how to use this dataset.
Details
Each row contains a png file in the dbytes column.
Rows: 216642
Size: 211.2 GB
Data type: png
Format: Knowledge graph using NetworkX with alpaca text box
Schema
The png is in the dbytes column:
{
"dbytes": "binary"… See the full description on the dataset page: https://huggingface.co/datasets/matlok/python-image-copilot-training-using-import-knowledge-graphs.OSReward
OSReward Benchmark
OSReward evaluates whether a multimodal judge can determine if a computer-use
agent completed a user's task. This release contains binary outcome labels only:
SUCCESS and FAIL.
The benchmark has two evaluation configurations:
Configuration
Trajectories
Unique task IDs
SUCCESS
FAIL
full
1,019
656
440
579
hard
284
257
86
198
Load the dataset
Load either benchmark configuration with datasets:
from datasets import load_dataset… See the full description on the dataset page: https://huggingface.co/datasets/OS-Copilot/OSReward.python-image-copilot-training-using-class-knowledge-graphs
Python Copilot Image Training using Class Knowledge Graphs
This dataset is a subset of the matlok python copilot datasets. Please refer to the Multimodal Python Copilot Training Overview for more details on how to use this dataset.
Details
Each row contains a png file in the dbytes column.
Rows: 312277
Size: 304.3 GB
Data type: png
Format: Knowledge graph using NetworkX with alpaca text box
Schema
The png is in the dbytes column:
{
"dbytes": "binary"… See the full description on the dataset page: https://huggingface.co/datasets/matlok/python-image-copilot-training-using-class-knowledge-graphs.
