android
Datasets
All datasets matching “android”android-times-articles
Android Times — Articles Dataset Archive
Private dataset containing synthesized and processed article archives, multi-language transcripts, metadata, and editorial assets for Android Times.
Dataset Structure
articles/
├── en-US/ # English (United States) localized articles & scripts
├── ja-JP/ # Japanese localized articles & scripts
├── en-AU/ # Australian localized articles
├── en-CN/ # China localized English… See the full description on the dataset page: https://huggingface.co/datasets/aoiandroid/android-times-articles.AndroidtvasstestAndroid-in-the-Wild
Android in the Wild (AITW)
This is a mirror of Google's Android in the Wild (AITW) dataset, re-hosted on Hugging Face for easier community access.
Original Source
Paper: Android in the Wild: A Large-Scale Dataset for Android Device Control
Original Repository: google-research/google-research/tree/master/android_in_the_wild
Dataset Description
Android in the Wild (AITW) is a large-scale dataset for Android device control. It contains human demonstrations of… See the full description on the dataset page: https://huggingface.co/datasets/leosltl/Android-in-the-Wild.androidlife-public
AndroidLife — Public 60-task preview + run trajectories
This dataset publishes the public 60-task AndroidLife sample (task definitions +
sidecars) together with the full per-run execution artifacts from real-phone
benchmarks (OnePlus CPH2423 via ADB/MobileRun + real LLMs).
The full 530-task corpus lives in
YuvrajSingh9886/androidlife-530.
Dataset preview (tasks)
Open the Dataset Viewer above (config tasks) for a table of all 60 public
tasks: task_id, day… See the full description on the dataset page: https://huggingface.co/datasets/YuvrajSingh9886/androidlife-public.androidlife-530
AndroidLife-530 — Android agent benchmark (real phone, real LLM)
AndroidLife runs Android agent tasks against a real phone (via ADB/MobileRun)
and a real LLM, and grades the agent on reaching a verifiable device end-state.
This repo ships the 530-task corpus plus everything needed to reproduce runs.
Benchmark, or template — your call. The 530 tasks are an extended version
of the benchmark, usable as a larger evaluation set for further benchmarking of
models beyond the 60-task… See the full description on the dataset page: https://huggingface.co/datasets/YuvrajSingh9886/androidlife-530.androidlife-trajectories
AndroidLife300 — trajectory replays
Public trajectory assets for the AndroidLife Android-agent benchmark, served to the published site via resolve/main URLs.
index.json — per-task availability manifest (model, success, steps, gif/data URLs)
data/trajectories/<set>/<day>/<task>.json — condensed step streams (thoughts, tool calls, screenshots)
trajectories/<set>/<day>/<task>/trajectory.gif + screenshots/ — screen replay + per-step frames
Regenerate/republish with uv run python… See the full description on the dataset page: https://huggingface.co/datasets/YuvrajSingh9886/androidlife-trajectories.
