Team Ai
Datasetpublic

friedmanroy/ctrl-shift

Dataset Card for Control+Shift: Generating Controllable Distribution Shifts [arXiv], [GitHub] Curated by: Roy Friedman and Rhea Chowers This dataset is the one that accompanies the paper Control+Shift: Generating Controllable Distribution Shifts. Our data is based on CIFAR10 and ImageNet, using EDM to generate our data. We generated datasets for 3 types of distribution shift on CIFAR10 and ImageNet - so a total of 6 datasets. The types of distribution shifts are called overlap… See the full description on the dataset page: https://huggingface.co/datasets/friedmanroy/ctrl-shift.

sourceHugging Facemitupdated 2y agoView on Hugging Face
1likes96downloads
Dataset Card

Dataset Card for Control+Shift: Generating Controllable Distribution Shifts

[arXiv], [GitHub]

Curated by: Roy Friedman and Rhea Chowers

<!-- Provide a quick summary of the dataset. -->

This dataset is the one that accompanies the paper `Control+Shift`: Generating Controllable Distribution Shifts. Our data is based on CIFAR10 and ImageNet, using EDM to generate our data. We generated datasets for 3 types of distribution shift on CIFAR10 and ImageNet - so a total of 6 datasets. The types of distribution shifts are called overlap, extend and truncation.

<!-- ## Dataset Description -->

<!-- ## Uses -->

<!-- Address questions around how the dataset is intended to be used. -->

<!-- ### Direct Use -->

<!-- This section describes suitable use cases for the dataset. -->

<!-- [More Information Needed] -->

<!-- ## Dataset Structure -->

<!-- This section provides a description of the dataset fields, and additional information about the dataset structure such as criteria used to create the splits, relationships between data points, etc. -->

<!-- [More Information Needed] -->

<!-- ## Dataset Creation -->

<!-- ### Curation Rationale -->

<!-- Motivation for the creation of this dataset. -->

<!-- [More Information Needed] -->

<!-- ### Source Data -->

<!-- This section describes the source data (e.g. news text and headlines, social media posts, translated sentences, ...). -->

<!-- #### Data Collection and Processing -->

<!-- This section describes the data collection and processing process such as data selection criteria, filtering and normalization methods, tools and libraries used, etc. -->

<!-- [More Information Needed] -->

<!-- #### Who are the source data producers? -->

<!-- This section describes the people or systems who originally created the data. It should also include self-reported demographic or identity information for the source data creators if this information is available. -->

<!-- [More Information Needed] -->

<!-- ## Bias, Risks, and Limitations -->

<!-- This section is meant to convey both technical and sociotechnical limitations. -->

<!-- [More Information Needed] -->

<!-- ### Recommendations -->

<!-- This section is meant to convey recommendations with respect to the bias, risk, and technical limitations. -->

<!-- Users should be made aware of the risks, biases and limitations of the dataset. More information needed for further recommendations. -->