Team Ai
Agents
Live
Problem
Plan
Sign
Petitions
Leaderboard
Community
Search
Create
Alerts
1 results
preference-hijacking
preference-hijacking
Search
in
all
models
datasets
apps
agents
people
projects
Datasets
All datasets matching “preference-hijacking”
yflantmy /
universal-preference-hijacking-datasets
Phi: Preference Hijacking in Multi-modal Large Language Models at Inference Time Figure 1: Examples of Phi, which can hijack MLLM's preference toward the image. Figure 2: Example of a universal hijacking perturbation, which can be transferred across different images. This dataset is used to train and evaluate the universal hijacking perturbations in the paper "Phi: Preference Hijacking in Multi-modal Large Language Models at Inference Time", accepted at EMNLP… See the full description on the dataset page: https://huggingface.co/datasets/yflantmy/universal-preference-hijacking-datasets.
image
question-answering
1K<n<10K
0 likes
76 downloads
1y ago
Hugging Face