Team Ai
20 results

Tweet

cardiffnlp /tweet_eval Dataset Card for tweet_eval Dataset Summary TweetEval consists of seven heterogenous tasks in Twitter, all framed as multi-class tweet classification. The tasks include - irony, hate, offensive, stance, emoji, emotion, and sentiment. All tasks have been unified into the same benchmark, with each dataset presented in the same format and with fixed training, validation and test splits. Supported Tasks and Leaderboards text_classification: The dataset can be… See the full description on the dataset page: https://huggingface.co/datasets/cardiffnlp/tweet_eval.texttext-classification100K<n<1M152 likes25k downloads3y agoHugging Facemteb /tweet_sentiment_extraction TweetSentimentExtractionClassification An MTEB dataset Massive Text Embedding Benchmark Task category t2c Domains Social, Written Reference https://www.kaggle.com/competitions/tweet-sentiment-extraction/overview How to evaluate on this task You can evaluate an embedding model on this dataset using the following code: import mteb task = mteb.get_tasks(["TweetSentimentExtractionClassification"]) evaluator = mteb.MTEB(task) model =… See the full description on the dataset page: https://huggingface.co/datasets/mteb/tweet_sentiment_extraction.texttext-classification10K<n<100K38 likes7.6k downloads1y agoHugging FaceDDSC /angry-tweets Dataset Card for AngryTweets Dataset Summary This dataset consists of anonymised Danish Twitter data that has been annotated for sentiment analysis through crowd-sourcing. All credits go to the authors of the following paper, who created the dataset: Pauli, Amalie Brogaard, et al. "DaNLP: An open-source toolkit for Danish Natural Language Processing." Proceedings of the 23rd Nordic Conference on Computational Linguistics (NoDaLiDa). 2021 Supported Tasks and… See the full description on the dataset page: https://huggingface.co/datasets/DDSC/angry-tweets.texttext-classification1K<n<10K4 likes3.1k downloads3y agoHugging Facecardiffnlp /tweet_topic_multilingual Dataset Card for "cardiffnlp/tweet_topic_multilingual" Dataset Summary This is the official repository of X-Topic (Multilingual Topic Classification in X: Dataset and Analysis, EMNLP 2024), a topic classification dataset based on X (formerly Twitter), featuring 19 topic labels. The classification task is multi-label, with tweets available in four languages: English, Japanese, Spanish, and Greek. The dataset comprises 4,000 tweets (1,000 per language), collected between… See the full description on the dataset page: https://huggingface.co/datasets/cardiffnlp/tweet_topic_multilingual.texttext-classification10K<n<100K3 likes2.6k downloads1y agoHugging FaceSetFit /tweet_eval_stance_abortiontextn<1K0 likes2.2k downloads4y agoHugging Facecardiffnlp /tweet_sentiment_multilingual Dataset Card for cardiffnlp/tweet_sentiment_multilingual Dataset Summary Tweet Sentiment Multilingual consists of sentiment analysis dataset on Twitter in 8 different lagnuages. arabic english french german hindi italian portuguese spanish Supported Tasks and Leaderboards text_classification: The dataset can be trained using a SentenceClassification model from HuggingFace transformers. Dataset Structure Data Instances An instance from… See the full description on the dataset page: https://huggingface.co/datasets/cardiffnlp/tweet_sentiment_multilingual.texttext-classification10K<n<100K24 likes2k downloads4y agoHugging Face