Team Ai
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01bitext /Bitext-customer-support-llm-chatbot-training-dataset Bitext - Customer Service Tagged Training Dataset for LLM-based Virtual Assistants Overview This hybrid synthetic dataset is designed to be used to fine-tune Large Language Models such as GPT, Mistral and OpenELM, and has been generated using our NLP/NLG technology and our automated Data Labeling (DAL) tools. The goal is to demonstrate how Verticalization/Domain Adaptation for the Customer Support sector can be easily achieved using our two-step approach to LLM… See the full description on the dataset page: https://huggingface.co/datasets/bitext/Bitext-customer-support-llm-chatbot-training-dataset.textquestion-answering10K<n<100K199 likes11k downloads2y agoHugging Face02Tobi-Bueck /customer-support-tickets Featuring Labeled Customer Emails and Support Responses 🔧 Synthetic IT Ticket Generator — Custom Dataset Create a dataset tailored to your own queues & priorities (no PII). 👉 Generate custom data Define your queues, priorities, language Need an on-prem AI to auto-classify tickets?→ Open Ticket AI There are 2 Versions of the dataset, the new version has more tickets, but only languages english and german. So please look at both files, to find what best fits… See the full description on the dataset page: https://huggingface.co/datasets/Tobi-Bueck/customer-support-tickets.texttext-classification10K<n<100K43 likes5.4k downloads4mo agoHugging Face03Kaludi /Customer-Support-Responsestextn<1K13 likes1.2k downloads4y agoHugging Face04bumblebearhug /insuff_supported_argumentstext10K<n<100K1 likes152 downloads3y agoHugging Face05aakash0017 /it-support-llmtext1K<n<10K3 likes76 downloads3y agoHugging Face06crossingminds /bitext_customer_support_mcq Multiple-Choice Formatted Version of Bitext Customer Support Dataset This repository contains a modified version of the Bitext - Customer Service Tagged Training Dataset for LLM-based Virtual Assistants dataset. The dataset has been transformed into a multiple-choice format aimed at training and evaluating intent classification models. Overview The original dataset consists of customer support instructions paired with labeled intents. In this variant, each… See the full description on the dataset page: https://huggingface.co/datasets/crossingminds/bitext_customer_support_mcq.texttext-classification10K<n<100K1 likes66 downloads2y agoHugging Face07FunDialogues /customer-service-robot-support This Dialogue Comprised of fictitious examples of dialogues between a customer encountering problems with a robotic arm and a technical support agent. Check out the example below: "id": 1, "description": "Robotic arm calibration issue", "dialogue": "Customer: My robotic arm seems to be misaligned. It's not picking objects accurately. What can I do? Agent: It appears that the arm may need recalibration. Please follow the instructions in the user manual to reset the calibration… See the full description on the dataset page: https://huggingface.co/datasets/FunDialogues/customer-service-robot-support.tabularquestion-answeringn<1K2 likes57 downloads3y agoHugging Face08gradium /tts-eval-customer-support-202608 TTS Hard Cases — Customer Support A compact, text-only evaluation set of hard cases for text-to-speech models, focused on the customer support use case. Published by Gradium. Most TTS benchmarks measure naturalness on ordinary prose, where modern models are already close to saturated. Production voice agents fail somewhere else: on the literal payload of a support call — the order number, the email address, the spelled-out surname, the date, the amount refunded. A voice that… See the full description on the dataset page: https://huggingface.co/datasets/gradium/tts-eval-customer-support-202608.texttext-to-speechn<1K1 likes56 downloads1mo agoHugging Face09Prady06 /customer-support-tickets Featuring Labeled Customer Emails and Support Responses 🔧 Synthetic IT Ticket Generator — Custom Dataset Create a dataset tailored to your own queues & priorities (no PII). 👉 Generate custom data Define your queues, priorities, language Need an on-prem AI to auto-classify tickets?→ Open Ticket AI There are 2 Versions of the dataset, the new version has more tickets, but only languages english and german. So please look at both files, to find what best fits your needs.… See the full description on the dataset page: https://huggingface.co/datasets/Prady06/customer-support-tickets.texttext-classification10K<n<100K1 likes53 downloads7mo agoHugging Face10anirudhhari /customer-support-ticketstabular1K<n<10K0 likes50 downloads5mo agoHugging Face11vasu1111 /customer-support-tickets Featuring Labeled Customer Emails and Support Responses 🔧 Synthetic IT Ticket Generator — Custom Dataset Create a dataset tailored to your own queues & priorities (no PII). 👉 Generate custom data Define your queues, priorities, language Need an on-prem AI to auto-classify tickets?→ Open Ticket AI There are 2 Versions of the dataset, the new version has more tickets, but only languages english and german. So please look at both files, to find what best fits your needs.… See the full description on the dataset page: https://huggingface.co/datasets/vasu1111/customer-support-tickets.texttext-classification10K<n<100K0 likes45 downloads6mo agoHugging Face12jonathansuru /customer_support_auto_completiontexttable-question-answering1K<n<10K2 likes34 downloads3y agoHugging Face13ljoaql /Bitext-customer-support-llm-chatbot-training-dataset Bitext - Customer Service Tagged Training Dataset for LLM-based Virtual Assistants Overview This hybrid synthetic dataset is designed to be used to fine-tune Large Language Models such as GPT, Mistral and OpenELM, and has been generated using our NLP/NLG technology and our automated Data Labeling (DAL) tools. The goal is to demonstrate how Verticalization/Domain Adaptation for the Customer Support sector can be easily achieved using our two-step approach to LLM… See the full description on the dataset page: https://huggingface.co/datasets/ljoaql/Bitext-customer-support-llm-chatbot-training-dataset.textquestion-answering10K<n<100K0 likes34 downloads6mo agoHugging Face14SID2702 /Bitext-customer-support-llm-chatbot-training-dataset Bitext - Customer Service Tagged Training Dataset for LLM-based Virtual Assistants Overview This hybrid synthetic dataset is designed to be used to fine-tune Large Language Models such as GPT, Mistral and OpenELM, and has been generated using our NLP/NLG technology and our automated Data Labeling (DAL) tools. The goal is to demonstrate how Verticalization/Domain Adaptation for the Customer Support sector can be easily achieved using our two-step approach to LLM… See the full description on the dataset page: https://huggingface.co/datasets/SID2702/Bitext-customer-support-llm-chatbot-training-dataset.textquestion-answering10K<n<100K0 likes32 downloads6mo agoHugging Face15mohamed0071 /Bitext-customer-support-llm-chatbot-training-dataset Bitext - Customer Service Tagged Training Dataset for LLM-based Virtual Assistants Overview This hybrid synthetic dataset is designed to be used to fine-tune Large Language Models such as GPT, Mistral and OpenELM, and has been generated using our NLP/NLG technology and our automated Data Labeling (DAL) tools. The goal is to demonstrate how Verticalization/Domain Adaptation for the Customer Support sector can be easily achieved using our two-step approach to LLM… See the full description on the dataset page: https://huggingface.co/datasets/mohamed0071/Bitext-customer-support-llm-chatbot-training-dataset.textquestion-answering10K<n<100K0 likes31 downloads4d agoHugging Face16harshgarg2006 /Bitext-customer-support-llm-chatbot-training-dataset Bitext - Customer Service Tagged Training Dataset for LLM-based Virtual Assistants Overview This hybrid synthetic dataset is designed to be used to fine-tune Large Language Models such as GPT, Mistral and OpenELM, and has been generated using our NLP/NLG technology and our automated Data Labeling (DAL) tools. The goal is to demonstrate how Verticalization/Domain Adaptation for the Customer Support sector can be easily achieved using our two-step approach to LLM… See the full description on the dataset page: https://huggingface.co/datasets/harshgarg2006/Bitext-customer-support-llm-chatbot-training-dataset.textquestion-answering10K<n<100K0 likes31 downloads3d agoHugging Face17unsloth /Support-Bot-Recommendationtextn<1K7 likes30 downloads1y agoHugging Face18manojroyal23 /customer-support-tickets Featuring Labeled Customer Emails and Support Responses 🔧 Synthetic IT Ticket Generator — Custom Dataset Create a dataset tailored to your own queues & priorities (no PII). 👉 Generate custom data Define your queues, priorities, language Need an on-prem AI to auto-classify tickets?→ Open Ticket AI There are 2 Versions of the dataset, the new version has more tickets, but only languages english and german. So please look at both files, to find what best fits your needs.… See the full description on the dataset page: https://huggingface.co/datasets/manojroyal23/customer-support-tickets.texttext-classification10K<n<100K0 likes29 downloads7mo agoHugging Face19abhi23457 /Bitext-customer-support-llm-chatbot-training-dataset Bitext - Customer Service Tagged Training Dataset for LLM-based Virtual Assistants Overview This hybrid synthetic dataset is designed to be used to fine-tune Large Language Models such as GPT, Mistral and OpenELM, and has been generated using our NLP/NLG technology and our automated Data Labeling (DAL) tools. The goal is to demonstrate how Verticalization/Domain Adaptation for the Customer Support sector can be easily achieved using our two-step approach to LLM… See the full description on the dataset page: https://huggingface.co/datasets/abhi23457/Bitext-customer-support-llm-chatbot-training-dataset.textquestion-answering10K<n<100K0 likes29 downloads1mo agoHugging Face20achrafgasmi /Bitext-customer-support-llm-chatbot-training-dataset Bitext - Customer Service Tagged Training Dataset for LLM-based Virtual Assistants Overview This hybrid synthetic dataset is designed to be used to fine-tune Large Language Models such as GPT, Mistral and OpenELM, and has been generated using our NLP/NLG technology and our automated Data Labeling (DAL) tools. The goal is to demonstrate how Verticalization/Domain Adaptation for the Customer Support sector can be easily achieved using our two-step approach to LLM… See the full description on the dataset page: https://huggingface.co/datasets/achrafgasmi/Bitext-customer-support-llm-chatbot-training-dataset.textquestion-answering10K<n<100K0 likes28 downloads8mo agoHugging Face21kshitij230 /emotional-supporttext1M<n<10M1 likes26 downloads1y agoHugging Face22sharmila122125 /customer-support-csattabular10K<n<100K0 likes26 downloads1y agoHugging Face23Mermeid /Bitext-customer-support-llm-chatbot-training-dataset Bitext - Customer Service Tagged Training Dataset for LLM-based Virtual Assistants Overview This hybrid synthetic dataset is designed to be used to fine-tune Large Language Models such as GPT, Mistral and OpenELM, and has been generated using our NLP/NLG technology and our automated Data Labeling (DAL) tools. The goal is to demonstrate how Verticalization/Domain Adaptation for the Customer Support sector can be easily achieved using our two-step approach to LLM… See the full description on the dataset page: https://huggingface.co/datasets/Mermeid/Bitext-customer-support-llm-chatbot-training-dataset.textquestion-answering10K<n<100K0 likes26 downloads5mo agoHugging Face24Nilesh987 /Bitext-customer-support-llm-chatbot-training-dataset Bitext - Customer Service Tagged Training Dataset for LLM-based Virtual Assistants Overview This hybrid synthetic dataset is designed to be used to fine-tune Large Language Models such as GPT, Mistral and OpenELM, and has been generated using our NLP/NLG technology and our automated Data Labeling (DAL) tools. The goal is to demonstrate how Verticalization/Domain Adaptation for the Customer Support sector can be easily achieved using our two-step approach to LLM… See the full description on the dataset page: https://huggingface.co/datasets/Nilesh987/Bitext-customer-support-llm-chatbot-training-dataset.textquestion-answering10K<n<100K0 likes24 downloads7mo agoHugging Face25karanverma19 /Multilingual_Customer_Support_Intent_Dataset_for_Indian_Contexts Multilingual Customer Support Intent Dataset for Indian Contexts Overview This dataset contains multilingual customer support queries across Indian contexts including ecommerce, banking, telecom, travel, and technology. Features Multilingual: English, Hindi, Hinglish, Punjabi Intent labeled (refund, payment_issue, account_issue, delivery_issue, complaint) Real-world customer queries Use Cases Customer support chatbots Intent classification models… See the full description on the dataset page: https://huggingface.co/datasets/karanverma19/Multilingual_Customer_Support_Intent_Dataset_for_Indian_Contexts.textn<1K0 likes23 downloads6mo agoHugging Face26olympiamantsiou /Bitext-customer-support-llm-chatbot-training-dataset Bitext - Customer Service Tagged Training Dataset for LLM-based Virtual Assistants Overview This hybrid synthetic dataset is designed to be used to fine-tune Large Language Models such as GPT, Mistral and OpenELM, and has been generated using our NLP/NLG technology and our automated Data Labeling (DAL) tools. The goal is to demonstrate how Verticalization/Domain Adaptation for the Customer Support sector can be easily achieved using our two-step approach to LLM… See the full description on the dataset page: https://huggingface.co/datasets/olympiamantsiou/Bitext-customer-support-llm-chatbot-training-dataset.textquestion-answering10K<n<100K0 likes23 downloads5mo agoHugging Face27Roy229 /fsfh6410-support-tickets-0c136b Support Tickets Support ticket records for triage processing. Each row contains the ticket id, customer name, category, priority, status, refund amount (where applicable), created date, and a short description. tickets.csv: the ticket records. policy.md: the triage policy to apply. This dataset is used by the customer support operations team for automated triage. The output report is published to a repository with the prefix fsfh6410-triage-report. textn<1K0 likes23 downloads2mo agoHugging Face28yunjaeys /Contextual_Response_Evaluation_for_ESL_and_ASD_Support Dataset Card for "Contextual Response Evaluation for ESL and ASD Support💜💬🌐"" Dataset Description 📖 Dataset Summary 📝 Curated by Eric Soderquist, this dataset is a collection of English prompts and responses generated by the Phi-2 model, designed to evaluate and improve NLP models for supporting ESL (English as a Second Language) and ASD (Autism Spectrum Disorder) user bases. Each prompt is paired with multiple AI-generated responses and evaluated using a… See the full description on the dataset page: https://huggingface.co/datasets/yunjaeys/Contextual_Response_Evaluation_for_ESL_and_ASD_Support.texttext-generationn<1K0 likes22 downloads3y agoHugging Face29aibabyshark /insurance_customer_support_conversation Dataset Card for Insurance Customer Support Conversation Dataset This is a synthetic dataset generated by ChatGPT-4o. Dataset Details Field Descriptions conversation: Type: StringDescription: The entire text of the conversation between the customer and the agent, including both parties' dialogues.Example: "Customer: Hello, I need to discuss the status of my insurance claim. It's been over a month, and I haven't received any updates. Agent: Good afternoon… See the full description on the dataset page: https://huggingface.co/datasets/aibabyshark/insurance_customer_support_conversation.tabularn<1K1 likes22 downloads2y agoHugging Face30gorges-haha /Bitext-customer-support-llm-chatbot-training-dataset Bitext - Customer Service Tagged Training Dataset for LLM-based Virtual Assistants Overview This hybrid synthetic dataset is designed to be used to fine-tune Large Language Models such as GPT, Mistral and OpenELM, and has been generated using our NLP/NLG technology and our automated Data Labeling (DAL) tools. The goal is to demonstrate how Verticalization/Domain Adaptation for the Customer Support sector can be easily achieved using our two-step approach to LLM… See the full description on the dataset page: https://huggingface.co/datasets/gorges-haha/Bitext-customer-support-llm-chatbot-training-dataset.textquestion-answering10K<n<100K0 likes22 downloads3mo agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.