Team Ai
20 results

telecom

Telecom-Paris /iamd_v0 Internet Archive Music Dataset (IAMD v0) ~4.2M thirty-second music segments (34,469 hours) sourced from Creative-Commons audio on the Internet Archive, each paired with machine-generated natural-language captions and the original item metadata. Segments 4.2M Audio 34k hours Segment length 30 s nominal (mean 29.22 s) Format MP3, 320 kbps CBR, native channels + sample rate Shards 2,320 Parquet files Download size 4.53 TB Loading A… See the full description on the dataset page: https://huggingface.co/datasets/Telecom-Paris/iamd_v0.audioaudio-classification1M<n<10M9 likes3k downloads7d agoHugging FaceAliMaatouk /TelecomTS 📡 TelecomTS: A Multi-Modal Telecom Dataset TelecomTS is a large-scale, high-resolution, multi-modal dataset derived from a 5G telecommunications testbed. It is the first public observability dataset to preserve deanonymized observability metrics with absolute scale information, encompassing by design various downstream tasks beyond forecasting such as anomaly detection, root-cause analysis, and multi-modal reasoning. Observability data, particularly in… See the full description on the dataset page: https://huggingface.co/datasets/AliMaatouk/TelecomTS.text10K<n<100K15 likes1.2k downloads5mo agoHugging Facetelecomadm1145 /sakuragpt_synthetic_ja_zh Dataset Card for skr_trans_distill 本数据集由 SakuraLLM 大模型生成机器翻译结果,主要用于模型蒸馏训练。数据集包含日文原文及其对应的中文翻译,适用于日译中任务的模型训练与蒸馏。 Dataset Details Dataset Description 本数据集使用 SakuraLLM 大模型对日文文本进行机器翻译,生成日译中的平行语料,主要用于知识蒸馏场景下的小模型训练。 Curated by: telecomadm1145 Shared by [optional]: telecomadm1145 Language(s) (NLP): Japanese (ja), Chinese (zh) License: MIT Dataset Sources [optional] Repository: telecomadm1145/skr_trans_distill Paper [optional]: [More… See the full description on the dataset page: https://huggingface.co/datasets/telecomadm1145/sakuragpt_synthetic_ja_zh.texttranslation10M<n<100M0 likes794 downloads2mo agoHugging FaceShashkovich /Telecommunication_SMS_time_series SMS Time series data for traffic and fraud forecasting. TeleWhale vendor collected data This dataset contains various time series from vendors. Shashkov A.A. Vendor A: 01.03.23-14.08.23 TS_*_all - Count of all SMS Vendor A: January TS_*_fraud - Count of fraud TS_*_all - Count of all SMS TS_*_hlrDelay - Mean values of hlr delay Vendor B: January 1-8 1-8_TS_*_fraud - Count of fraud 1-8_TS_*_all - Count of all SMS 1-8_TS_*_hlrDelay -… See the full description on the dataset page: https://huggingface.co/datasets/Shashkovich/Telecommunication_SMS_time_series.imagetime-series-forecasting3 likes285 downloads10mo agoHugging Facetalkmap /telecom-conversation-corpus Telecom 200k Dataset Overview This dataset consists of 200,000 synthetically generated conversations in a customer service setting for the telecom industry. There are two speakers: a customer, and an agent. texttext-generation1M<n<10M22 likes172 downloads3y agoHugging Facetelecomadm1145 /test13tabular1M<n<10M0 likes162 downloads1mo agoHugging Face