Team Ai
17 results

multi-context

artefactory /ledger-long-context-multi-kpi the LEDGER Long-Context Multi-KPI extraction datasets and benchmarks. OCR'd annual reports with ground-truth KPI values for financial information extraction benchmarking. Dataset Description This dataset pairs OCR-extracted annual report text (from DeepSeek OCR) with structured KPI ground-truth values. It is designed for evaluating LLM-based financial information extraction, retrieval, and needle-in-a-haystack tasks. Configs Config Reports… See the full description on the dataset page: https://huggingface.co/datasets/artefactory/ledger-long-context-multi-kpi.imagetable-question-answering1K<n<10K17 likes247 downloads3mo agoHugging Facenbtpj /multi-context-long-answer-datasettext1M<n<10M13 likes214 downloads4y agoHugging FaceTreeAILab /Multi-turn_Long-context_Benchmark_for_LLMs LoopServe: An Adaptive Dual-phase LLM Inference Acceleration System for Multi-Turn Dialogues Arxiv: https://www.arxiv.org/abs/2507.13681 Huggingface: https://huggingface.co/papers/2507.13681 Introduction LoopServe Multi-Turn Dialogue Benchmark is a comprehensive evaluation dataset comprising multiple diverse datasets designed to assess large language model performance in realistic conversational scenarios. Unlike traditional benchmarks that place queries only at the end… See the full description on the dataset page: https://huggingface.co/datasets/TreeAILab/Multi-turn_Long-context_Benchmark_for_LLMs.textquestion-answering1K<n<10K0 likes183 downloads1y agoHugging Facekothasuhas /multi_news_long_contexttext10K<n<100K1 likes21 downloads10mo agoHugging Facekothasuhas /multi_news_long_context_validationtext1K<n<10K0 likes14 downloads10mo agoHugging Facekothasuhas /multi_news_long_context_tokenized_n25722_ctx409610K<n<100K0 likes13 downloads10mo agoHugging Face