lima
Nous-Capybara-limarpv3-34B-i1-GGUFMixtral-8x7B-Instruct-v0.1-LimaRP-ZLoss-i1-GGUFllama-2-13B-chat-limarp-v2-merged-GGUFairoboros-2.2.1-limarpv3-y34b-i1-GGUFMixtral-8x7B-Instruct-v0.1-LimaRP-ZLoss-GGUFairoboros-2.2.1-limarpv3-y34b-GGUFNous-Capybara-limarpv3-34B-GGUFF1-Chimera-Hybrid-LimaRP-8B-i1-GGUF
Datasets
All datasets matching “lima”limaA high-quality dataset for efficient instruction tuning.LimaRP-G4-26B-KCPPPJMixers-Dev/lemonilia_LimaRP-Simple-CustomShareGPT with responses generated using non-thinking bartowski/google_gemma-4-26B-A4B-it-GGUF/google_gemma-4-26B-A4B-it-Q4_K_M.gguf.
Used default recommended generation settings: temp=1, top_k=64, top_p=0.95.
thinking-cap-tier-lima-dense
Thinking Cap Tier Curricula — LIMA Hyper-Dense Reasoning Alignment Suite (TCS v4)
[!IMPORTANT]
Dataset Release v1.2 (Sept 2026) — Clean Delimiters & Zero-Padding Architecture:
In v1.2, all 5,500 SFT and 2,000 SimPO records have undergone a complete token purge:
Zero <|pad|> batch residues: 100% eliminated across all records.
Zero reasoning leakage into final answers: Deliberation stays strictly inside <think>...</think>, and answers provide direct conclusions.
Native ChatML… See the full description on the dataset page: https://huggingface.co/datasets/Davd-b01/thinking-cap-tier-lima-dense.LimAgents
LimAgents Data
This dataset contains scientific paper metadata and extracted limitation information prepared for use with LLM Agents.The data comes from NeurIPS 2021–2022 papers and related OpenReview reviews, enriched with Cited in and Cited by information.
Dataset Structure
The repository contains two main directories:
1. NeurIPS_21_22_Lim_OPR_with_cited_in_by_papers
This directory includes one JSON file per paper. Each file contains:
title:… See the full description on the dataset page: https://huggingface.co/datasets/IbrahimAlAzhar/LimAgents.LimaRP
LIMA ERP data (LimaRP)
Following the principles highlighted in arXiv:2305.11206 by Zhou et al.
and replicated in some aspects by Kaiokendev with SuperHOT,
the archive in this repository contains about 2000 manually selected and curated 1-on-1 human-human
roleplaying conversations and associated LLM-generated persona and scenario data. The RP conversations
all feature only two human participants, although occasionally the participants may play the role of more
than one character.
The… See the full description on the dataset page: https://huggingface.co/datasets/lemonilia/LimaRP.LimAgents_limitation_data_scientific_papers_with_cited_papers
LimAgents Data
This dataset contains scientific paper metadata and extracted limitation information prepared for use with LLM Agents.The data comes from NeurIPS 2021–2022 papers and related OpenReview reviews, enriched with Cited in and Cited by information.
Dataset Structure
The repository contains two main directories:
1. NeurIPS_21_22_Lim_OPR_with_cited_in_by_papers
This directory includes one JSON file per paper. Each file contains:
title: Original paper… See the full description on the dataset page: https://huggingface.co/datasets/iaadlab/LimAgents_limitation_data_scientific_papers_with_cited_papers.
