Team Ai
20 results

text-2-sql

CoIR-Retrieval /synthetic-text2sqlEmploying the MTEB evaluation framework's dataset version, utilize the code below for assessment: import mteb import logging from sentence_transformers import SentenceTransformer from mteb import MTEB logger = logging.getLogger(__name__) model_name = 'intfloat/e5-base-v2' model = SentenceTransformer(model_name) tasks = mteb.get_tasks( tasks=[ "AppsRetrieval", "CodeFeedbackMT", "CodeFeedbackST", "CodeTransOceanContest", "CodeTransOceanDL"… See the full description on the dataset page: https://huggingface.co/datasets/CoIR-Retrieval/synthetic-text2sql.text100K<n<1M0 likes1.5k downloads2y agoHugging Facetext2sql-eval-toolkit /text2sql-eval-results Text2SQL Evaluation Toolkit — Pre-computed Results Pre-computed inference and evaluation results produced by the IBM/text2sql-eval-toolkit across six text-to-SQL benchmarks. These artefacts power the toolkit's evaluation dashboard and analysis scripts without requiring you to re-run multi-hour inference pipelines. Quick start Install the toolkit and fetch all results (~3.8 GB): pip install text2sql-eval-toolkit text2sql-eval-toolkit results fetch Fetch a single… See the full description on the dataset page: https://huggingface.co/datasets/text2sql-eval-toolkit/text2sql-eval-results.imagen<1K0 likes251 downloads23d agoHugging Facetrl-lab /SQaLe-2-text-to-SQL-SchemasSQaLe: schemas and databases Project page · Questions and SQL · Trained models · Python library · Citation This dataset holds the 9,259 populated databases of SQaLe, a large semi-synthetic text-to-SQL dataset grounded in real-world database schemas, introduced in the paper SQaLe: a large realistic dataset to empower small specialised text-to-SQL models. Each row is one database: its DDL, extended from a real schema in SchemaPile, and the generated rows of its tables. The… See the full description on the dataset page: https://huggingface.co/datasets/trl-lab/SQaLe-2-text-to-SQL-Schemas.texttext-generation1K<n<10K0 likes231 downloads4d agoHugging FaceCoIR-Retrieval /synthetic-text2sql-qrels Dataset Card for "synthetic-text2sql-qrels" More Information needed text100K<n<1M0 likes172 downloads2y agoHugging Facetrl-lab /SQaLe-2-text-to-SQL-QueriesSQaLe: questions and SQL Project page · Schemas and databases · Trained models · Python library · Citation SQaLe is a large semi-synthetic text-to-SQL dataset grounded in real-world database schemas, introduced in the paper SQaLe: a large realistic dataset to empower small specialised text-to-SQL models. It pairs 1,408,056 natural-language questions with 176,761 distinct SQL queries over 9,259 populated SQLite databases. The schemas come from SchemaPile, a collection of database… See the full description on the dataset page: https://huggingface.co/datasets/trl-lab/SQaLe-2-text-to-SQL-Queries.texttext-generation100K<n<1M0 likes166 downloads4d agoHugging Facefahmiaziz /text2sql-dataset Dataset We built this dataset from several sources combining examples from: Wikisql Bird Spider Synthetic SQL samples This dataset has been cleaned and filtered by: Removing DDL/DML examples (INSERT, UPDATE, DELETE, etc.) De-duplicating examples based on hashing semantics of SQL and queries Filtering only SELECT-style analytical queries texttext-generation100K<n<1M1 likes161 downloads1y agoHugging Face