Text2SQL
Code-Llama-2-13B-instruct-text2sql-GGUFArctic-Text2SQL-R1-7BArctic-Text2SQL-R1-7B-GGUFArctic-Text2SQL-R1-7B-i1-GGUFXeAI_-_LLaMa_3.2_3B_Instruct_Text2SQL_Legacy-ggufhuyhoangt2201_-_llama3.2-1b-text2SQL-finetuned-multitableJidouka2.1-ggufSiriusAI-Text2SQL-27b-agentic-v1-GGUFNESPED-GEN-Llama-3.2-text2SQL-v0-GGUF
Datasets
All datasets matching “Text2SQL”synthetic-text2sqlEmploying the MTEB evaluation framework's dataset version, utilize the code below for assessment:
import mteb
import logging
from sentence_transformers import SentenceTransformer
from mteb import MTEB
logger = logging.getLogger(__name__)
model_name = 'intfloat/e5-base-v2'
model = SentenceTransformer(model_name)
tasks = mteb.get_tasks(
tasks=[
"AppsRetrieval",
"CodeFeedbackMT",
"CodeFeedbackST",
"CodeTransOceanContest",
"CodeTransOceanDL"… See the full description on the dataset page: https://huggingface.co/datasets/CoIR-Retrieval/synthetic-text2sql.text2sql-eval-results
Text2SQL Evaluation Toolkit — Pre-computed Results
Pre-computed inference and evaluation results produced by the
IBM/text2sql-eval-toolkit
across six text-to-SQL benchmarks.
These artefacts power the toolkit's evaluation dashboard and analysis scripts
without requiring you to re-run multi-hour inference pipelines.
Quick start
Install the toolkit and fetch all results (~3.8 GB):
pip install text2sql-eval-toolkit
text2sql-eval-toolkit results fetch
Fetch a single… See the full description on the dataset page: https://huggingface.co/datasets/text2sql-eval-toolkit/text2sql-eval-results.text2sql-loop-engineering-data
text2sql-loop-engineering course data
Data package for the course repository https://github.com/mushan-shine/text2sql-loop-engineering .
Load it into your own Databricks workspace from a notebook with scripts/load_course_data.py.
Contents: the BEAVER dw data warehouse (97 tables) and the benchmark tables prepared in step 1
(questions, table metadata, gold results validated against BEAVER's official MySQL engine), plus the
dev set and step-1 result files.
Source and… See the full description on the dataset page: https://huggingface.co/datasets/MuShan795/text2sql-loop-engineering-data.RBAC-Text2SQL-Benchmark
RBAC-Text2SQL Benchmark
Role-conditioned Text-to-SQL instances for evaluating whether LLMs generate SQL that
respects Role-Based Access Control (RBAC) constraints. Each instance pairs a natural
language question with a role policy; the model must either produce a correct SQL query
that touches only authorized resources, or refuse with Sorry, I cannot answer.
Code, evaluation harness, and reproduction instructions:
https://github.com/2020dfff/RBAC-Text2SQL-Benchmark… See the full description on the dataset page: https://huggingface.co/datasets/sharkiefff/RBAC-Text2SQL-Benchmark.synthetic-text2sql-qrels
Dataset Card for "synthetic-text2sql-qrels"
More Information needed
synthetic-text2sql-queries-corpusEmploying the CoIR evaluation framework's dataset version, utilize the code below for assessment:
import coir
from coir.data_loader import get_tasks
from coir.evaluation import COIR
from coir.models import YourCustomDEModel
model_name = "intfloat/e5-base-v2"
# Load the model
model = YourCustomDEModel(model_name=model_name)
# Get tasks
#all task ["codetrans-dl","stackoverflow-qa","apps","codefeedback-mt","codefeedback-st","codetrans-contest","synthetic-
# text2sql","cosqa","codesearchnet"… See the full description on the dataset page: https://huggingface.co/datasets/CoIR-Retrieval/synthetic-text2sql-queries-corpus.
