Bioinformatics
mst_geBioinformatics
🔭 Overview
R2MED: First Reasoning-Driven Medical Retrieval Benchmark
R2MED is a high-quality, high-resolution synthetic information retrieval (IR) dataset designed for medical scenarios. It contains 876 queries with three retrieval tasks, five medical scenarios, and twelve body systems.
Dataset
#Q
#D
Avg. Pos
Q-Len
D-Len
Biology
103
57359
3.6
115.2
83.6
Bioinformatics77
47473
2.9
273.8
150.5
Medical Sciences
88
34810
2.8
107.1
122.7
MedXpertQA-Exam
97… See the full description on the dataset page: https://huggingface.co/datasets/R2MED/Bioinformatics.master-thesis-method2-patstackexchange_bioinformaticsBioinformatics-Toolkit-HYA
Bioinformatics Toolkit v1.0
Bioinformatics is plagued by software dependency drama. The Bioinformatics Toolkit makes it easy to set up a computational software environment. This allows researchers and students to focus on science rather than software troubleshooting.
It's a pre-built bioinformatics env for macOS that bundles the Python 3.12 interpreter and the uv binary. Packages are auto downloaded during the first run. The app is launched with a double-click.
This is a Hybrid App… See the full description on the dataset page: https://huggingface.co/datasets/vbookshelf/Bioinformatics-Toolkit-HYA.bioinformatics-qa-dataset
Bioinformatics QA Dataset
A curated question-answer dataset for bioinformatics and computational biology model training.
Summary
Total examples: 5880
Unique topics: 65
Columns: id, topic, question, answer
Source format: CSV converted to Hugging Face Dataset
Intended Use
This dataset is intended for:
Instruction tuning and domain adaptation for biomedical and bioinformatics LLMs
QA benchmarking in life-science terminology
Prompt-response training pipelines… See the full description on the dataset page: https://huggingface.co/datasets/yashm/bioinformatics-qa-dataset.
