数据集导航
医疗数据集导航 - 全部数据集(第 162 页)
汇集 20 大疾病系统 5035 个医疗数据集,支持搜索与分类筛选,助力医疗AI研发选型
全部 5035
肿瘤 579
神经系统疾病 507
心血管系统疾病 462
耳鼻咽喉疾病 451
眼病 412
皮肤与结缔组织疾病 399
泌尿生殖系统疾病 362
内分泌系统疾病 352
医学文本与大模型 285
呼吸道疾病 245
口颌系统疾病 219
消化系统疾病 200
免疫系统疾病 137
通用医学数据集 136
血液与淋巴系统疾病 106
感染 96
创伤与损伤 66
肌肉骨骼疾病 14
化学诱发性障碍 4
环境因素所致疾病 3
共 5035 个数据集 · 第 162 / 168 页
| 数据集名称 | 系统分类 | 病种 | 数据模态 | 任务类型 | 描述 | 来源 | 原始地址 | 操作 |
|---|---|---|---|---|---|---|---|---|
| AnesSuite | 医学文本与大模型 | 文本 | 麻醉学问答 | 4427 anesthesiology MCQ items focused on complex decision-making, specialized knowledge QA 2025 | GitHub | https://arxiv.org/pdf/2504.02404 | 访问 | |
| TracSum | 医学文本与大模型 | 文本 | 医疗摘要 | 500 abstracts resulting in 3500 summary-citation traceable pairs, aspect-based summarization EMNLP 2025 | GitHub | https://arxiv.org/pdf/2508.13798 | 访问 | |
| MedAgentBoard | 通用医学数据集 | 多模态 | 多智能体协作 | 8 benchmark categories for multi-agent reasoning tasks, text/image/EHR multi-agent collaboration 2025 | Project Page | https://arxiv.org/pdf/2505.12371 | 访问 | |
| HEAL-MedVQA | 医学文本与大模型 | 多模态 | 视觉问答 | 11000+ samples requiring localization prior to medical answering, grounded medical VQA IJCAI 2025 | Project Page | https://arxiv.org/pdf/2505.00744 | 访问 | |
| BRIDGE | 医学文本与大模型 | 文本 | 多任务评估 | 1.4 million samples covering 87 tasks in 9 languages, real-world clinical practice text 2025 | Project Page | https://arxiv.org/pdf/2504.19467 | 访问 | |
| LLMEval-Med | 医学文本与大模型 | 文本 | 临床问答验证 | ~1000 real-world cases validated via physician-in-the-loop audits, clinical QA validation 2025 | GitHub | https://arxiv.org/pdf/2506.04078 | 访问 | |
| CSEDB | 医学文本与大模型 | 文本 | 安全性评估 | 30 criteria across 26 specialties based on expert physician consensus, safety-effectiveness eval 2025 | GitHub | https://arxiv.org/pdf/2507.23486 | 访问 | |
| SSG-VQA | 医学文本与大模型 | 多模态 | 手术视觉问答 | 1300+ scene graph samples focused on instrument-tissue interaction, surgical VQA 2025 | GitHub | https://arxiv.org/pdf/2506.06232 | 访问 | |
| PET2Rep | 医学文本与大模型 | 多模态 | 报告生成 | 565 whole-body paired PET/CT data combinations with detailed radiology reports, report generation 2025 | GitHub | https://arxiv.org/pdf/2508.04062 | 访问 | |
| MedTVT-QA | 医学文本与大模型 | 多模态 | 视觉问答 | 3232 VQA pairs encompassing 15 medical specialties, 3 categories and 8 sub-categories clinical tasks ACL 2025 | GitHub | https://arxiv.org/pdf/2506.18512 | 访问 | |
| MediConfusion | 医学文本与大模型 | 多模态 | 视觉问答 | 176 confusing pairs of two images sharing same question but different correct answers, VQA reliability ICLR 2025 | Hugging Face | https://arxiv.org/abs/2409.15477 | 访问 | |
| GMAI-MMBench | 医学文本与大模型 | 多模态 | 视觉问答 | 26K QA pairs, 38 modality types, comprehensive multimodal evaluation benchmark NeurIPS 2024 | Hugging Face | https://huggingface.co/datasets/OpenGVLab/GMAI-MMBench | 访问 | |
| PathMMU | 医学文本与大模型 | 多模态 | 病理推理 | 33428 QAs, 24067 images, massive multimodal expert-level pathology understanding and reasoning 2024 | Hugging Face | https://huggingface.co/datasets/jamessyx/PathMMU | 访问 | |
| CARES | 眼病 | 多模态 | 可信度评估 | 41K QA pairs, trustworthiness benchmark for medical vision language models, open/closed QA NeurIPS 2024 | GitHub | https://github.com/richard-peng-xia/CARES | 访问 | |
| MultiMedEval | 眼病 | 多模态 | 多任务评估 | 6 tasks, 23 datasets, benchmark and toolkit for evaluating medical vision-language models 2024 | GitHub | https://github.com/corentin-ryr/MultiMedEval | 访问 | |
| medical-o1-reasoning-SFT | 医学文本与大模型 | 文本 | 医学推理 | 19.7k QA pairs, HuatuoGPT-o1 medical complex reasoning, medical VQA & reasoning ACL 2025 | Hugging Face | https://huggingface.co/datasets/FreedomIntelligence/medical-o1-reasoning-SFT | 访问 | |
| ReasonMed | 医学文本与大模型 | 文本 | 医学推理 | 370K multi-agent generated dataset from 1.75M CoT paths, medical reasoning, QA, chain-of-thought 2025 | Hugging Face | https://arxiv.org/abs/2506.09513 | 访问 | |
| Lingshu Train | 通用医学数据集 | 多模态 | 多模态训练 | ~9.3M training samples from 60+ datasets, generalist foundation model for unified multimodal medical understanding 2025 | Project Page | https://arxiv.org/pdf/2506.07044 | 访问 | |
| MedEvalKit Lingshu test | 通用医学数据集 | 多模态 | 基准评测 | 152066 evaluation samples from 16 benchmark datasets, VQA, report generation, medical text QA 2025 | GitHub | https://arxiv.org/pdf/2507.04289 | 访问 | |
| MIRIAD | 医学文本与大模型 | 文本 | 医学问答 | 5.82M/4.48M medical query-response pairs, augmenting LLMs with medical query-response, RAG & hallucination detection 202 | Hugging Face | https://arxiv.org/abs/2506.06091 | 访问 | |
| ClinBench-HPB | 医学文本与大模型 | 文本 | 临床问答 | 3535 MCQs and 337 clinical cases covering 465+ hepato-pancreato-biliary diseases 2025 | Project Page | https://arxiv.org/abs/2506.00095 | 访问 | |
| SurgVLM-DB | 眼病 | 视频 | 手术问答 | 1.81M frames, 7.79M dialogues, large-scale multimodal surgical database 2025 | GitHub | https://github.com/jinlab-imvr/SurgVLM | 访问 | |
| EndoBench | 消化系统疾病 | 多模态 | 内镜分析 | 6832 clinically validated VQA samples, gastroscopy/colonoscopy/capsule/surgical endoscopy, 12 tasks 2025 | Hugging Face | https://arxiv.org/abs/2505.23601v1 | 访问 | |
| MedXpertQA | 医学文本与大模型 | 多模态 | 专家级问答 | 4460 questions (text 2455 / image 2005), expert-level medical reasoning and understanding ICML 2025 | Hugging Face | https://huggingface.co/datasets/TsinghuaC3I/MedXpertQA | 访问 | |
| MedCaseReasoning | 医学文本与大模型 | 文本 | 诊断推理 | 14489 QA pairs, diagnostic reasoning from clinical case reports 2025 | GitHub | https://github.com/kevinwu23/Stanford-MedCaseReasoning | 访问 | |
| DrVD-Bench | 眼病 | 多模态 | 诊断推理 | 7789 image-text QA pairs, vision-language models reasoning like human doctors in medical image diagnosis 2025 | GitHub | https://github.com/Jerry-Boss/DrVD-Bench | 访问 | |
| MedS-Ins | 医学文本与大模型 | 文本 | 指令微调 | 5M samples, 19K instructions, versatile LLMs for medicine instruction tuning 2025 | Hugging Face | https://huggingface.co/datasets/Henrychur/MedS-Ins | 访问 | |
| MedS-Bench | 医学文本与大模型 | 文本 | 临床任务评估 | 11 categories of clinical tasks, benchmark for medical LLMs 2025 | Hugging Face | https://huggingface.co/datasets/Henrychur/MedS-Bench | 访问 | |
| MM-Skin | 皮肤与结缔组织疾病 | 多模态 | 皮肤病问答 | Image-text dataset derived from textbooks for dermatology VLM enhancement 2025 | GitHub | https://github.com/ZwQ803/MM-Skin | 访问 | |
| AlphaMed19K | 医学文本与大模型 | 文本 | 医学推理 | 19K QA pairs, medical LLM reasoning with minimalist rule-based RL 2025 | Hugging Face | https://huggingface.co/datasets/che111/AlphaMed19K | 访问 |