medical-o1-reasoning-SFT FreedomIntelligence

News [2025/04/22] We split the data and kept only the medical SFT dataset (medical_o1_sft.json). The file medical_o1_sft_mix.json contains a mix of medical and general instruction data. [2025/02/22] We released the distilled dataset from Deepseek-R1 based on medical verifiable problems. You can use it to initialize your models with the reasoning chain from Deepseek-R1. [2024/12/25] We open-sourced the medical reasoning dataset for SFT, built on medical verifiable problems and an… See the full description on the dataset page: https://huggingface.co/datasets/FreedomIntelligence/medical-o1-reasoning-SFT.

类型
dataset
许可
apache-2.0
语言
en
下载量
15,617
点赞
1,167
访问
public
文件
0

标签

  • 指令微调
  • 文本生成
  • 问答
  • 蒸馏
  • 大语言模型
  • 医学推理
  • 合成数据
  • 对话数据

摘要

该数据集用于微调医疗大模型 HuatuoGPT-o1,支持医学高级推理。数据基于可验证的医学问题由 GPT-4o 构建,并通过医学验证器校验,还包含从 DeepSeek-R1 蒸馏得到的推理链数据。适用于医学问答、复杂医学推理等医疗领域大模型的指令微调场景。

README

--- license: apache-2.0 task_categories: - question-answering - text-generation language: - en - zh tags: - medical - biology configs: - config_name: en data_files: medical_o1_sft.json - config_name: zh data_files: medical_o1_sft_Chinese.json - config_name: en_mix data_files: medical_o1_sft_mix.jso…

查看完整页面 · 查看原文