WiNGPT-3.0 Technical Report

Boqin Zhuang; Chenxiao Song; Huitong Lu; Jiacheng Qiao; Mingqian Liu; Mingxing Yu; Ping Hong; Rui Li; Xiaoxia Song; Xiangjun Xu; Xu Chen; Yaoyao Ma; Yujie Gao

arXiv:2505.17387·cs.CL·June 6, 2025

WiNGPT-3.0 Technical Report

Boqin Zhuang, Chenxiao Song, Huitong Lu, Jiacheng Qiao, Mingqian Liu, Mingxing Yu, Ping Hong, Rui Li, Xiaoxia Song, Xiangjun Xu, Xu Chen, Yaoyao Ma, Yujie Gao

PDF

1 Repo

TL;DR

WiNGPT-3.0 is a 32-billion parameter language model designed to improve medical reasoning and clinical applicability, demonstrating strong performance and the effectiveness of reinforcement learning with limited data for healthcare AI.

Contribution

The paper introduces WiNGPT-3.0, a large language model tailored for medical reasoning, utilizing a multi-stage training pipeline with reinforcement learning to enhance clinical accuracy.

Findings

01

Achieved 66.6 on MedCalc and 87.1 on MedQA-USMLE

02

Improved clinical reasoning score from 58.1 to 62.5 with targeted training

03

Reinforcement learning effective with only a few thousand examples

Abstract

Current Large Language Models (LLMs) exhibit significant limitations, notably in structured, interpretable, and verifiable medical reasoning, alongside practical deployment challenges related to computational resources and data privacy. This report focused on the development of WiNGPT-3.0, the 32-billion parameter LLMs, engineered with the objective of enhancing its capacity for medical reasoning and exploring its potential for effective integration within healthcare IT infrastructures. The broader aim is to advance towards clinically applicable models. The approach involved a multi-stage training pipeline tailored for general, medical, and clinical reasoning. This pipeline incorporated supervised fine-tuning (SFT) and reinforcement learning (RL), leveraging curated Long Chain-of-Thought (CoT) datasets, auxiliary reward models, and an evidence-based diagnostic chain simulation.…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

winninghealth/WiNGPT3
noneOfficial

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.