MedMobile: A mobile-sized language model with expert-level clinical capabilities
Vishwanath, Krithik, Stryker, Jaden, Alyakin, Anton, Alber, Daniel Alexander, Oermann, Eric Karl
–arXiv.org Artificial Intelligence
Language models (LMs) have demonstrated expert-level reasoning and recall abilities in medicine. However, computational costs and privacy concerns are mounting barriers to wide-scale implementation. We introduce a parsimonious adaptation of phi-3-mini, MedMobile, a 3.8 billion parameter LM capable of running on a mobile device, for medical applications. We demonstrate that MedMobile scores 75.7% on the MedQA (USMLE), surpassing the passing mark for physicians (~60%), and approaching the scores of models 100 times its size. We subsequently perform a careful set of ablations, and demonstrate that chain of thought, ensembling, and fine-tuning lead to the greatest performance gains, while unexpectedly retrieval augmented generation fails to demonstrate significant improvements
arXiv.org Artificial Intelligence
Oct-11-2024
- Country:
- North America > United States
- Texas > Travis County
- Austin (0.14)
- New York > New York County
- New York City (0.05)
- Missouri > St. Louis County
- St. Louis (0.04)
- Texas > Travis County
- Europe > Italy
- Calabria > Catanzaro Province > Catanzaro (0.04)
- North America > United States
- Genre:
- Research Report (1.00)
- Industry:
- Health & Medicine
- Therapeutic Area > Oncology (1.00)
- Diagnostic Medicine (1.00)
- Health & Medicine
- Technology: