GreyBrain School of AI
Back to the Pulse
GreyBrain Visual Model Card Verified · Hugging Face

Med42-v2 - A Suite of Clinically-aligned Large Language Models

m42-health/Llama3-Med42-8B

Med42-v2 is a suite of open-access clinical large language models (LLM) instruct and preference-tuned by M42 to expand access to medical knowledge. Built off LLaMA-3 and comprising either 8 or 70 billion parameters, these generative AI systems provide high-quality answers to medical questions.

Open on Hugging Face

From the model card

Taken straight from the Hugging Face model card, quoted not rewritten.

What it's for

The Med42-v2 suite of models is being made available for further testing and assessment as AI assistants to enhance clinical decision-making and access to LLMs for healthcare use. Potential use cases include:

  • Medical question answering
  • Patient record summarization
  • Aiding medical diagnosis
  • General health Q&A

Limitations & safe use

  • The Med42-v2 suite of models is not ready for real clinical use. Extensive human evaluation is undergoing as it is required to ensure safety.
  • Potential for generating incorrect or harmful information.
  • Risk of perpetuating biases in training data.

Use this suite of models responsibly! Do not rely on them for medical usage without rigorous safety testing.

Benchmarks

  • Med42-v2-70B outperforms GPT-4.0 in most of the MCQA tasks.
  • Med42-v2-70B achieves a MedQA zero-shot performance of 79.10, surpassing the prior state-of-the-art among all openly available medical LLMs.
  • Med42-v2-70B sits at the top of the Clinical Elo Rating Leaderboard.
ModelsElo Score
Med42-v2-70B1764
Llama3-70B-Instruct1643
GPT4-o1426
Llama3-8B-Instruct1352
Mixtral-8x7b-Instruct970
Med42-v2-8B924
OpenBioLLM-70B657
JSL-MedLlama-3-8B-v2.0447
Read the full model card

AI models can make mistakes. All details on this card are taken from the Hugging Face model card.