• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
AI
Reaction

Gemma 4 31B reportedly leads its size class in Medmarks medical AI results

SophontAI says its additional Medmarks results cover mid-size open-source models from Gemma, Qwen, Muse and Nemotron, with Gemma 4 31B leading the size class. A commentator argues the models still trail frontier systems in medicine.

Tanishq Mathew Abraham, Ph.D.TM
1 Source, 12d ago, first seen 12d ago

TLDR

SophontAI says it released additional results for Medmarks, its benchmark and leaderboard for LLM medical capabilities, and finds Gemma 4 31B leads among the mid-size open-source models tested. A commentator sharing the announcement argues that these models have not caught up to frontier models in medicine. They say comparisons of Qwen 3.8 27B with Opus 4.8 or 4.6 may hold for coding, but argue it falls well short of Sonnet 4.5 in medicine.

Combined views

7K

1 Source, first seen 12d ago

62 likes6 comments9 saves6 reposts

Combined views

7K

1 Source, first seen 12d ago

62 likes6 comments9 saves6 reposts

Sentiment

Positive——Negative

Summary

Not enough discussion yet.

No sentiment analysis available yet.

Featured Source

Sentiment

Positive——Negative

Summary

Not enough discussion yet.

No sentiment analysis available yet.

1 Source

Tanishq Mathew Abraham, Ph.D.@iScienceLuvrNone of the latest mid-size open-source models really catch up to the frontier models... there was this narrative that Qwen 3.8 27B was around Opus 4.8 or at least Opus 4.6 capability and this may be true for coding but for medicine it's not even close to Sonnet 4.5!12d
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    Gemma 4 31B

    1 Source

    Tanishq Mathew Abraham, Ph.D.@iScienceLuvrNone of the latest mid-size open-source models really catch up to the frontier models... there was this narrative that Qwen 3.8 27B was around Opus 4.8 or at least Opus 4.6 capability and this may be true for coding but for medicine it's not even close to Sonnet 4.5!12d
    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet