Post

Log inSign up

Post

  • user avatar
    Medical Sphere
    @MedicalSphereAI
    We benchmarked Muse Spark 1.1 and GPT-5.6 Sol on HealthBench Professional, OpenAI's benchmark of 525 real clinician tasks 馃彞馃┖ Muse Spark 1.1 tops our board: better overall score than GPT-5.6 Sol, statistically on par on the length-adjusted score at a fraction of the cost ($1.25/$4.25 vs $5/$30 per M tokens in/out, ~7脳 cheaper on output).
    9:11 PM 路 Jul 13, 2026184.1KViews
  • user avatar
    Mark Santos
    @markksantos
    Jul 14
    wheres fable 5?
    166
  • user avatar
    Luke
    @groccy1
    Jul 14
    Awesome. What about Grok 4.5?
    414
  • user avatar
    Emily
    @IamEmily2050
    Jul 14
    Promising results 馃檹
    487

Log in or sign up for X

See what鈥檚 happening and join the conversation

Continue with phone
or
Log in with username or email

Relevant people

Avatar
Medical Sphere@MedicalSphereAIFollow
The global community for advancing AI in healthcare Tag @AskMedSphere to test AI models on medical cases

Trending now

Terms路Privacy路Cookies路Accessibility路Ads Info路漏 2026 X Corp.