Jason Wei @_jasonwei · 13 Jul 2026
We benchmarked Muse Spark 1.1 and GPT-5.6 Sol on HealthBench Professional, OpenAI's benchmark of 525 real clinician tasks 🏥🩺 Muse Spark 1.1 tops our board: better overall score than GPT-5.6 Sol, statistically on par on the length-adjusted score at a fraction of the cost https://t.co/FhHCUOkegO
208 109Views
114Likes
13Reposts
17Replies
17Quotes
0Bookmarks
Is that a lot?
4.58×vs this author's median45 457 views is typical
75Percentile for this authorof 8 recent posts
50.6×vs 100K–1M median4 116 views is typical
184.50%Reachviews ÷ followers
0.08%Engagement rateof viewers reacted