SV Angel @svangel · 02 Sep 2026
We are releasing FrontierSWE v2, our updated ultra-long horizon coding benchmark V2 features an expanded task suite and improved methodology. We see large performance gaps between frontier models, with Claude Fable 5.1 leading by a wide margin https://t.co/VBXoMgMI0e
142 663Views
898Likes
102Reposts
71Replies
68Quotes
0Bookmarks
Is that a lot?
63.0×vs this author's median2 263 views is typical
78Percentile for this authorof 9 recent posts
51.9×vs 10K–100K median2 751 views is typical
3.8× audienceReachviews ÷ followers
0.80%Engagement rateof viewers reacted