Modal @modal · 27 Aug 2026
We're releasing Terminal-Bench-Science: a benchmark for evaluating AI agents on research workflows across scientific domains. An ongoing Stanford-led community effort, built by the team behind Terminal-Bench together with scientific domain experts at research institutions https://t.co/9YXMDrivGi
205 234Views
740Likes
135Reposts
56Replies
69Quotes
0Bookmarks
Is that a lot?
5.55×vs this author's median36 967 views is typical
67Percentile for this authorof 6 recent posts
70.9×vs 10K–100K median2 896 views is typical
5.7× audienceReachviews ÷ followers
0.49%Engagement rateof viewers reacted