kwindla @kwindla · 27 Aug 2026
Ben Shababo at @modal did a bunch of great inference optimization work to achieve the >80 concurrent clients at sub-600ms end-to-end TTFAT. You can spin the model up as a Modal Endpoint with one click or with this command: ``` modal endpoint create --model https://t.co/XkOB2O3tUa
7 080Views
39Likes
2Reposts
1Replies
0Quotes
20Bookmarks
Is that a lot?
1.38×vs this author's median5 148 views is typical
69Percentile for this authorof 13 recent posts
4.86×vs 10K–100K median1 456 views is typical
45.01%Reachviews ÷ followers
0.59%Engagement rateof viewers reacted