Peter Gostev @petergostev · 24 Feb 2026
I've got a fun new benchmark for you where most LLMs are doing pretty badly - "Bullshit Benchmark". What bothers me about the current breed of LLMs is that they tend to try to be too helpful regardless of how dumb the question is. So I've built 55 'bullshit' questions that don't https://t.co/4o4quN5EFR
906 471Views
4 723Likes
413Reposts
258Replies
182Quotes
2 164Bookmarks
Is that a lot?
109×vs this author's median8 295 views is typical
100Percentile for this authorof 5 recent posts
313×vs 10K–100K median2 896 views is typical
37.4× audienceReachviews ÷ followers
0.61%Engagement rateof viewers reacted