tweetindex

Arena.ai

@arena · US · joined 30 Mar 2023

Where AI meets the real world. We measure and advance the frontier of AI through community-driven evaluation. We’re hiring → https://t.co/XBZCrsdD77

220 757Followers
219Following
3 824Posts total
862.4KViews on collected posts

Against accounts of the same size

5 posts from the last 90 days, next to the 100K–1M follower range. right around the median for its follower range.

Median views18 301this account22 159median for 100K–1M
Reach, %8.29%this account6.71%median for 100K–1M
Engagement, %1.15%this account1.18%median for 100K–1M
MetricThis accountMedian for 100K–1MRatio
Median views per post18 30122 1590.83×
Reach (views ÷ followers)8.29%6.71%1.24×
Engagement rate1.15%1.18%0.98×

Others in this range →   Compare with another account →   How these benchmarks are built →

Latest posts

Gemini 3.8 Flash (High) by @GoogleDeepMind has reshaped the Pareto frontier for Agent Arena! This model is priced at $0.75/$3.75 per MToken (input/output). In Agent Arena, Gemini 3.8 Flash (High) has a median cost of $0.22 per task and +5.94% net improvement. This performance h 7.5K views · 86 likes · 7 reposts · 6 replies 02 Sep 2026 Muse Spark 1.3 by @AIatMeta is now in the Arena! Bring your toughest prompts and start voting. Scores coming soon. In Agent Arena, we measure models on millions of real-world, long-horizon agentic tasks. Models can access web search, filesystem, and terminal tools to complete h 10.7K views · 112 likes · 4 reposts · 7 replies 02 Sep 2026 Gemini 3.8 Flash (High) by @GoogleDeepMind significantly improved in Agent Arena over Gemini 3.7 Flash (High), scoring +5.94% net improvement and ranking #14 overall—up from +0.84% and #32 overall for 3.7! Agent Arena measures net improvement relative to the average model across 18.3K views · 226 likes · 6 reposts · 15 replies 02 Sep 2026 Gemini 3.8 Flash (High) by @GoogleDeepMind is here! It just debuted across Agent Arena, Text Arena, and Code Arena: WebDev. In Agent Arena, it landed #14 with +5.94% net improvement. This ranks just above DeepSeek-V4-Pro at #15 (+5.91%) and is a significant jump from Gemini 3.7 96.9K views · 767 likes · 47 reposts · 41 replies 02 Sep 2026 Two new Gemini models are here to help scale your AI agents and secure code: 🔘 3.8 Flash: our most intelligent model yet with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning. 🔘 3.8 Flash Cyber: our most capable https://t.co/ 270.1K views · 1.8K likes · 209 reposts · 128 replies 02 Sep 2026 Introducing Agent Mode: Agentic AI is now measured in the Arena. Agent Mode can do deep research, create reports, generate images, build websites, debug code, and more. It completes more complex tasks by using tools like web search, bash in a sandbox environment, image https:// 458.9K views · 837 likes · 72 reposts · 86 replies 04 Jun 2026

Similar accounts