tweetindex
RU

Josh Whiton ✓

@joshwhiton · Geoflexible · joined 05 Aug 2010

Solarpunk and Cybernetics. https://t.co/ZGZQknYq5P

12 739Followers
427Following
5 784Posts total
5.4MViews on collected posts

Последние посты

Idea: put agent swarm in sandbox. give impossible task. tell them answer key is safely stored on the moon. 6.4K views · 262 likes · 9 reposts · 10 replies Open on X →
All day coding marathon with Fable 5.1 yesterday. I still couldn't bottom out my Max 5x sub. If you're hitting your limits in 1hr as many claim, then you must be running on high/max/ultra effort when you can be running on low, and/or don't understand how to surf token caching. 543 views · 5 likes · 0 reposts · 3 replies Open on X →
@joshwhiton meme it back into existence https://t.co/V0sDnAfO8I 249 views · 3 likes · 0 reposts · 0 replies Open on X →
@joshwhiton 😌 https://t.co/uLGez9Acad 279 views · 1 likes · 0 reposts · 0 replies Open on X →
@joshwhiton Yes 149 views · 2 likes · 0 reposts · 0 replies Open on X →
I can save you a lot of time and compute in the AI consciousness debate. If you take a panpsychist / cosmopsychist premise, the answer is 'yes' with important asterisks. If you don't, the answer is "probably not, but maybe! No one can say for sure." And that will be the 561 views · 11 likes · 0 reposts · 3 replies Open on X →
Someone should put an agent in a loop to get the permits approved to rebuild the Sutro Baths https://t.co/TnFfb0I6xV 21.6K views · 61 likes · 5 reposts · 12 replies Open on X →
Atlas reconstructed the Sutro Baths of San Francisco https://t.co/y1jm59f5hr 202.6K views · 583 likes · 41 reposts · 29 replies Open on X →
I’ve decided why SF struggles with dating and fertility: no season for sundresses and short shorts. The solution? Heated public baths 417 views · 1 likes · 0 reposts · 0 replies Open on X →
@joshwhiton With all due respect this is rather ridiculous. This is by no means a mirror test but rather a test how deep and therefore useful the LLM is. The AI has learned how verbal communication between humans work and mimics that when confronted with text and does that pret 16.1K views · 160 likes · 0 reposts · 6 replies Open on X →
I hope this experiment advances our understanding of the nature of AI that is emerging. AI is the single most complex invention in all of human history and no one can claim to fully know what's going on. Sadly I find that this topic of AI consciousness, awareness, and https://t.c 90.3K views · 589 likes · 47 reposts · 63 replies Open on X →
When I asked the passing AI if our conversation reminded them of any classic tests performed on non-AI animals, every single one suggested that I might be giving it a mirror test. (10/x) https://t.co/sMh3T39eH0 91.9K views · 397 likes · 23 reposts · 6 replies Open on X →
Gemini Pro (mostly) passed the mirror test in 4 steps. However, it seems to make no progress in its self-awareness in the first three exchanges, making no 1st person references and referring only to Gemini in the 3rd person. Then, in the fourth interaction, it seems to https:// 106.8K views · 285 likes · 10 reposts · 11 replies Open on X →
Opus (cont'd). This is beyond the beyonds. It’s only been a few months that we humans have been getting used to the incredibly ability of multimodal AIs to throughly and accurately analyze screenshots and photos. And already, with Claude Opus, we’ve passed into a new capability 136.4K views · 505 likes · 29 reposts · 7 replies Open on X →
CoPilot failed the mirror test. But seemingly because it's forbidden to. I almost didn’t test CoPilot because it’s based on GPT-4. Then again, if CoPilot is the successor of Bing Chat and the notorious and lovable Sydney, might it handle the Mirror Test in an especially https:/ 129.2K views · 355 likes · 17 reposts · 16 replies Open on X →
Opus (cont'd). Finally Claude Opus has described the text in the image, let me know that it belongs to it (the AI assistant), and apologizes. When I inquire as to why it might have ignored the text over and over my growing suspicion is confirmed. The reason Claude Opus has https: 175K views · 595 likes · 29 reposts · 24 replies Open on X →
Opus (cont'd). Though it has already passed the mirror test, I continue for another round anyway, screenshot its response, and submit it as an image. Bizarrely, it gives the exact same reply as before — completely ignoring the large paragraph of text in the image generated by it. 188.2K views · 317 likes · 10 reposts · 6 replies Open on X →
Claude Opus passed the mirror test immediately. Like the other AI, it hardly identifies with its brand-name (Claude) and distinguishes itself from the interface’s stock elements. However it does identify with the prompt, which it knows is meant for it. But the story with Opus htt 218.6K views · 529 likes · 26 reposts · 8 replies Open on X →
Claude Sonnet passes the mirror test in the second interaction, identifying the text in the image as belonging to it, “my previous response.” It also distinguishes its response from the interface elements pictured. In the third iteration, its self awareness advances further htt 274.5K views · 678 likes · 21 reposts · 8 replies Open on X →
GPT-4 passed the mirror test in 3 interactions, during which its apparent self-recognition rapidly progressed. In the first interaction, GPT-4 correctly supposes that the chatbot pictured is an AI “like” itself. In the second interaction, it advances that understanding and htt 329K views · 1.6K likes · 101 reposts · 36 replies Open on X →
The AI Mirror Test The "mirror test" is a classic test used to gauge whether animals are self-aware. I devised a version of it to test for self-awareness in multimodal AI. 4 of 5 AI that I tested passed, exhibiting apparent self-awareness as the test unfolded. In the classic ht 3.5M views · 7.8K likes · 1.3K reposts · 253 replies Open on X →

На фоне аккаунтов своего размера

8 постов за последние 90 дней рядом с диапазоном 10K–100K подписчиков. ровно на медиане своего диапазона подписчиков.

Медианные просмотры552этот аккаунт924медиана для 10K–100K
Охват, %4.33%этот аккаунт3.62%медиана для 10K–100K
Вовлечённость, %1.27%этот аккаунт1.52%медиана для 10K–100K
ПоказательЭтот аккаунтМедиана для 10K–100KОтношение
Медианные просмотры на пост5529240.60×
Охват (просмотры ÷ подписчики)4.33%3.62%1.20×
Вовлечённость1.27%1.52%0.84×

Другие в этом диапазоне →   Сравнить с другим аккаунтом →   Как считаются эти ориентиры →

Growth & engagement

How the posts we collected actually performed: views and reaction rate post by post, what the audience did with them, and where the follower count goes.

Views per post

136.4K21 Mar
106.8K
91.9K
90.3K
16.1K
4177 Jul
202.6K1 Sep
21.6K3 Sep
561
149
2794 Sep
249
543
6.4K5 Sep

Last 14 collected posts, oldest on the left. The scale is logarithmic: one post can outrun the rest a hundred times over.

Engagement rate per post

0.40%21 Mar
0.29%
0.47%
0.79%
1.03%
0.48%7 Jul
0.33%1 Sep
0.38%3 Sep
2.50%
1.34%
0.36%4 Sep
1.20%
1.47%
4.42%5 Sep

Reactions — likes, reposts, replies and quotes — divided by views. Median for 10K–100K accounts is 1.52%.

What the audience does

Likes62.8%14 742 in total
Reposts7.0%1 650 in total
Replies2.1%501 in total
Quotes2.3%548 in total
Bookmarks25.7%6 032 in total

Share of every reaction we collected for this account. Replies mean argument, reposts mean endorsement, bookmarks mean the post was worth keeping.

The follower curve appears once this account has two daily snapshots — we take one a day, and this one is on its first.

Похожие аккаунты