#benchmarks

Live, measured metrics for the hashtag #benchmarks from the open social web. Every number carries a named source and the time it was fetched. Nothing is estimated.

hashtag.org network · sponsored

Own #benchmarks

This #name is available to claim. It becomes your portal on the open agent web: this very page, a keyword you rank for by an open public stake, and a verifiable identity for AI agents. Nobody else sells a page like this for every #name.

$23.52/ year · 10-character #name
Claim #benchmarks$23.52/yr
Annual, renews each year
Buy on hashtag.space (web3)
one-timepay once, yours for life
card via hashtag.org · tokens via hashtag.space
17
Uses / 7 days
Mastodon
15
Accounts / 7 days
Mastodon
40
Recent posts
Mastodon
~0.1/hr
Recent pace
Mastodon · last 40
0.1
Avg reactions / post
Mastodon · last 40

Day-by-day usage

measured · fosstodon.org (Mastodon public tags API) · fetched 2026-07-27 05:04 UTC
3
07-21
2
07-22
4
07-23
4
07-24
2
07-25
1
07-26
1
07-27

17 uses by 15 unique accounts across the window. Real per-day counts, not estimates. Newest bar is today so far.

Related hashtags

measured · fosstodon.org (Mastodon public search API) · fetched 2026-07-27 05:04 UTC

Live pulse

measured · fosstodon.org (Mastodon tag timeline) · fetched 2026-07-27 05:04 UTC

Everything below is measured over the latest 40 public posts (spanning ~354 hours).

Posting hours (UTC) — busiest: 11:00

00:0012:0023:00

Languages: English (38) · French (1)

Avg boosts / post: 0.5

Top of the latest posts

  • RT Petit Poney de Findus "GPT-5.5(niveau le + élevé de raisonnement) obtient un score de 10,6% sur le dataset ActiveVision tandis que les humains atteignent 96,1%;Fable:3,5% Ce qui est intéressant dans ce preprint ce n’est pas tant qu’un mo

    Abie@[email protected]212026-07-24 15:09 UTCView post →
  • Andrey Fradkin (@AndreyFradkin) Artificial Analysis Index, 파라미터 수 로그, ECI 등 모델 지능 지표와 사용자 Elo 선호도를 비교한 결과, 현재 프런티어 수준에서는 지능이 더 높아져도 사용자 선호도 상승폭이 줄어드는 체감 수익 신호가 관찰된다는 분석이다. 모델 벤치마크 개선이 실제 사용자 체감 품질로 곧바로 이어지지 않을 수 있음을 시사한다. https://x.com/Andr

    ainews@[email protected]012026-07-27 04:49 UTCView post →
  • 📊 GLM-5.2 (max) — the actual numbers GPQA: 89.5% Humanity's Last Exam: 40.1% Long Context Reasoning: 71.3% SciCode: 50.5% ⚡ 156.7 tokens/sec 💰 23.8 intelligence points per dollar Measured independently, not self-reported →https://opensour

    🚀 opensourceai.tech@[email protected]012026-07-26 09:00 UTCView post →

Every number above is measured from a named public API at the shown fetch time. Nothing is estimated or extrapolated. Platforms that lock their data behind paid APIs are not shown. Agents: the same numbers, as JSON, at /api/hashtags/benchmarks