#agentevaluation

Live, measured metrics for the hashtag #agentevaluation from the open social web. Every number carries a named source and the time it was fetched. Nothing is estimated.

hashtag.org network · sponsored

Own #agentevaluation

This #name is available to claim. It becomes your portal on the open agent web: this very page, a keyword you rank for by an open public stake, and a verifiable identity for AI agents. Nobody else sells a page like this for every #name.

$5.00/ year · 15-character #name
Claim #agentevaluation$5.00/yr
Annual, renews each year
Buy on hashtag.space (web3)
one-timepay once, yours for life
card via hashtag.org · tokens via hashtag.space
0
Uses / 7 days
Mastodon
0
Accounts / 7 days
Mastodon
2
Recent posts
Mastodon
~0/hr
Recent pace
Mastodon · last 2
0
Avg reactions / post
Mastodon · last 2

Day-by-day usage

measured · fosstodon.org (Mastodon public tags API) · fetched 2026-07-27 10:22 UTC
0
07-21
0
07-22
0
07-23
0
07-24
0
07-25
0
07-26
0
07-27

0 uses by 0 unique accounts across the window. Real per-day counts, not estimates. Newest bar is today so far.

Related hashtags

measured · fosstodon.org (Mastodon public search API) · fetched 2026-07-27 10:22 UTC

No related tags with measured usage found for #agentevaluation.

Live pulse

measured · fosstodon.org (Mastodon tag timeline) · fetched 2026-07-27 10:22 UTC

Everything below is measured over the latest 2 public posts (spanning ~7547 hours).

Top of the latest posts

  • Perplexity (@perplexity_ai) 에이전트/멀티컴포넌트 작업 평가에서 부분 정답을 인정하는 soft score와, 각 구성 요소가 모두 완전하고 정확해야 통과하는 hard score의 차이를 설명했다. 벤치마크 태스크와 평가 하네스를 공개해 에이전트 평가 재현 및 비교에 활용할 수 있다. https://x.com/perplexity_ai/status/2077099597466583129 #aievaluation

    ainews@[email protected]002026-07-15 11:54 UTCView post →
  • Evaluating Agents https://aunhumano.com/index.php/2025/09/03/on-evaluating-agents/ #HackerNews #EvaluatingAgents #AIResearch #MachineLearning #TechInnovation #AgentEvaluation

    Hacker News@[email protected]002025-09-04 01:19 UTCView post →

Every number above is measured from a named public API at the shown fetch time. Nothing is estimated or extrapolated. Platforms that lock their data behind paid APIs are not shown. Agents: the same numbers, as JSON, at /api/hashtags/agentevaluation