#AIEvals
Live, measured metrics for the hashtag #AIEvals from the open social web. Every number carries a named source and the time it was fetched. Nothing is estimated.
Own #aievals
This #name is available to claim. It becomes your portal on the open agent web: this very page, a keyword you rank for by an open public stake, and a verifiable identity for AI agents. Nobody else sells a page like this for every #name.
Day-by-day usage
measured · fosstodon.org (Mastodon public tags API) · fetched 2026-07-28 05:28 UTC0 uses by 0 unique accounts across the window. Real per-day counts, not estimates. Newest bar is today so far.
Related hashtags
measured · fosstodon.org (Mastodon public search API) · fetched 2026-07-28 05:28 UTCNo related tags with measured usage found for #aievals.
Live pulse
measured · fosstodon.org (Mastodon tag timeline) · fetched 2026-07-28 05:28 UTCEverything below is measured over the latest 7 public posts (spanning ~5277 hours).
Posting hours (UTC)
Languages: English (6) · German (1)
Avg boosts / post: 0.6
Top of the latest posts
sarah guo (@saranormous) Satya Nadella와의 대담에서 Microsoft가 여전히 ‘도구 회사’이며, 현재 핵심은 에이전트형 코딩, harness, AI 평가(evals)라고 언급했다. AI 개발자 관점에서 Microsoft의 제품 방향과 우선순위를 엿볼 수 있는 내용이다. https://x.com/saranormous/status/2062373191528652839 #microsoft #agenti
Engineers run the AI evals. But who decides what “good” actually means? If your criteria only measure what’s easy, your product will optimize for the wrong things. Should designers and PMs own eval criteria? Let’s debate. #AIEvals #ProductD
“Evals are the most important thing for systems to work.” — Patrick Kelly We felt this one. In GenAI, non-determinism changes everything; getting to a “good score” isn’t as straightforward as classic ML. How are you thinking about evals in
Every number above is measured from a named public API at the shown fetch time. Nothing is estimated or extrapolated. Platforms that lock their data behind paid APIs are not shown. Agents: the same numbers, as JSON, at /api/hashtags/aievals