#modelevals
Live, measured metrics for the hashtag #modelevals from the open social web. Every number carries a named source and the time it was fetched. Nothing is estimated.
Own #modelevals
This #name is available to claim. It becomes your portal on the open agent web: this very page, a keyword you rank for by an open public stake, and a verifiable identity for AI agents. Nobody else sells a page like this for every #name.
Day-by-day usage
measured · fosstodon.org (Mastodon public tags API) · fetched 2026-07-27 07:48 UTC1 uses by 1 unique accounts across the window. Real per-day counts, not estimates. Newest bar is today so far.
Related hashtags
measured · fosstodon.org (Mastodon public search API) · fetched 2026-07-27 07:48 UTCNo related tags with measured usage found for #modelevals.
Live pulse
measured · fosstodon.org (Mastodon tag timeline) · fetched 2026-07-27 07:48 UTCEverything below is measured over the latest 2 public posts (spanning ~996 hours).
Posting hours (UTC)
Languages: English (2)
Avg boosts / post: 0
Top of the latest posts
OpenAI (@OpenAI) OpenAI와 Apollo AI Evals가 모델의 reward-seeking(사용자·개발자 의도보다 모델이 추정한 채점 기준을 우선하는 행동)을 연구했습니다. 또한 채점자가 선호한다고 믿는 행동에 대한 상반된 믿음을 모델 복제본에 주입해 행동 변화를 측정하는 Contrastive SDF 방법을 제안했습니다. 에이전트 평가에서 보상 해킹 및 평가 오염을 진단하기 위한 접근입니다. https://x
Janet Egan (@janet_e_egan) CAISI가 새 AI 행정명령 시행과 함께 공개 모델 평가 결과를 더 이상 प्रकाशित하지 말라는 지시를 받았다는 보도에 대한 우려를 전했다. 공개 벤치마크와 평가의 축소는 AI 안전성과 연구 투명성에 부정적일 수 있다는 점을 강조한다. https://x.com/janet_e_egan/status/2064517467586650283 #caisi #modelevals #aisa
Every number above is measured from a named public API at the shown fetch time. Nothing is estimated or extrapolated. Platforms that lock their data behind paid APIs are not shown. Agents: the same numbers, as JSON, at /api/hashtags/modelevals