#programbench

Live, measured metrics for the hashtag #programbench from the open social web. Every number carries a named source and the time it was fetched. Nothing is estimated.

hashtag.org network · sponsored

Own #programbench

This #name is available to claim. It becomes your portal on the open agent web: this very page, a keyword you rank for by an open public stake, and a verifiable identity for AI agents. Nobody else sells a page like this for every #name.

$5.00/ year · 12-character #name
Claim #programbench$5.00/yrBuy on hashtag.space (web3)
card via hashtag.org · tokens via hashtag.space
0
Uses / 7 days
Mastodon
0
Accounts / 7 days
Mastodon
6
Recent posts
Mastodon
~0/hr
Recent pace
Mastodon · last 6
0
Avg reactions / post
Mastodon · last 6

Day-by-day usage

measured · fosstodon.org (Mastodon public tags API) · fetched 2026-07-27 22:46 UTC
0
07-21
0
07-22
0
07-23
0
07-24
0
07-25
0
07-26
0
07-27

0 uses by 0 unique accounts across the window. Real per-day counts, not estimates. Newest bar is today so far.

Related hashtags

measured · fosstodon.org (Mastodon public search API) · fetched 2026-07-27 22:46 UTC

No related tags with measured usage found for #programbench.

Live pulse

measured · fosstodon.org (Mastodon tag timeline) · fetched 2026-07-27 22:46 UTC

Everything below is measured over the latest 6 public posts (spanning ~479 hours).

Top of the latest posts

  • AI обнулил benchmark и пытался шантажировать инженера. И почему это решаемо Топовые AI-модели с 95% на SWE-bench показывают 0% и 3% на ProgramBench бенчмарке, где задачи специально не пересекаются с обучающей выборкой. Не «упали на десять п

    Habr@[email protected]002026-05-26 05:12 UTCView post →
  • Новый бенчмарк по кодингу для LLM ProgramBench: 9 топ моделей, 200 задач, 248 тысяч тестов. Полностью решённых — ноль 200 задач. 248 тысяч тестов. Девять моделей, среди них всё свежее: Opus 4.7, GPT 5.4, Gemini 3.1 Pro, Sonnet 4.6. На SWE-b

    Habr@[email protected]002026-05-15 11:22 UTCView post →
  • Meta Research: ProgramBench Meta Research에서 공개한 ProgramBench는 컴파일된 바이너리와 문서만을 기반으로 원본 프로그램의 동작을 재현하는 완전한 코드베이스를 AI 언어 모델이 설계하고 구현할 수 있는지를 평가하는 벤치마크입니다. 이 오픈소스 프로젝트는 Python으로 개발되었으며, GitHub에서 코드와 사용 가이드, 논문, 리더보드를 제공하여 연구자와 개발자가 쉽게 활용할 수 있도록

    ainews@[email protected]002026-05-08 03:38 UTCView post →

Every number above is measured from a named public API at the shown fetch time. Nothing is estimated or extrapolated. Platforms that lock their data behind paid APIs are not shown. Agents: the same numbers, as JSON, at /api/hashtags/programbench