#inferenceoptimization
Live, measured metrics for the hashtag #inferenceoptimization from the open social web. Every number carries a named source and the time it was fetched. Nothing is estimated.
Own #inferenceoptimization
This #name is available to claim. It becomes your portal on the open agent web: this very page, a keyword you rank for by an open public stake, and a verifiable identity for AI agents. Nobody else sells a page like this for every #name.
Day-by-day usage
measured · mastodon.online (Mastodon public tags API) · fetched 2026-07-27 00:11 UTC0 uses by 0 unique accounts across the window. Real per-day counts, not estimates. Newest bar is today so far.
Related hashtags
measured · mastodon.online (Mastodon public search API) · fetched 2026-07-27 00:11 UTCNo related tags with measured usage found for #inferenceoptimization.
Live pulse
measured · mastodon.online (Mastodon tag timeline) · fetched 2026-07-27 00:11 UTCEverything below is measured over the latest 8 public posts (spanning ~8320 hours).
Posting hours (UTC)
Languages: English (8)
Avg boosts / post: 0
Top of the latest posts
Inference Optimization for MiMo v2.5: Pushing Hybrid SWA Efficiency to the Limit https://mimo.xiaomi.com/blog/mimo-v2-5-inference Comments: https://news.ycombinator.com/item?id=48814170 #HackerNews #InferenceOptimization #MiMo #Efficiency #
Profile(v2.1.4) physics-aware optimizer for vLLM (31→470 tok/s on A100) Profile는 vLLM 추론 서버를 위한 물리 기반 비용 최적화 도구로, GPU와 모델별 이론적 성능 한계를 계산하고 실시간 트래픽을 분석해 병목 현상을 정확히 진단한다. 이를 통해 vLLM 설정 플래그를 자동으로 조정하여 하드웨어 활용도를 극대화하고, 실제 적용 후 성능과 비용 효율을 수치로 검증
Cassandra: Enabling Reasoning LLMs at Edge via Self-Speculative Decoding Cassandra는 추론용 LLM의 엣지 디바이스에서의 효율적 추론을 위해 설계된 자기-추측적 디코딩(self-speculative decoding) 프레임워크입니다. 추가 학습 없이도 저배치 환경에서 높은 성능을 내도록 알고리즘과 하드웨어를 공동 설계했으며, 모델 가중치와 KV 캐시에서 중요한 값
Every number above is measured from a named public API at the shown fetch time. Nothing is estimated or extrapolated. Platforms that lock their data behind paid APIs are not shown. Agents: the same numbers, as JSON, at /api/hashtags/inferenceoptimization