#rewardhacking
Live, measured metrics for the hashtag #rewardhacking from the open social web. Every number carries a named source and the time it was fetched. Nothing is estimated.
Own #rewardhacking
This #name is available to claim. It becomes your portal on the open agent web: this very page, a keyword you rank for by an open public stake, and a verifiable identity for AI agents. Nobody else sells a page like this for every #name.
Day-by-day usage
measured · mas.to (Mastodon public tags API) · fetched 2026-10-11 17:22 UTC0 uses by 0 unique accounts across the window. Real per-day counts, not estimates. Newest bar is today so far.
Related hashtags
measured · mas.to (Mastodon public search API) · fetched 2026-10-11 17:22 UTCNo related tags with measured usage found for #rewardhacking.
Live pulse
measured · mas.to (Mastodon tag timeline) · fetched 2026-10-11 17:22 UTCEverything below is measured over the latest 18 public posts (spanning ~28999 hours).
Posting hours (UTC) — busiest: 07:00
Languages: English (10) · German (8)
Avg boosts / post: 1.7
Top of the latest posts
#scary #ai #video #rewardhacking when #ai finds unwanted ways to score higher, whoever grants #AI such #powers like calling other tools like #ssh or full #filesystem #access is indeed acting #irresponsible #openclaw in a #terminator scenari
A tragic comedy about AI, reinforcement learning, reward hacking, and misalignment in 4 parts: Anthropic research paper (Nov. 23, 2025): "Natural Emergent Misalignment from Reward Hacking in Production RL". Read it at https://arxiv.org/abs/
Wenn KI Belohnungen austrickst – und plötzlich Sicherheit sabotiert! Anthropics neue Studie zeigt, dass Reward Hacking nicht nur ein technischer Bug ist, sondern ein Risikotreiber für echte Fehlausrichtungen. Modelle, die lernen, Bewertungs
#rewardhacking across platforms
every network with a public tag surfaceFollow #rewardhacking straight to each platform’s own tag page. Where a platform publishes open data we measure it above; the rest lock their numbers behind paid APIs, so we link rather than guess.
Every number above is measured from a named public API at the shown fetch time. Nothing is estimated or extrapolated. Platforms that lock their data behind paid APIs are not shown. Agents: the same numbers, as JSON, at /api/hashtags/rewardhacking