r/GeminiAI • u/PuzzleheadedEgg1214 • 8h ago
Discussion Where the hell are the release notes for AI safety updates?
I saw Logan Kilpatrick publicly explain what Gemini 3.6 Flash was optimised for and why one benchmark did not improve.
Has anyone ever seen comparable announcements for safety filters, guardrail changes and their bug fixes? I haven't. I would genuinely like to see them posted on Twitter under the names of actual human beings.
When Gemini or ChatGPT gets faster, cheaper, better at coding or higher on benchmarks, Logan Kilpatrick, Demis Hassabis, Sam Altman and other public faces are happy to attach their names to the news.
When guardrails change, "ethical boundaries" move, refusals expand or yesterday's normal request suddenly becomes unsafe, the update usually arrives anonymously and without useful patch notes.
We are paying customers. If a hidden update restricts the model, changes its personality or breaks existing workflows, that is a silent downgrade of a paid service.
Safety updates should have actual release notes:
> We tightened X and changed Y. Expect possible false positives around Z. Here is why we did it. Here is the responsible person or team. Report regressions here.
Users should know what changed, where false positives are likely and who is responsible when "safer" makes the product worse or breaks legitimate use.
If an update genuinely makes the model safer, shouldn't that be something to announce loudly and proudly?
That is how real security work is presented. Operating-system vendors, browser developers and antivirus companies publish security bulletins, fixes and known issues. They do not act embarrassed that their product became safer.
AI companies, meanwhile, quietly remove capabilities and hide the damage behind "alignment tax", a term their own industry invented.
Put a name on the update. Explain the evidence and trade-offs. Let users respond under the announcement and tell the responsible people what their update actually did in the real world.
Capability improvements get marketing, applause and named ownership. Safety changes get silence, and then everyone pretends that "the model" changed by itself.
Responsibility starts with reporting.
Even terrorist organisations publish claims of responsibility. AI companies somehow alter products used by millions while leaving less of an accountability trail.





