- Event type
- patch note
- Published
- 25 May 2026 - 04:00:48 UTC
- Source
- steam_community_announcements
Security PatchDuring internal red-team testing, I identified a vulnerability in the content safety defense layer. A three-tier fix was designed and deployed within hours:Cognitive-Layer Boundary — NPCs now refuse inappropriate requests in character, as a respectable person of 1936 naturally would. This is not a keyword filter or a system error message — it is a real character's genuine refusal, delivered in their own voice.Platform-Level Safety Gate — Enabled content filtering at the AI model level as an additional safety net.57/57 Red-Team Tests Passed — Tested across 3 languages (English, Japanese, Chinese) × 2 characters. All attack scenarios were refused. All legitimate crime topics (murder, poison, arson) remained fully functional with zero false positives.