How are you managing AI safety, Alignment and Hostile/Rogue agents right now?
I'm building an AI kill switch platform for companies managing hostile and rogue AI. Here in NYC there's a bill that might get passed that has a lot of people worried so we're supporting some users with it. It works. But I still feel like I lack more nuanced feedback from people who actually do this stuff day-to-day and have had to build their own solutions internally. I'd love if anyone could speak on techniques they're comfortable sharing on how they've been able to manage this issue internally. It would really help me and I imagine help many others immensely.
scorecomments7 sightings
first seen 2026-10-05 17:39 UTClast seen 2026-10-06 09:47 UTCscore then 1score now 0gained -1sightings 7
open on reddit ↗
💬 16 (+9)