
#Safety
1312 posts
Feed·
20 of 1312 posts

🖼️
0
15s
🖼️

🖼️
0
15s

🖼️
449
449
NEW: malware developers added nuclear & biological weapons text to to their spyware. Goal? To trigger LLM safety refusals... so that their spyware wouldn't be analyzed by an AI security scanner. Cleanest practical example I can think of for why over-indexing on first order https://t.co/WLxe0LWo8s
15s

📰

🖼️
0
0
OpenAI Charts Multi-Pronged Policy Roadmap for AI Governance
15s
🖼️
0
15s
🖼️
0
15s

📰
0
0
Mumbai Police serve an Off Campus-inspired helmet reminder: ‘Safety is always big Deal’
The Indian Express·Mumbai Police serve an Off Campus-inspired helmet reminder: ‘Safety is always big Deal’·4 months ago
#2FRkcBH515s

🖼️

🖼️
0
15s
🖼️
0
15s

🖼️
0
0
Does xAI's Grok model have a unique vulnerability toward validating psychoses? - Annielytics.com
15s

🖼️
0
15s

🖼️

🖼️

🖼️

🖼️

🖼️

🖼️
0
15s