
#Safety
1311 posts
Feed·
20 of 1311 posts

🖼️
🖼️

🖼️
0
15s

🖼️
449
449
NEW: malware developers added nuclear & biological weapons text to to their spyware. Goal? To trigger LLM safety refusals... so that their spyware wouldn't be analyzed by an AI security scanner. Cleanest practical example I can think of for why over-indexing on first order https://t.co/WLxe0LWo8s
Hacker News·4 months ago
#U4ElWg4W15s

📰

🖼️
0
0
OpenAI Charts Multi-Pronged Policy Roadmap for AI Governance
15s
🖼️
0
15s
🖼️
0
15s

📰

🖼️

🖼️
0
0
Queens Man Faces Federal Charges Over Pokémon Card Armed Robbery Cases
15s
🖼️
0
15s

🖼️
0
0
Does xAI's Grok model have a unique vulnerability toward validating psychoses? - Annielytics.com
15s

🖼️
0
15s

🖼️

🖼️

🖼️

🖼️

🖼️
0
0
FAA Chief Rewrites Trump Air Traffic Control History — Then Blames Airlines For The System He Runs
15s

🖼️
0
15s