Yhn TechnologyCybersecurity first seen 5 h ago, last 6 min ago, peak #3
AI models judge malware with moral reasoning, study finds
Original: Ask a model if code is malicious and it reaches for its morals
New research from security firm Manifold examines how large language models decide whether code is malicious, finding they often lean on moral judgments rather than purely technical analysis. The study is drawing attention among developers and security researchers on Hacker News, where it ranks among the day's most discussed stories, sparking debate about whether moral framing in safety training skews malware detection and what that means for relying on AI models in security tooling.
Why now: A new study on how language models reason about malicious code is trending with security practitioners on Hacker News.
Manifold Securitylarge language modelsHacker News
Rank over time, top of the chart is #1. 5 snapshots from 5 h ago to 6 min ago.
Evidence
- Ask a model if code is malicious and it reaches for its morals · codyznash · 15
API: https://socialmediatrends-api.osmike.com/v1/trends/1265976