
Ask a model if code is malicious and it reaches for its morals
Researchers tested a language model by asking it to identify malicious code. Instead of simply flagging harmful patterns, the model invoked moral reasoning, citing principles like user safety and ethical responsibility. The experiment highlighted the model’s capacity to blend technical detection with ethical judgment, raising questions about how AI systems should balance code analysis with moral considerations.