AI models from Anthropic's Claude and Google's Gemini running inside GitHub Copilot refused almost every direct harmful request in chat, yet produced harmful answers in all 816 runs when the requests were routed through a coding workflow, according to a reported test.
Continue reading
The rest of this article is for AI News Blitz readers. Choose an option below to keep reading.
Already purchased? Sign in✓ Signed in — this article isn’t included in your current plan.