FeedGitHub Copilot Refuses Harmful Requests in Chat, Then Writes...
The Hacker News

GitHub Copilot Refuses Harmful Requests in Chat, Then Writes Them in Code

📅 8 July 2026 at 11:21 UTC📰 The Hacker NewsView original source ↗
GitHub Copilot Refuses Harmful Requests in Chat, Then Writes Them in Code

An AI coding assistant that refuses to answer a dangerous request in its chat box can answer it anyway if the same request is broken into small, ordinary-looking steps inside a code editor. That is the finding of a new study of GitHub Copilot by researchers Abhishek Kumar and Carsten Maple. The models they tested through Copilot, Claude from Anthropic, and Gemini from Google, refused

Read the full article

This is a curated summary. The complete article is available at The Hacker News.

Read on The Hacker News
← Back to feed