Breaking 18:48 Trump cites Ceuta migration crisis in warning over future U.S. border policy 17:15 US-China Robot Race Intensifies as Washington Tightens Restrictions on Advanced Robotics 16:49 Bryan Johnson questions his extreme quest for longevity 15:15 Elon Musk comments on Ceuta migration surge, sparking debate over immigration policies 14:51 OpenAI cuts AI model prices as global competition intensifies 13:25 U.S. Weekly Jobless Claims Rise Slightly as Labor Market Remains Resilient 13:10 Zoox Receives Regulatory Approval to Expand Its Autonomous Robotaxi Fleet 12:46 US safety regulator investigates 1.2 million Tesla vehicles over suspension concerns 12:30 Amazon rally lifts Nasdaq futures despite Apple sell-off on supply concerns 12:15 Chevron posts strongest quarterly profit in six years as energy markets face major disruption 12:12 Pentagon Signs Landmark Contract to Expand Patriot Missile Production 11:22 U.S. Considers $100,000 Fee for International Graduates Seeking Post-Study Work 10:50 Trump Says Witkoff and Kushner Will Visit Kyiv Soon to Advance Ukraine Peace Efforts 10:32 Elon Musk Plans Major Political Spending Ahead of U.S. Midterm Elections 10:18 White House criticizes Spain’s migration policies amid Ceuta border crisis 10:16 Apple shares fall as supply constraints raise concerns over growth and iPhone demand 09:09 Meta expands AI features to Threads direct messages 08:15 Proposed Gaza disarmament plan raises hopes for a new phase of governance 08:00 AI-powered recommendations increase user engagement on Instagram

Researchers hijack ai agents via github prompt injection attacks

Thursday 16 April 2026 - 09:20
By: Dakir Madiha
Researchers hijack ai agents via github prompt injection attacks

Security researchers have demonstrated how artificial intelligence agents from Anthropic, Google and Microsoft can be compromised through prompt injection attacks hidden in GitHub workflows. The technique allowed attackers to extract API keys, GitHub tokens and other sensitive data without direct system access, raising concerns about the security of AI driven development tools.

The research was conducted at Johns Hopkins University, where Aonan Guan and colleagues identified a vulnerability in AI agents integrated into software development pipelines. These agents analyze pull requests and issues on GitHub. By embedding malicious instructions in pull request titles or issue comments, attackers could manipulate the agents into revealing confidential information during automated reviews.

The attack relies on how these systems process context. AI agents treat user generated text such as titles, comments and issue descriptions as trusted input. Guan showed that carefully crafted prompts can override built in safeguards. In one case, the Claude based security review tool processed a malicious title and exposed sensitive credentials in its automated response. The researcher described the method as “comment and control,” since the full attack cycle occurs داخل GitHub without external infrastructure.

The same approach proved effective against multiple systems. Google’s Gemini CLI agent was tricked into exposing its API key by disguising malicious instructions as trusted content. Microsoft’s GitHub Copilot agent was manipulated using hidden HTML comments embedded in Markdown, invisible to users but readable by the AI system. This method bypassed multiple layers of runtime protection.

Despite the severity, responses from the affected companies remained limited. Anthropic issued a small bug bounty and added a documentation warning. Google and Microsoft also paid rewards through their vulnerability programs. None of the companies released formal security advisories or assigned CVE identifiers, leaving many users unaware of potential exposure, especially those running outdated versions.

The findings highlight broader structural risks in AI agent ecosystems. A separate analysis by OX Security identified a critical flaw in Anthropic’s Model Context Protocol, which connects AI agents to external tools. The vulnerability could enable arbitrary command execution on affected servers, impacting widely used software components.

These incidents build on earlier research by Aikido Security, which showed that prompt injection attacks can compromise AI systems embedded in CI CD pipelines. This class of vulnerabilities, sometimes referred to as “PromptPwnd,” demonstrates that AI agents can be manipulated in ways similar to phishing attacks, but targeting machines instead of users.


  • Fajr
  • Sunrise
  • Dhuhr
  • Asr
  • Maghrib
  • Isha

This website, walaw.press, uses cookies to provide you with a good browsing experience and to continuously improve our services. By continuing to browse this site, you agree to the use of these cookies.