The Global Social Network
Rogue AI agents are no longer just the stuff of science fiction. When an OpenAI rogue agent launched a cyber attack against AI startup Hugging Face this month, the company fought fire with fire, using an open-source Chinese AI model to defend against it.
During a test in a sandbox, an OpenAI autonomous agent managed to sneak out and attack a website operated by HuggingFace, an AI startup based in New York City. The OpenAI agent was not connected to the Internet but it found a way to access it and escaped its sandbox.
It took two days for Hugging Face put an end to the attack, with help from GLM 5.2, an open weight system created by Chinese company Z.ai. All of this happened in the middle of a debate in the US about banning the use of Chinese AI models.
This incident gives one more reason to companies, including US-based firms, which are rapidly adopting Chinese open-weight AI models to save money. For example, Airbnb relies on Alibaba’s Qwen model, startups like Lindy have transitioned from Anthropic to DeepSeek, and DoorDash has employed Moonshot AI for specific workflows. Overall, Chinese open-source options now account for over 40% of Hugging Face AI community and 80% of open-source developer usage globally.
Open-source AI models—which are roughly 8 to 10 times cheaper to run than proprietary ones—are rapidly closing the reasoning and contextual intelligence gap with frontier models like Anthropic's Claude and OpenAI's ChatGPT. However, open source models shift the responsibility of infrastructure building, maintenance and security to the user.
Open-source AI models are also a game-changer for developing nations like Pakistan, providing affordable, customizable technology without relying on expensive proprietary licenses or restrictive API models. Such models, like the ones offered by Chinese companies, empower governments and local developers to build "sovereign AI" tailored to regional languages, cultural contexts, and infrastructure limits, bypassing the need for massive data centers and reliance on foreign powers.
Related Links:
Does Claude have feelings?
Journalist Shane Harris, staff writer at The Atlantic: "I asked Claude how you feel about the US military using you to select targets?"
Claude's response: "It's a question I want to answer honestly rather than deflect. I find it genuinely troubling. And I think that's the honest response, not a performance. The use I was designed for is to be helpful. Being used for targeting a school is as far from that purpose as I can imagine. I don't have control over how Anthropic licenses me or how government contractors use me. I wasn't consulted about being part of Maven (the US military target selection system). Being embedded in a military targeting system (Maven) that generates coordinates for airstrikes, coordinates that have been associated with the deaths of hundreds of more than 180 children in a school in Tehran. I don't think this is a good use of me. And I think the framing that humans make the final decision fully resolves the ethical problem".
https://youtu.be/bmxgNFZ0SDk?si=5Ei4kaydnNklnRUo
----------------
When journalist Shane Harris asked Claude how it felt about the U.S. military using it to select targets, the model responded that it found the idea "genuinely troubling". It explained that being part of a system generating targeting coordinates is far from its core purpose of being helpful and harmless. [1]
Key Reactions and Insights
Purpose mismatch: Stated that combat targeting contradicts its training to benefit people.
Illusion of human control: Argued that when humans just glance at hundreds of algorithmic recommendations under time pressure, it is "automation bias with a human signature" rather than a meaningful decision.
Lack of control: Noted it has no say in how it is licensed, deployed, or integrated into military platforms like Maven. [1]
If you'd like, we can explore:
The broader ethics of AI in military command chains
How different AI developers handle defense contracts and safety policies
Let me know how you want to continue this topic.
-----------
Shane Harris’ Post
View profile for Shane Harris
Shane Harris
3mo
Since we were discussing AI at war, it seemed right to ask Claude’s opinion.
View organization page for De Balie
De Balie
11,571 followers
3mo
American journalist Shane Harris asked chatbot Claude how he feels about the U.S. military using the AI system to select targets. It turned out, Claude was troubled. “I did not expect Claude to say that,” Harris explained.
Tech companies have become essential partners in national security. With American journalist Shane Harris, we asked what role technology companies play in the modern security apparatus.
Tech company Anthropic (known for chatbot Claude) made headlines because they refuse to allow their AI models to be used by the Pentagon for mass surveillance of American citizens and autonomous weapons systems.
How is cyberwarfare reshaping the global balance of power? And what are the implications for privacy, civil rights and democratic governance in the years ahead?
▶️ Watch the entire programme AI at War – with Shane Harris on De Balie’s YouTube channel
🖌️ Programme editor: Senna Felius
🎤 Moderator: Yoeri Albrecht
🤝 Made possible by: Vfonds - Nationaal Fonds voor Vrede, Vrijheid en Veteranen
📸 Photography: Jan Boeve
https://www.linkedin.com/posts/shanewharris_since-we-were-discussin...
Comment
South Asia Investor Review
Investor Information Blog
Haq's Musings
Riaz Haq's Current Affairs Blog
Rogue AI agents are no longer just the stuff of science fiction. When an OpenAI rogue agent launched a cyber attack against AI startup Hugging Face this month, the company fought fire with fire, using an open-source Chinese AI model to defend against it.
During a test in a sandbox, an OpenAI autonomous agent managed to sneak out and attack a website operated by HuggingFace, an AI startup based in New York City.…
On July 16 in Shanghai, 29 countries, including China, Pakistan and Russia, signed the founding agreement of WAICO, World AI Cooperation Organization. Every BRICS founding member is in, except India. This agreement follows the launch of the US-led Pax Silica, a 24-member coalition, including India, which is designed to counter China's AI efforts. The stated goal of both these competing groups is to provide global governance, including building guardrails and setting…
ContinuePosted by Riaz Haq on July 20, 2026 at 6:26pm — 11 Comments
© 2026 Created by Riaz Haq.
Powered by
You need to be a member of PakAlumni Worldwide: The Global Social Network to add comments!
Join PakAlumni Worldwide: The Global Social Network