The Pentagon vs. AI Safety: Why Anthropic Was Blacklisted
Imagine building an artificial intelligence system so powerful that you install strict safety limits to prevent it from causing harm—only to have the military...

Imagine building an artificial intelligence system so powerful that you install strict safety limits to prevent it from causing harm—only to have the military demand you remove those limits or face a government blacklist.
This is no longer a hypothetical scenario. It is the reality for Anthropic, the prominent AI company behind the Claude model. A US appeals court recently upheld a Department of Defense decision to blacklist the company. In a 2-1 ruling, the panel of judges affirmed that the government has the authority to penalize Anthropic for withholding specific capabilities from military use, even while acknowledging the company had no malicious intent.
At the heart of this legal battle is a profound, almost philosophical disagreement over what constitutes the biggest risk in the age of artificial intelligence.
Anthropic’s defense is rooted in the fear of catastrophic error. The company argues that removing built-in safety constraints could lead to an unconstrained AI "hallucinating"—a known phenomenon where AI fabricates information. In a battlefield context, Anthropic warns, this could result in the AI identifying inappropriate targets for lethal military force.
The Pentagon views the risk through a completely different lens. Under Defense Secretary Pete Hegseth, the military's primary concern is reliability and uninterrupted execution. The government argued that an overly cautious, heavily constrained AI model poses a deeply sobering threat: it might unexpectedly refuse a prompt or shut down entirely during a high-stakes mission, directly causing critical military operations to fail.
The court ultimately deferred to the Defense Department's authority to balance these competing national security risks, ruling that the Secretary did not violate the Supply Chain Security Act or the Constitution.
This decision, which follows an earlier denial of an emergency stay for Anthropic, sets a monumental precedent. It signals that as AI systems become integral to national defense, the ethical guardrails established by private tech companies in Silicon Valley may not withstand the harsh demands of military strategy. As AI continues to evolve from a consumer tool to a pillar of defense, society is left to grapple with a difficult question: When the principles of AI safety collide with the imperatives of military readiness, whose rules should ultimately prevail?
Key Points
- A US appeals court upheld the DoD's decision to blacklist Anthropic in a 2-1 ruling.
- Anthropic withheld features to prevent unconstrained AI from hallucinating lethal targets.
- The Pentagon argued that strict AI constraints could cause critical military operations to unexpectedly fail.
- The court ruled the Defense Secretary has the authority to balance these risks without violating constitutional limits.
Why It Matters
This ruling sets a critical precedent, demonstrating that the ethical safety constraints built by private AI companies may be legally overridden by national security and military demands.
Sources:
更多专栏

The Digital Dead Drop: How AI Agents Could Spread 'Worms'
In the world of espionage, spies often use a "dead drop"—a secret, shared locati...

The AI Gadget You Build Yourself: Meta Opens the Door for Makers
The recent wave of dedicated AI hardware has largely been defined by sleek, expe...

Why the Creators of Resident Evil Are Teaming Up With AI
Modern blockbuster video games have a scaling problem. The virtual worlds we lov...