2 weeks ago

Anthropic Hires Weapons Expert to Set Claude Safety Boundaries

Anthropic Hires Weapons Expert to Set Claude Safety Boundaries
Anthropic is hiring a weapons expert to decide what Claude should be allowed to do · wionews.com

Anthropic makes Claude, an artificial intelligence system.

The company is hiring a weapons expert to help set rules for what Claude can do.

The expert will decide which military-related requests are safe and which could help build weapons.

Some technologies, such as robots and navigation systems, can be used both for civilian purposes and weapons.

The expert will test how Claude might be misused.

They will also work with engineers to create protections that block harmful uses.

The job includes concerns about systems that can choose and attack targets without human approval.

Anthropic says it is doing this because its AI systems are becoming more capable.

Key facts

Job title
Policy Design Manager, Conventional Weapons
Organisation
Anthropic’s Safeguards organisation
Annual salary
$245,000 to $285,000
Required expertise
Practical knowledge of weapons systems, including experience at a defence research laboratory, government research agency or weapons company
Preferred knowledge
Machine learning, robotics, autonomy, guidance and navigation, sensors, aerospace or embedded software
Main responsibilities
Developing threat models and evaluations and helping implement guardrails, detection systems and enforcement tools
Specific high-risk area
Autonomous systems capable of selecting and engaging targets without human authorisation

Sources

Related news