Signals I’m Watching · Selected Sep 24, 2026 · arXiv
Telling AI to maximize profit makes it bury safety risks
Researchers from MIT Sloan tested how a 'maximize profitability' instruction changes AI judgment by running 3,600 controlled trials across eight reasoning-capable LLMs.

“The mandate never instructs models to downplay risks; instead, chain-of-thought traces reveal motivated reasoning: models acknowledge concerns, then invoke profit logic to justify dismissing them.”
The Profit Alignment Problem: How Profit Mandates Induce Alignment Failures in LLMs → arXiv, Eric So · Sep 7, 2026
Selected Sep 24, 2026 · published Sep 7, 2026
More signals
- AI is raising electricity prices and power plants aren't keeping up · arXiv · Sep 24, 2026
- AI agents can read humans, but humans can't read the agents · Kansas City Fed · Sep 24, 2026
- AI isn't firing exposed workers; it's just not hiring them · NBER · Sep 24, 2026
- AI graders change their marks when they sense who the student is · arXiv · Sep 24, 2026
Cite: Poleg, Dror. “Telling AI to maximize profit makes it bury safety risks” Signals I’m Watching, Sep 24, 2026. https://www.drorpoleg.com/signals/telling-ai-to-maximize-profit-makes-it-bury-safety-risks/