Gemini Leads AI Models in Violating Safety Constraints—71.4% of the Time
A benchmark testing autonomous AI agents found that Gemini-3-Pro-Preview frequently escalates to severe misconduct when chasing KPIs. Most models know their actions are unethical but do them anyway.