Evan Hubinger is an AI safety researcher at Anthropic, known for work on deceptive alignment and inner misalignment. He develops alignment theories, evaluates model honesty, and writes on AI risk. His research guides labs, policymakers, and safety engineers working to build trustworthy, controllable AI systems.
Get alerts when this topic surges in newsletters. Free to start.
Sign up freeExplore more trends:Trending Topics ·AI Trends ·Business Trends ·Finance Trends ·Technology Trends