All Tags
Browse through all available tags to find articles on topics that interest you.
Browse through all available tags to find articles on topics that interest you.
Showing 2 results for this tag.
Falling Behind Drives Unsafe Development in an Idealised AI Race Experiment
This paper explores how competitive pressure in AI development influences safety choices. Through a behavioral experiment and an evolutionary model, it demonstrates that unsafe development is primarily driven by fear of falling behind and strategic responses to competitors, rather than individual risk preferences alone.
Empathy Modeling in Active Inference Agents for Perspective-Taking and Alignment
This paper introduces an active inference computational framework for empathy in AI agents, enabling explicit perspective-taking through a self-other model transformation. It demonstrates that empathic perspective-taking can induce robust cooperation in strategic dilemmas like the Iterated Prisoner's Dilemma, highlighting empathy as a structural prior for socially aligned AI.