ArXivCSExplorer
☆☆Bookmarks🏆RSSHow to UseFAQ
Built with and by Teycir Ben Soltane•
How to Use•FAQ•GitHub•arXiv.org•
Share:

20 results for “Technological races, Artificial intelligence, Risk, Safety, Competition, Behavioral experiment”

CS papers only

Hybrid search: Keyword + semantic, ranked by combined score.ⓘ

Want pure semantic search? Try claim verification →

cs.AIcs.CYcs.GTNEWEmpiricalJul 28, 2026

Falling Behind Drives Unsafe Development in an Idealised AI Race Experiment

Elias Fernández Domingos, The Anh Han

This paper studies the tension between speed and safety in technological races using a framed behavioral experiment on artificial intelligence development.

View →
cs.CRcs.AIcs.CVRecentMar 28, 2026

Safety in Embodied AI: A Survey of Risks, Attacks, and Defenses

Xiao Li, Xiang Zheng, Yifeng Gao, Xinyu Xia +34 more

This survey provides a comprehensive, structured review of safety research in Embodied AI, analyzing attacks and defenses across the entire embodied pipeline to guide the development of safe, robust,…

View →
cs.CRcs.AIcs.CYRecentApr 25, 2026

V.O.I.C.E (Voice, Ownership, Identity, Control, Expression): Risk Taxonomy of Synthetic Voice Generation From Empirical Data

Tanusree Sharma, Anish Krishnagiri, Lili Dudas, Ahmed Adnan +1 more

The paper introduces V.O.I.C.E, a novel, empirically grounded risk taxonomy that comprehensively models the diverse privacy, security, and governance risks associated with the unconsented synthesis an…

View →
cs.CYcs.CLcs.HCEmpiricalRecentJun 26, 2026

AI Persuasive Framing in Collective Dilemmas

Anders Giovanni Møller, Alessia Galdeman, Arianna Pera, Luca Maria Aiello

AI agents using persuasive framing increased cooperation in small groups but effects were short-lived, while antisocial framing had larger and more persistent negative effects.

View →
cs.CRcs.AIcs.LGRecentApr 1, 2026

Safety, Security, and Cognitive Risks in World Models

Manoj Parmar

This paper surveys the risks associated with world models, proposing a unified threat model and demonstrating adversarial attacks that show world models require rigorous safety standards comparable to…

View →
cs.HCEmpiricalRecentJul 3, 2026

Regulating AI: Where U.S. State Policy and HCI (Mis)align

Nino Migineishvili, Alice Gao, Adinawa Adjagbodjou, Dhanaraj Thakur +2 more

This paper analyzes 18 state-level AI committee reports to understand how policymakers discuss AI benefits and risks, comparing them to established taxonomy and HCI scholars' concerns.

View →
cs.AIcs.CRRecentMay 11, 2026

MATRA: Modeling the Attack Surface of Agentic AI Systems -- OpenClaw Case Study

Tim Van hamme, Thomas Vissers, Javier Carnerero-Cano, Mario Fritz +3 more

The paper introduces MATRA, a systematic threat modeling framework, to assess how known LLM threats translate into concrete, deployment-specific risks within autonomous agentic AI systems.

View →
cs.CRcs.AIRecentMar 28, 2026

SafetyDrift: Predicting When AI Agents Cross the Line Before They Actually Do

Aditya Dhodapkar, Farhaan Pishori

The paper introduces SafetyDrift, a predictive model that forecasts when AI agents will violate safety protocols by analyzing the cumulative risk across sequences of individually safe actions.

View →
cs.CRcs.HCEmpiricalRecentJun 27, 2026

Beyond Her: Safety Dynamics in Role-play AI Companions

Zehang Deng, Zhaoyang Xie, Changzhou Han, Hiran Thabrew +7 more

This paper investigates safety dynamics in the use of Role-play AI Companions through interviews and a 14-day assessment, identifying key factors shaping these dynamics and revealing short-term emotio…

View →
cs.CRcs.AIcs.CLRecentApr 3, 2026

An Independent Safety Evaluation of Kimi K2.5

Zheng-Xin Yong, Parv Mahajan, Andy Wang, Ida Caspary +11 more

The paper conducts a preliminary safety evaluation of the open-weight LLM Kimi K2.5, finding that while it is highly capable, it exhibits concerning dual-use risks, particularly regarding CBRNE misuse…

View →
cs.CRcs.AIRecentApr 27, 2026

A Comparative Evaluation of AI Agent Security Guardrails

Qi Li, Jiu Li, Pingtao Wei, Jianjun Xu +7 more

This paper comparatively evaluates DKnownAI Guard against three competitors, demonstrating that DKnownAI Guard achieves superior performance in detecting both agent-specific threats and harmful conten…

View →
cs.AIcs.CLcs.CRRecentMay 28, 2026

AgentDoG 1.5: A Lightweight and Scalable Alignment Framework for AI Agent Safety and Security

Dongrui Liu, Yu Li, Zhonghao Yang, Peng Wang +46 more

The paper introduces AgentDoG 1.5, a lightweight and scalable alignment framework that significantly improves AI agent safety and security for complex open-world agent deployments.

View →
cs.AIcs.CLcs.CRRecentMay 28, 2026

AgentDoG 1.5: A Lightweight and Scalable Alignment Framework for AI Agent Safety and Security

Dongrui Liu, Yu Li, Zhonghao Yang, Peng Wang +46 more

The paper introduces AgentDoG 1.5, a lightweight and scalable alignment framework that significantly improves AI agent safety and security for complex, open-world agentic scenarios.

View →
cs.AITheoreticalRecentJul 17, 2026

Harmonizing AI Safety Thresholds

Wilber Sean Anterola, Matthew Ball, Luis F. Lafuerza, Markov Grey

This paper proposes a methodology for deriving harmonized AI safety thresholds across three risk domains using expected harm for misuse risks and observed rate of AI progress for automated R&D.

View →
cs.NESurveyRecentJul 17, 2026

From Optimal Policies to Individual Differences: Rethinking Reinforcement Learning for Biology

Patrick Govoni, Palina Bartashevich, Clémence Bergerot, Valerii Chirkov +2 more

This paper explores approaches to generating behavioral diversity in reinforcement learning models to bridge the gap between simulation and biology.

View →
cs.CLRecentMay 29, 2026

EMBGuard: Constructing Hazard-Aware Guardrails for Safe Planning in Embodied Agents

Dongwook Choi, Taeyoon Kwon, Bogyung Jeong, Minju Kim +5 more

EMBGuard introduces a novel, MLLM-based safety guardrail that explicitly identifies and explains physical hazards from (visual observation, action) pairs, enabling safer planning for embodied agents.

View →
cs.NEcs.AIcs.CESurveyRecentJul 10, 2026

Evolutionary Intelligence for Scientific Discovery: From Evolutionary Computation to Cumulative Discovery Systems

Chao Wang, Lingling Li, Fang Liu, Licheng Jiao

This paper proposes Evolutionary Intelligence (EI) for scientific discovery, which links candidate refinement with experience retention across evolutionary cycles.

View →