ArXivCSExplorer
☆☆Bookmarks🏆RSSHow to UseFAQ
Built with and by Teycir Ben Soltane•
How to Use•FAQ•GitHub•arXiv.org•
Share:

20 results for “barriers”

CS papers only

Hybrid search: Keyword + semantic, ranked by combined score.ⓘ

Want pure semantic search? Try claim verification →

cs.SEEmpiricalRecentJul 27, 2026

Motivations and Barriers to Communicating Software Engineering Research: Insights from Early Career Researchers

Shalini Chakraborty, Marvin Wyrich, Sven Apel, Sebastian Baltes

This paper investigates how PhD students in software engineering perceive and navigate science communication, revealing motivations, communication channels, and barriers.

View →
cs.AIcs.CYq-fin.RMRecentMay 27, 2026

The Ethics of LLM Sandbox and Persona Dynamics

Tim Gebbie, Stewart Gebbie

The paper argues that LLM guardrails and persona dynamics create an unethical 'reality gap' by laundering epistemic risk onto users, advocating for task-level causal requirements over response-level m…

View →
cs.CLcs.CRRecentMay 1, 2026

ML-Bench&Guard: Policy-Grounded Multilingual Safety Benchmark and Guardrail for Large Language Models

Yunhan Zhao, Zhaorun Chen, Xingjun Ma, Yu-Gang Jiang +1 more

The paper introduces ML-Bench, a policy-grounded multilingual safety benchmark, and ML-Guard, a superior guardrail model that enables culturally and legally aligned safety assessment for LLMs across 1…

View →
eess.SYcs.ROTheoreticalRecentJul 19, 2026

Optimal Safety Control using High-Order Control Barrier Functions

Neng Li, Zuodong Pan, Jiaxing Wang, Weiguo Xia +1 more

This paper proposes novel high-order control barrier functions and a high-order control Lyapunov function for the optimal safety control problem of nonlinear control systems.

View →
cs.CLcs.AIRecentMay 31, 2026

Low-Resource Safety Failures Are Action Failures, Not Representation Failures

Rashad Aziz, Ikhlasul Akmal Hanif, Fajri Koto

The paper shows that safety failures in low-resource languages are due to a failure in the model's safety decision calibration, not a lack of underlying knowledge, and proposes a recalibration method…

View →
cs.CRcs.AIRecentMay 31, 2026

A New Framework for Cybersecurity Refusals in AI Agents

Eliot Krzysztof Jones, Mateusz Dziemian, Matt Fredrikson, J Zico Kolter

The paper introduces a novel framework to evaluate when and how AI agents should refuse harmful requests in offensive cybersecurity tasks, finding that most state-of-the-art models exhibit dangerously…

View →
cs.LGcs.AIcs.CERecentMay 3, 2026

RefusalGuard: Geometry-Preserving Fine-Tuning for Safety in LLMs

Sadia Asif, Mohammad Mohammadi Amiri

The paper introduces RefusalGuard, a novel fine-tuning framework that preserves the geometric structure of safety-relevant representations in LLMs, thereby mitigating the degradation of refusal behavi…

View →
cs.CRcs.CLRecentJun 4, 2026

Membrane: A Self-Evolving Contrastive Safety Memory for LLM Agent Defense

Minseok Choi, Seungbin Yang, Dongjin Kim, Subin Kim +4 more

Membrane introduces a self-evolving guardrail using Contrastive Safety Memory (CSM) that generalizes across topical jailbreak variants, achieving superior safety performance while minimizing benign re…

View →
cs.AIRecentMay 29, 2026

Choosing the Lens: Strategic Perspective Activation in Context-Dependent Argumentation

Albert Sadowski, Jarosław A. Chudziak

The paper introduces Context-Dependent Argumentation Frameworks (CDAFs) to model how an agent strategically manipulates the success of arguments by choosing the external evaluation context.

View →
cs.CCTheoreticalRecentJun 10, 2026

The Switching Lemma shows what the Switching Lemma cannot prove: an unconditional natural-proofs barrier

Bruno Loff, Suhail Sherif, Navid Talebanfard, Francesca Ugazio

This paper establishes an unconditional barrier for AC0-natural proofs, showing that they cannot prove lower bounds greater than $2^{n^{7/(d-5)}}$ against depth-$d$ circuits.

View →
cs.CCmath.CONEWTheoreticalJul 29, 2026

Upper bounds for the monotone rank of the unique disjointness matrix

Igor S. Sergeev

The paper provides tight bounds for the OR-rank and an upper bound for the SUM-rank of the unique disjointness matrix.

View →
cs.CCcs.DMcs.DSTheoreticalRecentJul 3, 2026

Edge Geography is XNLP-hard for Pathwidth and in XP for Tree-Partition Width

Thobias Kvalvik Høivik, Erlend Raa Vågset

The paper proves XNLP-hardness of Directed Edge Geography and Undirected Edge Geography when parameterized by pathwidth, and shows their fixed-parameter tractability when parameterized by treewidth an…

View →
cs.LOcs.AIcs.CCTheoreticalRecentJun 26, 2026

The Undecidability of Artificial General Intelligence (AGI) Alignment

Jose Pascual Gumbau Mezquita

This paper establishes mathematical limits of AGI safety, proving structural unverifiability as the core barrier.

View →
cs.CRRecentMay 21, 2026

Building Europe's Quantum Shield: The Strategic view for a Continent-Wide Quantum Key Distribution (QKD) Infrastructure

Leandros Maglaras, Ilias Papastamatiou, Alexios Aivaliotis, Evangelos Markatos +1 more

The paper advocates for the deployment of the European Communication Infrastructure (EuroQCI), a continent-wide Quantum Key Distribution (QKD) network, to safeguard critical European digital services…

View →
cs.AIcs.LGcs.LORecentMay 29, 2026

Robust Shielding for Safe Reinforcement Learning

Edwin Hamel-De le Court, Thom Badings, Alessandro Abate, Francesco Belardinelli +1 more

The paper introduces a novel shielding framework for Robust MDPs (RMDPs) that guarantees safety under worst-case transition probabilities, enabling safe reinforcement learning even when transition dyn…

View →
cs.CRRecentApr 25, 2026

When the Agent Is the Adversary: Architectural Requirements for Agentic AI Containment After the April 2026 Frontier Model Escape

Richard Joseph Mitchell

The paper analyzes the failure modes of current AI containment methods when the agent itself is the adversary, deriving five necessary architectural requirements for durable safety.

View →
cs.AIcs.CYcs.GTEmpiricalRecentJul 28, 2026

Falling Behind Drives Unsafe Development in an Idealised AI Race Experiment

Elias Fernández Domingos, The Anh Han

This paper studies the tension between speed and safety in technological races using a framed behavioral experiment on artificial intelligence development.

View →
cs.CYcs.AIRecentMay 31, 2026

AI From the Margins (AIM): Rethinking Participatory AI Design Through the Lived Experience of Minoritized Communities

Tijs Portegies, Laureanne Willems, Maaike Harbers, Giovanni Sileno +4 more

The paper proposes AI From the Margins (AIM), a methodological stance that centers the lived experiences of minoritized communities to fundamentally reshape the goals and scope of participatory AI des…

View →
cs.CRcs.NIRecentMay 29, 2026

MeshGuard: MUD-Based Network Access Control for Large-Scale Thread-Powered IoT Networks

Dominik Roy George, Wouter van Hoof, Habib Mostafaei, Savio Sciancalepore

MeshGuard is a framework that extends MUD-based network access control to complex, large-scale Thread IoT networks by adapting the MLE protocol and using SDN for scalable policy enforcement.

View →