ArXivCSExplorer
☆☆Bookmarks🏆RSSHow to UseFAQ
Built with and by Teycir Ben Soltane•
How to Use•FAQ•GitHub•arXiv.org•
Share:

20 results for “Traceability graph”

CS papers only

Hybrid search: Keyword + semantic, ranked by combined score.ⓘ

Want pure semantic search? Try claim verification →

cs.SEEmpiricalRecentJul 21, 2026

TraceDev: A Traceability-Driven Multi-agent Framework for Requirement-to-Code Development

Mingyu Chen, Yakun Zhang, Zihao Xie, Yixing Luo +4 more

The paper proposes TraceDev, a multi-agent framework for automated software development grounded in use cases, achieving higher success rates than baseline approaches in repository-level code generati…

View →
cs.LGEmpiricalRecentJun 30, 2026

FedLAB: Traceable Semantic Codebooks for Federated Multimodal Graph Foundation Learning

Zekai Chen, Kairui Yang, Xuaner Chen, Xunkai Li +3 more

The paper proposes FedLAB, a traceable semantic codebook framework for federated multimodal graph foundation learning, which organizes multimodal graph knowledge into hierarchical codebooks and refine…

View →
cs.AIRecentMay 28, 2026

Citation-Closure Retrieval and Per-Rule Attribution for Real-World Regulatory Compliance Question Answering

Yeong-Joon Ju, Seong-Whan Lee

The paper introduces RefWalk, a novel framework designed to improve regulatory compliance question answering by ensuring rigorous citation traceability and explicit per-rule attribution across complex…

View →
cs.AIcs.CRcs.SERecentMay 24, 2026

Inverting the Shield: Systematically Generating Safety Tests from Policy Specifications

Xiaoyue Lu, Xianglin Yang, Haijun Liu, Jiahao Liu +3 more

The paper introduces POLARIS, a novel framework that systematically generates comprehensive and verifiable safety tests for LLMs by formalizing natural language policies into First-Order Logic and exp…

View →
cs.LOcs.CRcs.FLRecentMar 20, 2026

Agentproof: Static Verification of Agent Workflow Graphs

Melwin Xavier, Vaisakh M A, Melveena Jolly, Midhun Xavier

Agentproof is a system that provides static, pre-deployment verification of safety properties in agent workflow graphs by automatically extracting a unified graph model and applying structural and tem…

View →
cs.SEcs.AIcs.MAEmpiricalRecentJul 16, 2026

StructureClaw: Traceable LLM Agents and an Executable Benchmark for Structural Engineering Workflows

Sizhong Qin, Yi Gu, Yao Jiang, Ao Cai +12 more

This paper introduces StructureClaw, an artifact-centered workbench for evaluating structural-engineering agents, and presents StructureClaw-Bench, an executable benchmark for testing these agents.

View →
cs.IRcs.AIcs.CLTheoreticalRecentJul 3, 2026

TRIAGE: Trustworthy Retrieval Instrumentation And Graph Evaluation

Axel TahmasebiMoradi, Lucas Schott, Martin Royer

TRIAGE is a framework for evaluating and diagnosing failures in document-grounded graph-RAG systems by attaching stage-specific metrics.

View →
cs.DScs.CCTheoreticalRecentJun 11, 2026

Sketching Intersection Profiles: A Simple Proof and Three Applications

Flavio Chierichetti, Mirko Giacchini, Ravi Kumar, Alessandro Panconesi +2 more

This paper settles the complexity of three sketching problems in graphs and distributions.

View →
cs.SEcs.AIRecentMay 31, 2026

FVSpec: Real-World Property-Based Tests as Lean Challenges

Quinn Dougherty, Max von Hippel, Hazel Shackleton, Mike Dodds

The paper introduces FVSpec, a large-scale benchmark that translates thousands of real-world Python property-based tests into formal Lean 4 specifications to evaluate AI models for formal software ver…

View →
cs.ARcs.LGEmpiricalRecentJul 4, 2026

Weave: Verified Netlist-to-Schematic Conversion via Layered Graph Layout

Senol Gulgonul

This paper presents Weave, a deterministic method to convert SPICE netlists into LTspice schematics with guaranteed connectivity preservation.

View →
cs.CYcs.CLRecentMay 29, 2026

Traceable by Design: An LLM Pipeline and Dashboard for EU Regulatory Consultation Analysis

Thales Bertaglia, Haoyang Gui, Catalina Goanta, Gerasimos Spanakis

The paper presents an end-to-end LLM pipeline and interactive dashboard designed to automatically extract and structure topics from massive volumes of regulatory consultation submissions, ensuring ful…

View →
cs.NIEmpiricalRecentJun 23, 2026

Overconfident Coordinates: Quantifying Confidence in Traceroute Geolocation

Santiago Klein, Caleb J. Wang, Fabián E. Bustamante

This paper introduces Path Consistency Scoring (PCS), a framework that evaluates router geolocation as a path-level consistency problem and produces a path consistency score based on a Hidden Markov M…

View →
cs.SEcs.LOcs.PLEmpiricalRecentJul 1, 2026

Kani: A Model Checker for Rust

Rémi Delmas, Zyad Hassan, Qinheping Hu, Rahul Kumar +8 more

Kani is an open-source model checker for Rust that provides correctness guarantees for safety properties through compilation and a specification language.

View →
cs.AIcs.CLcs.LGRecentMay 28, 2026

Conformal Certification of Reasoning Trace Prefixes

Matt Y. Cheung, Ashok Veeraraghavan, Hanjie Chen, Guha Balakrishnan

The paper introduces CROP, a novel conformal procedure that provides rigorous statistical guarantees for certifying the longest safe prefix of a language model's reasoning trace, allowing for targeted…

View →
cs.SEcs.AIRecentMay 27, 2026

Tool Forge: A Validation-Carrying Toolchain for Governed Agentic Execution

Swanand Rao

Tool Forge is a validation-carrying toolchain that converts natural language capability intent into governed, sandbox-verified tool artifacts, significantly improving agent efficiency and reliability.

View →