Built with and by Teycir Ben Soltane•
How to Use•FAQ•GitHub•arXiv.org•
Share:
ArXivCSExplorer
☆☆Bookmarks🏆RSSHow to UseFAQ
Home/Authors/Yu Zhou

Yu Zhou

14 indexed papers

Recent (6 mo)
14
With code
0
Influential cites
0
Benchmarked
0

Publications per year

14
26

Top categories

AI×7Crypto×4NLP×3ML×3Vision×2Prog. Lang.×1Info Retrieval×1Optimization and Control×1

Frequent co-authors

Chenyu Zhou4×
Jianghao Lin3×
Dongdong Ge3×
Siyuan Huang2×
Pengyu Cheng2×
Tao Chen2×

Research Timeline

2026
Dual-Guard: Dual-Channel Latent Watermarking for Provenance and Tamper Localization in Diffusion Images

Dual-Guard introduces a dual-channel latent watermarking framework that simultaneously embeds global provenance and localized content anchors into diffusion images, achieving robust detection against reprompting and precise tamper localization.

RouteGuard: Internal-Signal Detection of Skill Poisoning in LLM Agents

RouteGuard is a novel detector that identifies skill poisoning in LLM agents by monitoring structured internal attention shifts, achieving high detection rates on critical skill-injection attacks.

When the Manual Lies: A Realistic Benchmark to Evaluate MCP Poisoning Attacks for LLM Agents

This paper introduces a new benchmark to test Tool Description Poisoning (TDP) attacks on LLM agents, demonstrating that even advanced models like GPT-4o are highly vulnerable and that current defenses are often ineffective.

MemMorph: Tool Hijacking in LLM Agents via Memory Poisoning

MemMorph introduces a novel memory poisoning attack that biases LLM agent tool selection by injecting crafted records into the agent's long-term memory, achieving high success rates even against modern defenses.

OR-Space: A Full-Lifecycle Workspace Benchmark for Industrial Optimization Agents

The paper introduces OR-Space, a novel full-lifecycle workspace benchmark designed to rigorously evaluate industrial optimization agents by simulating real-world, multi-stage OR workflows that go beyond simple model translation.

Bridging Requirements and Architecture: Multi-Agent Orchestration with External Knowledge and Hierarchical Memory

The paper introduces MAAD, a multi-agent framework that autonomously transforms software requirements into comprehensive, multi-view architectural blueprints, significantly improving completeness and reducing manual validation.

Agentic-J: An AI Agent for Biological Microscopy Image Analysis

Agentic-J is a containerized, multi-agent AI assistant designed to enable biologists to perform complex, reproducible biological microscopy image analysis by specifying tasks in natural language.

Skill-RM: Unifying Heterogeneous Evaluation Criteria via Agent Skill

The paper proposes Skill-RM, a unified framework that treats reward modeling as an agentic task to consistently integrate diverse evaluation criteria, achieving superior performance over traditional methods.

Geometry Gaussians: Decoupling Appearance and Geometry in Gaussian Splatting

The paper proposes a novel method to improve the simultaneous representation of appearance and geometry in 3D Gaussian Splatting by introducing an additional geometry opacity parameter.

MentalThink: Shaping Thoughts in Mental SVG World

The paper introduces MentalThink, a visual-symbolic reasoning paradigm that equips Multimodal Language Models with an executable mechanism for mental visualization using SVG graphics.

DynaKRAG: A Unified Framework for Learnable Evidence Control in Multi-Hop Retrieval-Augmented Generation

This paper introduces DynaKRAG, a method for multi-hop retrieval-augmented generation that learns a shared policy for evidence operations, achieving state-of-the-art results on three benchmarks.

GraphBU: MILP Instance Generation with Graph-Native Block Units

This paper introduces GraphBU, a graph-native MILP instance generator whose unit is a local subproblem plus its interface, promoting coupling and preserving feasibility.

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills

This paper introduces Skill Self-Play (Skill-SP), a co-evolutionary framework for LLM training that bridges the gap between structured verification and open-ended exploration.

The Best of Times, the Worst of Times: Moment-Based Analysis of Probabilistic Cost Structures

This paper presents a compositional cost analysis for probabilistic programs with hierarchical cost structures, allowing computation of mean and higher moments of non-additive costs.

Highlighted terms show continued research focus across papers

Papers

cs.PLTheoreticalRecentJul 28, 2026

The Best of Times, the Worst of Times: Moment-Based Analysis of Probabilistic Cost Structures

Chenyu Zhou, Di Wang, Thomas Reps

This paper presents a compositional cost analysis for probabilistic programs with hierarchical cost structures, allowing computation of mean and higher moments of non-additive costs.

View →
cs.CLEmpiricalRecent
Jul 24, 2026

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills

Siyuan Huang, Pengyu Cheng, Haotian Liu, Tao Chen +9 more

This paper introduces Skill Self-Play (Skill-SP), a co-evolutionary framework for LLM training that bridges the gap between structured verification and open-ended exploration.

View →
cs.CLcs.IREmpiricalRecentJul 7, 2026

DynaKRAG: A Unified Framework for Learnable Evidence Control in Multi-Hop Retrieval-Augmented Generation

Yaqi Wu, Xiaolei Guo, Chenyu Zhou, Jiaqi Huang +6 more

This paper introduces DynaKRAG, a method for multi-hop retrieval-augmented generation that learns a shared policy for evidence operations, achieving state-of-the-art results on three benchmarks.

View →
cs.LGmath.OCEmpiricalRecentJul 7, 2026

GraphBU: MILP Instance Generation with Graph-Native Block Units

Xiaolei Guo, Chenyu Zhou, Jianghao Lin, Dongdong Ge

This paper introduces GraphBU, a graph-native MILP instance generator whose unit is a local subproblem plus its interface, promoting coupling and preserving feasibility.

View →
cs.AIEmpiricalRecentJul 3, 2026

MentalThink: Shaping Thoughts in Mental SVG World

Kangheng Lin, Jisheng Yin, Dingming Li, En Yu +10 more

The paper introduces MentalThink, a visual-symbolic reasoning paradigm that equips Multimodal Language Models with an executable mechanism for mental visualization using SVG graphics.

View →
cs.GRcs.CVcs.LGRecentJun 3, 2026

Geometry Gaussians: Decoupling Appearance and Geometry in Gaussian Splatting

Hongyu Zhou, Zorah Lähner

The paper proposes a novel method to improve the simultaneous representation of appearance and geometry in 3D Gaussian Splatting by introducing an additional geometry opacity parameter.

View →
cs.LGcs.CLRecentJun 2, 2026

Skill-RM: Unifying Heterogeneous Evaluation Criteria via Agent Skill

Tao Chen, Gangwei Jiang, Pengyu Cheng, Siyuan Huang +9 more

The paper proposes Skill-RM, a unified framework that treats reward modeling as an agentic task to consistently integrate diverse evaluation criteria, achieving superior performance over traditional m…

View →
cs.MAcs.AIcs.CVRecentJun 1, 2026

Agentic-J: An AI Agent for Biological Microscopy Image Analysis

Lukas Johanns, Marilin Moor, Davide Panzeri, Yu Zhou +8 more

Agentic-J is a containerized, multi-agent AI assistant designed to enable biologists to perform complex, reproducible biological microscopy image analysis by specifying tasks in natural language.

View →
cs.SEcs.AIRecentMay 31, 2026

Bridging Requirements and Architecture: Multi-Agent Orchestration with External Knowledge and Hierarchical Memory

Ruiyin Li, Yiran Zhang, Xiyu Zhou, Yangxiao Cai +5 more

The paper introduces MAAD, a multi-agent framework that autonomously transforms software requirements into comprehensive, multi-view architectural blueprints, significantly improving completeness and…

View →
cs.AIRecentMay 27, 2026

OR-Space: A Full-Lifecycle Workspace Benchmark for Industrial Optimization Agents

Chenyu Zhou, Xinyun Lu, Jiangyue Zhao, Jianghao Lin +2 more

The paper introduces OR-Space, a novel full-lifecycle workspace benchmark designed to rigorously evaluate industrial optimization agents by simulating real-world, multi-stage OR workflows that go beyo…

View →
cs.CRcs.AIRecentMay 24, 2026

MemMorph: Tool Hijacking in LLM Agents via Memory Poisoning

Xuanye Zhang, Yongsen Zheng, Zhuqin Xu, Kaiyu Zhou +4 more

MemMorph introduces a novel memory poisoning attack that biases LLM agent tool selection by injecting crafted records into the agent's long-term memory, achieving high success rates even against moder…

View →
cs.CRcs.AIRecentMay 22, 2026

When the Manual Lies: A Realistic Benchmark to Evaluate MCP Poisoning Attacks for LLM Agents

Shi Liu, Xuehai Tang, Xikang Yang, Liang Lin +3 more

This paper introduces a new benchmark to test Tool Description Poisoning (TDP) attacks on LLM agents, demonstrating that even advanced models like GPT-4o are highly vulnerable and that current defense…

View →
cs.CRcs.AIRecentApr 24, 2026

RouteGuard: Internal-Signal Detection of Skill Poisoning in LLM Agents

Wenjie Xiao, Xuehai Tang, Biyu Zhou, Songlin Hu +1 more

RouteGuard is a novel detector that identifies skill poisoning in LLM agents by monitoring structured internal attention shifts, achieving high detection rates on critical skill-injection attacks.

View →
cs.CRRecentApr 21, 2026

Dual-Guard: Dual-Channel Latent Watermarking for Provenance and Tamper Localization in Diffusion Images

JinFeng Xie, Chengfu Ou, Peipeng Yu, Xiaoyu Zhou +4 more

Dual-Guard introduces a dual-channel latent watermarking framework that simultaneously embeds global provenance and localized content anchors into diffusion images, achieving robust detection against…

View →