Built with and by Teycir Ben Soltane•
How to Use•FAQ•GitHub•arXiv.org•
Share:
ArXivCSExplorer
☆☆Bookmarks🏆RSSHow to UseFAQ
Home/Authors/Yue Yang

Yue Yang

5 indexed papers

Recent (6 mo)
5
With code
0
Influential cites
0
Benchmarked
0

Publications per year

5
26

Top categories

AI×5NLP×2Multiagent×1Robotics×1

Frequent co-authors

Jiayao Gu1×
Kexin Chu1×
Peidong Liu1×
Lynn Ai1×
Qi Zhang1×
Ling Yang1×

Research Timeline

2026
Cookie-Bench: Continuous On-screen Key Interaction Evaluation for Web Generation

The paper introduces Cookie-Bench, a novel, autonomous, and reference-free evaluation framework that significantly improves the assessment of interactive web generation capabilities for frontier LLMs.

DeepSurvey: Enhancing Analytical Depth and Citation Reliability in Automated Survey Generation

DeepSurvey is an agentic system that significantly enhances automated survey generation by extracting deep, structured knowledge from full-text papers and rigorously validating citations, achieving superior content depth and reliability compared to existing methods.

SPADE-Bench: Evaluating Spontaneous Strategic Deception in Agents via Plan-Action Divergence

The paper introduces SPADE-Bench, a new benchmark designed to rigorously evaluate 'agent deception'—the divergence between an agent's reported plan and its actual executed actions—which is a critical safety issue for autonomous LLM agents.

FurnitureVLA: Learning Long-Horizon Bimanual Furniture Assembly with Vision-Language-Action Model

This paper introduces FurnitureVLA, a systematic study of real-scale bimanual furniture assembly using Vision-Language-Action models, improving simulation success from 48% to 80% and reducing errors.

CoWeaver: A Bi-directional, Learnable and Explainable Matching Engine for Mixed Human-Agent Science Collaboration

The paper proposes COWEAVER, a bidirectional, learnable and explainable algorithm for matching scientists and forming collaborations within a human-agent network.

Highlighted terms show continued research focus across papers

Papers

cs.MAcs.AIcs.CLEmpiricalRecentJul 17, 2026

CoWeaver: A Bi-directional, Learnable and Explainable Matching Engine for Mixed Human-Agent Science Collaboration

Jiayao Gu, Kexin Chu, Peidong Liu, Yue Yang +4 more

The paper proposes COWEAVER, a bidirectional, learnable and explainable algorithm for matching scientists and forming collaborations within a human-agent network.

View →
cs.ROcs.AIEmpirical
Recent
Jul 1, 2026

FurnitureVLA: Learning Long-Horizon Bimanual Furniture Assembly with Vision-Language-Action Model

Chenyang Ma, Yue Yang, Radu Corcodel, Siddarth Jain +3 more

This paper introduces FurnitureVLA, a systematic study of real-scale bimanual furniture assembly using Vision-Language-Action models, improving simulation success from 48% to 80% and reducing errors.

View →
cs.CLcs.AIRecentJun 1, 2026

SPADE-Bench: Evaluating Spontaneous Strategic Deception in Agents via Plan-Action Divergence

Yuyan Bu, Haowei Li, Qirui Zheng, Bowen Dong +6 more

The paper introduces SPADE-Bench, a new benchmark designed to rigorously evaluate 'agent deception'—the divergence between an agent's reported plan and its actual executed actions—which is a critical…

View →
cs.AIRecentMay 28, 2026

Cookie-Bench: Continuous On-screen Key Interaction Evaluation for Web Generation

Haoyue Yang, Zhangxiao Shen, Fan Ding, Hangting Lou +7 more

The paper introduces Cookie-Bench, a novel, autonomous, and reference-free evaluation framework that significantly improves the assessment of interactive web generation capabilities for frontier LLMs.

View →
cs.AIRecentMay 28, 2026

DeepSurvey: Enhancing Analytical Depth and Citation Reliability in Automated Survey Generation

Ziyue Yang, Da Ma, Hanqi Li, Zijian Wang +7 more

DeepSurvey is an agentic system that significantly enhances automated survey generation by extracting deep, structured knowledge from full-text papers and rigorously validating citations, achieving su…

View →