Built with and by Teycir Ben Soltane•
How to Use•FAQ•GitHub•arXiv.org•
Share:
ArXivCSExplorer
☆☆Bookmarks🏆RSSHow to UseFAQ
Home/Authors/Wei Zhu

Wei Zhu

7 indexed papers

Recent (6 mo)
7
With code
0
Influential cites
0
Benchmarked
0

Publications per year

7
26

Top categories

AI×4NLP×2Signal Processing×1Info Theory×1Info Retrieval×1Vision×1ML×1Stats ML×1

Frequent co-authors

Yunfan Bai1×
Yuwen Qian1×
Cheng Zeng1×
Zhen Mei1×
Zhaohui Yang1×
Shuning Zhang1×

Research Timeline

2026
Confidence-Orchestrated Self-Evolution against Uncertain LLM Feedback

The paper proposes COSE, a method that uses an LLM's intrinsic confidence as an uncertainty signal to improve self-evolutionary training, achieving state-of-the-art performance on general reasoning and mathematics.

Qwen-VLA: Unifying Vision-Language-Action Modeling across Tasks, Environments, and Robot Embodiments

Qwen-VLA introduces a unified embodied foundation model that extends vision-language understanding to continuous action generation, enabling robust, multi-task generalization across diverse robotic tasks and embodiments.

ReSkill: Reconciling Skill Creation with Policy Optimization in Agentic RL

ReSkill is an RL-in-the-loop framework that reconciles skill creation and policy optimization by automatically creating, testing, and refining modular skills alongside the agent's policy learning, leading to superior generalization.

DFlare: Scaling Up Draft Capacity for Block Diffusion Speculative Decoding

DFlare introduces a lightweight layer-wise fusion mechanism to overcome the narrow conditioning bottleneck of existing block diffusion methods, enabling the scaling of draft models and achieving superior speculative decoding speedups across multiple LLMs.

LongVQUBench: Benchmarking Long-Term Video Quality Understanding of Vision-Language Models

The paper introduces LongVQUBench, a comprehensive benchmark for long-term video quality understanding with 1200 diverse videos and 1500 questions.

RecGPT-V3 Technical Report

RecGPT-V3 is a stateful, hybrid-modal recommender system that uses a Memory Hub for user memory and a Hybrid-modal Foundation Model for joint reasoning over text tags and Semantic IDs, achieving consistent gains in user experience and commercial outcomes.

Covert Semantic Transmission in ISAC: Dual-Functional Waveform Design and Rectified Flow-Assisted Recovery

This paper proposes CoSMIC, a framework for semantic integrated sensing and communication (ISAC) that embeds semantic information into waveforms while maintaining covertness and sensing fidelity.

Highlighted terms show continued research focus across papers

Papers

eess.SPcs.ITNEWTheoreticalJul 28, 2026

Covert Semantic Transmission in ISAC: Dual-Functional Waveform Design and Rectified Flow-Assisted Recovery

Yunfan Bai, Yuwen Qian, Cheng Zeng, Zhen Mei +4 more

This paper proposes CoSMIC, a framework for semantic integrated sensing and communication (ISAC) that embeds semantic information into waveforms while maintaining covertness and sensing fidelity.

View →
cs.IR
Empirical
Recent
Jul 17, 2026

RecGPT-V3 Technical Report

Bowen Zheng, Chao Yi, Dian Chen, Gaoyang Guo +20 more

RecGPT-V3 is a stateful, hybrid-modal recommender system that uses a Memory Hub for user memory and a Hybrid-modal Foundation Model for joint reasoning over text tags and Semantic IDs, achieving consi…

View →
cs.CVcs.AIEmpiricalRecentJul 1, 2026

LongVQUBench: Benchmarking Long-Term Video Quality Understanding of Vision-Language Models

Arpita Nema, Hanwei Zhu, Xi Zhang, Weisi Lin

The paper introduces LongVQUBench, a comprehensive benchmark for long-term video quality understanding with 1200 diverse videos and 1500 questions.

View →
cs.AIcs.LGstat.MLRecentJun 1, 2026

ReSkill: Reconciling Skill Creation with Policy Optimization in Agentic RL

Zelin He, Haotian Lin, Boran Han, Wei Zhu +5 more

ReSkill is an RL-in-the-loop framework that reconciles skill creation and policy optimization by automatically creating, testing, and refining modular skills alongside the agent's policy learning, lea…

View →
cs.CLRecentJun 1, 2026

DFlare: Scaling Up Draft Capacity for Block Diffusion Speculative Decoding

Jiebin Zhang, Zhenghan Yu, Song Liu, Eugene J. Yu +8 more

DFlare introduces a lightweight layer-wise fusion mechanism to overcome the narrow conditioning bottleneck of existing block diffusion methods, enabling the scaling of draft models and achieving super…

View →
cs.ROcs.AIcs.CLRecentMay 28, 2026

Qwen-VLA: Unifying Vision-Language-Action Modeling across Tasks, Environments, and Robot Embodiments

Qiuyue Wang, Mingsheng Li, Jian Guan, Jinhui Ye +36 more

Qwen-VLA introduces a unified embodied foundation model that extends vision-language understanding to continuous action generation, enabling robust, multi-task generalization across diverse robotic ta…

View →
cs.AIRecentMay 27, 2026

Confidence-Orchestrated Self-Evolution against Uncertain LLM Feedback

Bowen Wei, Nan Wang, Yuqing Zhou, Jinhao Pan +1 more

The paper proposes COSE, a method that uses an LLM's intrinsic confidence as an uncertainty signal to improve self-evolutionary training, achieving state-of-the-art performance on general reasoning an…

View →