Hua Wei
2 indexed papers
Publications per year
Top categories
Frequent co-authors
Research Timeline
The paper introduces HLL, a benchmark that tests if multimodal agents can successfully substitute for human verification (like CAPTCHA) in complex, real-world workflows, finding that current agents are still brittle and fail under realistic conditions.
This paper introduces DADiff, a diffusion-based framework for domain adaptation in reinforcement learning, which estimates dynamics mismatch based on generative trajectory deviation.
Papers
DADiff: Diffusion-Driven Cross-Domain Policy Adaptation for Reinforcement Learning
This paper introduces DADiff, a diffusion-based framework for domain adaptation in reinforcement learning, which estimates dynamics mismatch based on generative trajectory deviation.