Built with and by Teycir Ben Soltane•
How to Use•FAQ•GitHub•arXiv.org•
Share:
ArXivCSExplorer
☆☆Bookmarks🏆RSSHow to UseFAQ
Home/Authors/Bin Zhang

Bin Zhang

11 indexed papers

Recent (6 mo)
11
With code
0
Influential cites
0
Benchmarked
0

Publications per year

11
26

Top categories

Info Retrieval×3Crypto×2AI×2NLP×2Info Theory×1Audio and Speech Processing×1Sound×1Multimedia×1

Frequent co-authors

Yu Cui2×
Ruiqing Yue2×
Sicheng Pan2×
Zhuoyu Sun2×
Baohan Huang2×
Haibin Zhang2×

Research Timeline

2026
Spore: Efficient and Training-Free Privacy Extraction Attack on LLMs via Inference-Time Hybrid Probing

The paper introduces extsc{Spore}, a novel, training-free, and highly efficient privacy extraction attack that targets sensitive information stored in the memory of LLM agents during inference, outperforming existing state-of-the-art methods.

HASTE: Hardware-Aware Dynamic Sparse Training for Large Output Spaces

HASTE introduces group-shared fixed fan-in sparsity for multi-label classification, achieving significant wall-clock speedups (up to 25x in backward pass) by enabling efficient GPU execution while maintaining high accuracy.

DFlare: Scaling Up Draft Capacity for Block Diffusion Speculative Decoding

DFlare introduces a lightweight layer-wise fusion mechanism to overcome the narrow conditioning bottleneck of existing block diffusion methods, enabling the scaling of draft models and achieving superior speculative decoding speedups across multiple LLMs.

HarnessForge: Joint Harness and Policy Evolution for Adaptive Agent Systems

HarnessForge introduces a meta-adaptive framework that jointly evolves the execution structure (harness) and the reasoning policy of LLM agents, significantly improving overall system performance across diverse tasks.

Hybrid Diffusion Transformer for Instruction-Guided Audio Editing via Rectified Flow

This paper proposes a hybrid two-stage diffusion transformer architecture for instruction-guided audio editing, balancing performance and efficiency.

GigaSpeechBench: A Real-World Multilingual Speech-to-Text Benchmark

The paper introduces GigaSpeechBench, a comprehensive multilingual and multidimensional ASR & AST benchmark with 680 hours of human-annotated speech, featuring 12 low-resource languages, 6 Chinese dialects, 6 English accents, dense terminology, older adult and child speech, and human-annotated translations.

From Bit to Block: Capacity Achievement via Product Coding

This paper presents a product coding scheme that converts bit-level reliability into block-level reliability, achieving the same asymptotic rate.

Log-Insight: Automating Microservice Incident Diagnosis via Neuro-Symbolic Log Analysis

Log-Insight is an automated incident-diagnosis system for large-scale microservice systems that reduces raw events by 1,000-7,000x while preserving statistically significant failure signals, achieving high accuracy in under a minute.

Refusal is Not Safety! Benchmarking Latent Safety Risks of LLM-Driven Content Humorization

This paper explores safety risks in humorization of large language models (LLMs) and introduces HumorSafe framework for evaluating latent safety risks.

MagicSelector: Joint Optimization for Agent Tool Selection via Counterfactual Decomposition and Progressive Reranking

MagicSelector is a framework for tool retrieval in agents using counterfactual task decomposition, progressive reranking, and dynamic Top-K.

UniRank: Benchmarking Ranking Models for Unified Sequential Modeling and Feature Interaction

This paper introduces UniRank, an open benchmark for comparing and studying unified ranking models that combine sequential modeling and feature interaction.

Highlighted terms show continued research focus across papers

Papers

cs.IREmpiricalRecentJul 22, 2026

UniRank: Benchmarking Ranking Models for Unified Sequential Modeling and Feature Interaction

Honghao Li, Xianquan Wang, Zibin Zhang, Yi Zhang +2 more

This paper introduces UniRank, an open benchmark for comparing and studying unified ranking models that combine sequential modeling and feature interaction.

View →
cs.IREmpiricalRecent
Jul 20, 2026

MagicSelector: Joint Optimization for Agent Tool Selection via Counterfactual Decomposition and Progressive Reranking

HONOR Agentic Search Team, Zhengzong Chen, Lei Tang, Lijun Liu +26 more

MagicSelector is a framework for tool retrieval in agents using counterfactual task decomposition, progressive reranking, and dynamic Top-K.

View →
cs.CREmpiricalRecentJul 17, 2026

Refusal is Not Safety! Benchmarking Latent Safety Risks of LLM-Driven Content Humorization

Yu Cui, Ruiqing Yue, Tingyu Li, Sicheng Pan +5 more

This paper explores safety risks in humorization of large language models (LLMs) and introduces HumorSafe framework for evaluating latent safety risks.

View →
cs.IREmpiricalRecentJul 9, 2026

Log-Insight: Automating Microservice Incident Diagnosis via Neuro-Symbolic Log Analysis

Carlos Garcia-Hernandez, Aymane Abdali, Guangyu Wu, Mingxue Wang +3 more

Log-Insight is an automated incident-diagnosis system for large-scale microservice systems that reduces raw events by 1,000-7,000x while preserving statistically significant failure signals, achieving…

View →
cs.ITTheoreticalRecentJul 7, 2026

From Bit to Block: Capacity Achievement via Product Coding

Bin Zhang

This paper presents a product coding scheme that converts bit-level reliability into block-level reliability, achieving the same asymptotic rate.

View →
eess.ASEmpiricalRecentJun 27, 2026

GigaSpeechBench: A Real-World Multilingual Speech-to-Text Benchmark

Yujie Tu, Yifan Yang, Tianrui Wang, Yanqiao Zhu +32 more

The paper introduces GigaSpeechBench, a comprehensive multilingual and multidimensional ASR & AST benchmark with 680 hours of human-annotated speech, featuring 12 low-resource languages, 6 Chinese dia…

View →
cs.SDcs.AIcs.MMEmpiricalRecentJun 18, 2026

Hybrid Diffusion Transformer for Instruction-Guided Audio Editing via Rectified Flow

Liting Gao, Yonggang Zhu, Yaru Chen, Dongyu Wang +4 more

This paper proposes a hybrid two-stage diffusion transformer architecture for instruction-guided audio editing, balancing performance and efficiency.

View →
cs.CLRecentJun 1, 2026

DFlare: Scaling Up Draft Capacity for Block Diffusion Speculative Decoding

Jiebin Zhang, Zhenghan Yu, Song Liu, Eugene J. Yu +8 more

DFlare introduces a lightweight layer-wise fusion mechanism to overcome the narrow conditioning bottleneck of existing block diffusion methods, enabling the scaling of draft models and achieving super…

View →
cs.CLRecentJun 1, 2026

HarnessForge: Joint Harness and Policy Evolution for Adaptive Agent Systems

Mingju Chen, Can Lv, Guibin Zhang, Heng Chang +1 more

HarnessForge introduces a meta-adaptive framework that jointly evolves the execution structure (harness) and the reasoning policy of LLM agents, significantly improving overall system performance acro…

View →
cs.LGcs.AIRecentMay 31, 2026

HASTE: Hardware-Aware Dynamic Sparse Training for Large Output Spaces

Nasib Ullah, Jinbin Zhang, Jean Lucien Randrianantenaina, Erik Schultheis +1 more

HASTE introduces group-shared fixed fan-in sparsity for multi-label classification, achieving significant wall-clock speedups (up to 25x in backward pass) by enabling efficient GPU execution while mai…

View →
cs.CRRecentApr 26, 2026

Spore: Efficient and Training-Free Privacy Extraction Attack on LLMs via Inference-Time Hybrid Probing

Yu Cui, Ruiqing Yue, Hang Fu, Sicheng Pan +5 more

The paper introduces extsc{Spore}, a novel, training-free, and highly efficient privacy extraction attack that targets sensitive information stored in the memory of LLM agents during inference, outpe…

View →