ArXivCSExplorer
☆☆Bookmarks🏆RSSHow to UseFAQ
Built with and by Teycir Ben Soltane•
How to Use•FAQ•GitHub•arXiv.org•
Share:

20 results for “AI summaries”

CS papers only

Hybrid search: Keyword + semantic, ranked by combined score.ⓘ

Want pure semantic search? Try claim verification →

cs.IRcs.HCEmpiricalRecentJul 3, 2026

AI Overviews in Academic Search: Evaluating AI-generated Summaries of Search Results in a Domain-specific Search Engine

Kevin Schott, Kanishka Silva, Ingo Frommholz, Philipp Mayr +2 more

This paper evaluates the use of AI-generated summaries on search engine results pages (SERPs) in academic search for social science information.

View →
cs.CLcs.AIcs.MAEmpiricalRecentJul 16, 2026

Dialogue Summarization with Emotion Dynamics Using Topic- and Participant-Centric Decomposition

Linyun Xiang, Mark Neerincx, Stephanie Tan

This paper proposes a framework for summarizing dialogues, modeling semantic and emotion dynamics using multimodal inputs and an adapted hierarchical Chain-of-Agents approach.

View →
cs.AIRecentMay 27, 2026

AIBuildAI-2: A Knowledge-Enhanced Agent for Automatically Building AI Models

Ruiyi Zhang, Peijia Qin, Qi Cao, Li Zhang +1 more

The paper introduces AIBuildAI-2, a knowledge-enhanced agent that significantly improves the automatic building of AI models by integrating an external, evolving knowledge system, achieving state-of-t…

View →
cs.AIcs.MAEmpiricalRecentJul 9, 2026

ASMR: Agentic Schema Generation for Ship Maintenance Report Writing

Sohrab Namazi Nia, Amogh Dalal, Ning Sa, Peter Ly +5 more

This paper proposes ASMR, a framework for automatically generating compact and informative schemas from historical ship maintenance reports using two specialized agents.

View →
cs.CLeess.ASEmpiricalRecentJul 19, 2026

Robust Summarization of Doctor-Patient Conversations: TalTech Systems for the Beyond Transcription Challenge

Aivo Olev, Tanel Alumäe

TalTech submitted top-ranking systems to the Beyond Transcription Challenge using fine-tuned Voxtral models and reinforcement learning against Open Medical Concept F1.

View →
cs.AIcs.DBRecentMay 27, 2026

A Query Engine for the Agents

Kenny Daniel

The paper introduces Hyperparam, a set of lightweight JavaScript libraries designed to enable direct, model-aware querying of unstructured data (like agent traces) within client-side AI applications.

View →
cs.CLcs.AIRecentMay 31, 2026

Understanding LLM Behavior in Multi-Target Cross-Lingual Summarization

Sangwon Ryu, Yihong Liu, Mingyang Wang, Yunsu Kim +3 more

The paper introduces a new benchmark for multi-target cross-lingual summarization (MTXLS) and proposes an activation steering method that significantly improves LLM performance by guiding the generati…

View →
cs.IREmpiricalRecentJul 24, 2026

The Prompt Is Not the Query: How Request State Evolves Across Multi-Turn AI Conversations

Benjamin Tannenbaum

This paper investigates how the final prompt in conversational AI-search evaluations differs from the conversation history, using two corpora of commercial and PRISM conversations.

View →
cs.CLcs.AIcs.HCEmpiricalRecentJun 15, 2026

PromptMN: Pseudo Prompting Language

Enkhzol Dovdon

This paper introduces PromptMN, a domain-specific language for annotating natural language prompts to clarify roles, goals, and constraints for AI models, reducing context ambiguities and repair cycle…

View →
cs.AIEmpiricalRecentJun 21, 2026

PaperClaw: Harnessing Agents for Autonomous Research and Human-in-the-Loop Refinement

Weiwei Ye, Hangchen Liu, Dongyuan Li, Renhe Jiang

PAPERCLAW is a multi-agent system that autonomously curates a domain, generates ideas, and writes venue-compliant papers using large language models and a stoppable hypothesis map.

View →
cs.CLRecentMay 30, 2026

I-WebGenBench : Evaluating Interactivity in LLM-Generated Scientific Web Applications

Dasen Dai, Biao Wu, Meng Fang, Shuoqi Li +1 more

The paper introduces I-WebGenBench, a framework and benchmark that converts static scientific papers into executable, interactive web systems, allowing users to dynamically explore the paper's mechani…

View →
cs.HCEmpiricalRecentJul 21, 2026

Evaluating a Visual Query Tracer and Builder for Learning Declarative Logic Programming

Julián Méndez, Lukas Gerlach, Tobias Wieland, Alex Ivliev +2 more

The authors conducted a user study to assess the effectiveness of their interactive visual query tracer and builder tools for Nemo, a Datalog reasoner, in helping students learn Datalog.

View →
cs.IREmpiricalRecentJul 1, 2026

When RAG Meets Query Planning: Logical Query Trees for Resolving Exploratory Reasoning Problems

Ganlin Xu, Linghao Zhang, Zhitao Yin, Hongda Xi +6 more

The paper introduces PlanRAG, a framework for Retrieval-Augmented Generation (RAG) that models exploratory reasoning problems as logical query trees, addressing representation and optimization gaps be…

View →
cs.CLcs.AIcs.CERecentMay 28, 2026

MOOSE-Copilot: A Web-Based Interactive Assistant for Unified Exploratory and Fine-Grained Scientific Hypothesis Discovery

Hongran An, Zonglin Yang

MOOSE-Copilot is a novel web-based framework that unifies scientific hypothesis discovery by formalizing human-AI interaction, significantly improving performance over autonomous LLM baselines.

View →
cs.CLRecentJun 1, 2026

Towards Multidisciplinary Summarization of Hospital Stays: Efficient Sentence-Level Clinical Provenance Categorization

Baris Karacan, Vaibhav Bhargava, Barbara Di Eugenio, Natalie Parde +20 more

The paper introduces a supervised fine-tuning pipeline using large language models to accurately categorize sentence-level clinical provenance across multi-disciplinary hospital notes, demonstrating t…

View →
cs.CLRecentJun 1, 2026

TVIR: Building Deep Research Agents Towards Text--Visual Interleaved Report Generation

Xinkai Ma, Zhiqi Bai, Dingling Zhang, Pei Liu +20 more

The paper introduces TVIR, a new benchmark and multi-agent framework for deep research, to evaluate and improve the generation of factually reliable, text-visual interleaved reports.

View →
cs.CLcs.AIcs.CYRecentMay 29, 2026

If LLMs Have Human-Like Attributes, Then So Does Age of Empires II

Adrian de Wynter

The paper argues that purported anthropomorphic attributes of LLMs are not unique to language models but are substrate-dependent, demonstrating this by training a neural network on the game Age of Emp…

View →
cs.SEPositionRecentJul 24, 2026

Code Review is a Conversation: Toward Conversational AI Review Assistants

Rosalia Tufano

This paper proposes conversational AI review assistants for code review, systems that engage in conversation with developers instead of just generating comments.

View →
cs.CLcs.AIcs.MAEmpiricalRecentJul 1, 2026

From Personas to Plot: Character-Grounded Multi-Agent Story Generation for Long-Form Narratives

Aayush Aluru, Chloe Ho, Muhammad Hammouri, Kerry Luo +4 more

This paper introduces MAGNET, a framework for long-form narrative generation and verification using a multi-agent goal-driven engine and a graph-based pipeline.

View →