~ similar to 2607.24387· 20 results
This paper evaluates Centaur, a foundation model trained on psychological experiments, for program comprehension tasks and compares its performance to Llama 3.1.
The paper investigates how AI coding assistants shift developers' security focus from proactive prevention to reactive review, finding that this structural change is reinforced by current tool interac…
Jia Liu, Veena Krishnaraj, Kateryna Vovk, Kosuke Aizawa +17 more
This paper investigates the ability of large language models to assist in scientific project planning and proposal evaluation, finding that human reviewers rate human- and AI-written proposals similar…
This paper synthesizes practitioner discourse at scale to build a causal model explaining the impact of AI on code review, recovering mechanisms behind observed trends.
This paper introduces Aleena, an open-source lifecycle alignment agent that uses GitHub to align research software engineering stakeholders and preserve decision continuity.
This paper provides the first large-scale characterisation of Domain-Driven Design (DDD) adoption and implementation on GitHub.
This paper analyzes 25,264 agentic pull requests from 2,361 GitHub repositories to investigate adoption, productivity, and collaboration patterns of agentic coding tools.
Seth Bernstein, Paul Denny, Juho Leinonen, Kush Patel +3 more
This paper explores the effectiveness of diverse LLM-generated explanations versus generic explanations in computer science education, finding that diverse explanations led to higher open-ended respon…
The study analyzes coding patterns in malware versus benign software, finding that malware code is optimized for quick evasion and secrecy rather than maintainability, though its metrics are not uniqu…
The paper introduces Clover, a code completion tool that logs students' interactions and offers attention checks to promote reflective engagement during programming tasks.
This paper compares four requirements elicitation approaches using AI-supported collaboration and evaluates their impact on requirements artifact quality and stakeholder perceptions.
This paper analyzes the shift in student adoption and accountability practices of GenAI tools in a graduate-level HCI prototyping course across three cohorts using GitHub data.
This paper conducts rapid reviews of peer-reviewed publications to assess the availability of evidence on software engineering practices in government agencies.
This paper investigates how developer attributes influence generated software using AI, finding significant differences in interface design, template content, and code structure based on age and gende…
Fariha Tanjim Shifat, Hariswar Baburaj, Ce Zhou, Jaydeb Sarker +1 more
The paper analyzes GitHub security advisories for LLM-integrated open-source systems, finding that while most vulnerabilities map to existing code-level weaknesses, the architectural risks like Supply…
This study provides an ecosystem-scale measurement of commit signing on GitHub, finding that current signing adoption rates are misleading and that developers struggle to maintain consistent, long-ter…
This paper analyzes online developer discussions to identify four major security concerns—data leakage, code licensing, adversarial attacks, and insecure suggestions—associated with using generative A…