cs.SE · 2026-06-29 · No. 38

Software Engineering, 2026-06-29.

3 new papers in cs.SE. Titles, authors, abstracts. Links to arXiv. Want this in your inbox every morning? Subscribe →

01 — The papers

3 entries
  1. 01

    Govern the Repository, Not the Agent: Measuring Ecosystem-Level Risk in AI-Native Software

    Daniel Russo

    cs.SE · cs.AI

    Autonomous coding agents now open and merge pull requests in shared repositories at scale, and the field evaluates them the way it has always evaluated components, one agent at a time, on isolated benchmark tasks. Yet agents that each pass their own tests still leave repositories that accumulate problems no single contribution accounts for. We ask whether this problem belongs to the individual agent or to the repository where it accumulates....

    arxiv.org/abs/2606.28235 · PDF

  2. 02

    Reasoning Beyond Prediction: From Data-Driven to Causal Software Engineering

    Roberto Pietrantuono, Luca Giamattei, Stefano Russo

    cs.SE · cs.AI

    Software engineering is an intellectually demanding, creative discipline that juggles a web of interdependent tasks to design, build, and assure the quality of increasingly complex systems. As our expectations from software soar - with demands spanning AI-driven products, pervasively distributed and cloud-native architectures, and deeply embedded cyber-physical environments - its complexity steadily increases. In response, a new wave of...

    arxiv.org/abs/2606.27960 · PDF

  3. 03

    Speculative Refinement: A Hybrid Autoregressive Diffusion Decoding Strategy and Its Behavior Across Benchmarks

    Aditi Gupta, Neel Mishra, Kushagra Trivedi, Pawan Kumar

    cs.SE · cs.AI

    How should we evaluate generation systems that combine autoregressive (AR) and diffusion decoding? We study this question through Speculative Refinement (SpecRef), a training-free hybrid method that warm-starts a masked diffusion language model from an AR draft using entropy-guided selective masking. Evaluating SpecRef across six benchmarks (HumanEval, MBPP, GSM8K, BBH, ARC-Challenge, HellaSwag) with three distinct evaluation protocols...

    arxiv.org/abs/2606.27474 · PDF

This edition is part of The Daily Abstract — cs.SE archive. Subscribe to receive these in your inbox each morning, automatically translated to Spanish, with reply-to-PDF: arxivdaily.ignorelist.com.

Colophon Set in Georgia, with system sans for interface chrome and a monospaced stack for code and paper identifiers. Sole accent: amber #D99C5E. Built and served on an always-free VM. The masthead is set 14% letterspaced because newspapers do that and it works.