14 papers · ranked by Valyu relevance
Tsong Yueh Chen, Shing-Chi Cheung, Siu-Ming Yiu
In software testing, a set of test cases is constructed according to some predefined selection criteria. The software is then examined against these test cases. Three interesting observations have been made on the current artifacts of software testing. Firstly, an error-revealing test case is considered useful while a…
Shou Li, Andrey Kan, Laurent Callot, Bharat Bhasker + 2 more
'Muhammad H. Rashid' 'Timothy B Esler'] As LLM-based agents exhibit exceptional capabilities in addressing complex problems, there is a growing focus on developing coding agents to tackle increasingly sophisticated tasks. Despite their promising performance, these coding agents often produce programs or modifications…
Colin S. Gordon, Chaewon Yun
We describe a new concrete approach to giving predictable error locations for sequential (flow-sensitive) effect systems. Prior implementations of sequential effect systems rely on either computing a bottom-up effect and comparing it to a declaration (e.g., method annotation) or leaning on constraint-based type…
Hridya Dhulipala, Xiaokai Rong, Tien N. Nguyen
In several software development scenarios, it is desirable to detect runtime errors and exceptions in code snippets without actual execution. A typical example is to detect runtime exceptions in online code snippets before integrating them into a codebase. In this paper, we propose Cerberus, a novel predictive…
Mohammed Bekkouche
A model checker can produce a trace of counterexample, for an erroneous program, which is often long and difficult to understand. In general, the part about the loops is the largest among the instructions in this trace. This makes the location of errors in loops critical, to analyze errors in the overall program. In…
Yuzhuo Bai, Shuzheng Si, Kangyang Luo, Qingyi Wang + 4 more
Large language models (LLMs) often hallucinate, yet most existing fact-checking methods treat factuality evaluation as a binary classification problem, offering limited interpretability and failing to capture fine-grained error types. In this paper, we introduce InFi-Check, a framework for interpretable and…
Shuai Fu, Tim Dwyer, Peter J. Stuckey, John Grundy
Statically typed languages offer significant advantages, such as bug prevention, enhanced code quality, and reduced maintenance costs. However, these benefits often come at the expense of a steep learning curve and a slower development pace. Haskell, known for its expressive and strict type system, poses challenges for…
Yanggyu Lee, Suchae Jeong, Jihie Kim
Relationship into Prompts Authors: ['Yanggyu Lee' 'Suchae Jeong' 'Jihie Kim'] Abstract. LLMs trained in the understanding of programming syntax are now providing effective assistance to developers and are being used in programming education such as in generation of coding problem examples or providing code…
Nima Shiri Harzevili, Mohammad Mahdi Mohajer, Jiho Shin, Moshi Wei + 7 more
'Gias Uddin' 'Jinqiu Yang' 'Junjie Wang' 'Song Wang' 'Zhen Ming' 'Jiang' 'Nachiappan Nagappan'] Abstract—Checker bugs in Deep Learning (DL) libraries are critical yet not well-explored. These bugs are often concealed in the input validation and error-checking code of DL libraries and can lead to silent failures…
В. Г. Данилов, Ilya S. Turuntaev
In this article we address the problem of automatic answer checking in interactive learning systems that support mathematical notation. This problem consists of the problem of establishing identities in formal mathematical systems and hence is formally unsolvable. However, there is a way to cope with the issue. We…
Ben Greenman, Alan Jeffrey, Shriram Krishnamurthi, Mitesh Shah
Context Roblox Studio lets millions of creators build interactive experiences by programming in a variant of Lua called Luau. The creators form a broad group, ranging from novices writing their first script to professional developers; thus, Luau must support a wide audience. As part of its efforts to support all kinds…
Safeeullah Soomro, Zahid Hussain, Ayaz Keerio
Precisely and automatically detection of faults in programs, is a software engineering dream. Every effort in this regard takes us one step closer to realizing it. Many efforts have been taken from the people of these areas on testing, verification and debugging. We are proposing such effort for the research community…
Scott Ballentine, Eitan Farchi
| 1 | Introduction and background | | 1 | | | --- | --- | --- | --- | --- | | 2 | Review mindset and pitfalls | | 4 | | | 3 | | Review preparation | 6 | | | 4 | | Review techniques | 8 | | | | 4.1 | Fagan reviews | 8 | | | | 4.2 | Paraphrasing | 11 | | | | 4.3 | Obligations and contracts | 12 | | | | 4.4 | Bug patterns…
Benjamin Loriot, Fernanda Madeiral, Martin Monperrus
—Ensuring code formatting conventions is an essential aspect of modern software quality assurance, because it helps in code readability. In this paper, we present STYLER, a tool dedicated to fix formatting errors raised by Checkstyle, a highly configurable format checker for Java. To fix formatting errors in a given…