Abstract:Symbolic execution is an effective but expensive technique for automated test generation. Over the years, a large number of refined symbolic execution techniques have been proposed to improve its efficiency. However, the symbolic execution efficiency problem remains, and largely limits the application of symbolic execution in practice. Orthogonal to refined symbolic execution, in this paper we propose to accelerate symbolic execution through semantic-preserving code transformation on the target programs. During the initial stage of this direction, we adopt a particular code transformation, compiler optimization, which is initially proposed to accelerate program concrete execution by transforming the source program into another semantic-preserving target program with increased efficiency (e.g., faster or smaller). However, compiler optimizations are mostly designed to accelerate program concrete execution rather than symbolic execution. Recent work also reported that unified settings on compiler optimizations that can accelerate symbolic execution for any program do not exist at all. Therefore, in this work we propose a machine-learning based approach to tuning compiler optimizations to accelerate symbolic execution, whose results may also aid further design of specific code transformations for symbolic execution. In particular, the proposed approach LEO separates source-code functions and libraries through our program-splitter, and predicts individual compiler optimization (i.e., whether a type of code transformation is chosen) separately through analyzing the performance of existing symbolic execution. Finally, LEO applies symbolic execution on the code transformed by compiler optimization (through our local-optimizer). We conduct an empirical study on GNU Coreutils programs using the KLEE symbolic execution engine. The results show that LEO significantly accelerates symbolic execution, outperforming the default KLEE configurations (i.e., turning on/off all compiler optimizations) in various settings, e.g., with the default training/testing time, LEO achieves the highest line coverage in 50/68 programs, and its average improvement rate on all programs is 46.48%/88.92% in terms of line coverage compared with turning on/off all compiler optimizations.

Symbolic Execution with Test Cases Generated by Large Language Models

Python Symbolic Execution with LLM-powered Code Generation

Steering Symbolic Execution to Less Traveled Paths

Learning to Accelerate Symbolic Execution via Code Transformation.

Symbolic Execution of Complex Program Driven by Machine Learning Based Constraint Solving

Machine Learning Steered Symbolic Execution Framework for Complex Software Code

Generating Data for Symbolic Language with Large Language Models

Can Large Language Models Understand Symbolic Graphics Programs?

Speculative Symbolic Execution

Speak It Out: Solving Symbol-Related Problems with Symbol-to-Language Conversion for Language Models

Symbol-LLM: Towards Foundational Symbol-centric Interface For Large Language Models

Large Language Models to Generate System-Level Test Programs Targeting Non-functional Properties

Investigating Symbolic Capabilities of Large Language Models

Divide, Conquer and Verify: Improving Symbolic Execution Performance

Large Language Models are Interpretable Learners

SCSE: Boosting Symbolic Execution Via State Concretization

Large Language Models Are Neurosymbolic Reasoners

Executing Natural Language-Described Algorithms with Large Language Models: An Investigation

Large Language Models as Code Executors: An Exploratory Study

Natural Symbolic Execution-Based Testing for Big Data Analytics

Automated Control Logic Test Case Generation using Large Language Models