Abstract:Symbolic execution is an effective but expensive technique for automated test generation. Over the years, a large number of refined symbolic execution techniques have been proposed to improve its efficiency. However, the symbolic execution efficiency problem remains, and largely limits the application of symbolic execution in practice. Orthogonal to refined symbolic execution, in this paper we propose to accelerate symbolic execution through semantic-preserving code transformation on the target programs. During the initial stage of this direction, we adopt a particular code transformation, compiler optimization, which is initially proposed to accelerate program concrete execution by transforming the source program into another semantic-preserving target program with increased efficiency (e.g., faster or smaller). However, compiler optimizations are mostly designed to accelerate program concrete execution rather than symbolic execution. Recent work also reported that unified settings on compiler optimizations that can accelerate symbolic execution for any program do not exist at all. Therefore, in this work we propose a machine-learning based approach to tuning compiler optimizations to accelerate symbolic execution, whose results may also aid further design of specific code transformations for symbolic execution. In particular, the proposed approach LEO separates source-code functions and libraries through our program-splitter, and predicts individual compiler optimization (i.e., whether a type of code transformation is chosen) separately through analyzing the performance of existing symbolic execution. Finally, LEO applies symbolic execution on the code transformed by compiler optimization (through our local-optimizer). We conduct an empirical study on GNU Coreutils programs using the KLEE symbolic execution engine. The results show that LEO significantly accelerates symbolic execution, outperforming the default KLEE configurations (i.e., turning on/off all compiler optimizations) in various settings, e.g., with the default training/testing time, LEO achieves the highest line coverage in 50/68 programs, and its average improvement rate on all programs is 46.48%/88.92% in terms of line coverage compared with turning on/off all compiler optimizations.

Boosting Symbolic Execution Via Constraint Solving Time Prediction (experience Paper)

Speculative Symbolic Execution

Symbolic Execution of Complex Program Driven by Machine Learning Based Constraint Solving

SCSE: Boosting Symbolic Execution Via State Concretization

Neuro-Symbolic Execution: The Feasibility of an Inductive Approach to Symbolic Execution.

Towards Optimal Concolic Testing

Machine Learning Steered Symbolic Execution Framework for Complex Software Code

Symbolic execution of floating-point programs: How far are we?

Optimal Refinement-based Array Constraint Solving for Symbolic Execution

Python Symbolic Execution with LLM-powered Code Generation

Dependence Guided Symbolic Execution.

Learning to Accelerate Symbolic Execution via Code Transformation.

Divide, Conquer and Verify: Improving Symbolic Execution Performance

Symbolic Execution with Test Cases Generated by Large Language Models

Steering Symbolic Execution to Less Traveled Paths

Neuro-Symbolic Execution: Augmenting Symbolic Execution with Neural Constraints

On Benchmarking the Capability of Symbolic Execution Tools with Logic Bombs

Natural Symbolic Execution-Based Testing for Big Data Analytics

A Static Analysis Tool with Optimizations for Reachability Determination

An Empirical Study on Constraint Optimization Techniques for Test Generation

Chronosymbolic Learning: Efficient CHC Solving with Symbolic Reasoning and Inductive Learning