Ranking LLM-Generated Loop Invariants for Program Verification

October 13, 2023 · Declared Dead · 🏛 Conference on Empirical Methods in Natural Language Processing

Authors Saikat Chakraborty, Shuvendu K. Lahiri, Sarah Fakhoury, Madanlal Musuvathi, Akash Lal, Aseem Rastogi, Aditya Senthilnathan, Rahul Sharma, Nikhil Swamy arXiv ID 2310.09342 Category cs.PL: Programming Languages Cross-listed cs.AI, cs.CL, cs.SE Citations 57 Venue Conference on Empirical Methods in Natural Language Processing Repository https://github.com/microsoft/NeuralInvariantRanker} Last Checked 1 month ago

Abstract

Synthesizing inductive loop invariants is fundamental to automating program verification. In this work, we observe that Large Language Models (such as gpt-3.5 or gpt-4) are capable of synthesizing loop invariants for a class of programs in a 0-shot setting, yet require several samples to generate the correct invariants. This can lead to a large number of calls to a program verifier to establish an invariant. To address this issue, we propose a {\it re-ranking} approach for the generated results of LLMs. We have designed a ranker that can distinguish between correct inductive invariants and incorrect attempts based on the problem definition. The ranker is optimized as a contrastive ranker. Experimental results demonstrate that this re-ranking mechanism significantly improves the ranking of correct invariants among the generated candidates, leading to a notable reduction in the number of calls to a verifier. The source code and the experimental data for this paper are available in \url{https://github.com/microsoft/NeuralInvariantRanker}.

📄 View on arXiv 🌐 View on ar5iv 📑 PDF 💻 Repository 🎉 Report Code Found

Community Contributions

Found the code? Know the venue? Think something is wrong? Let us know!

📜 Similar Papers

In the same crypt — Programming Languages

R.I.P. 👻 Ghosted

Ascertaining Uncertainty for Efficient Exact Cache Analysis

Valentin Touzeau, Claire Maïza, ... (+2 more)

cs.PL 🏛 CAV 📚 816 cites 8 years ago

R.I.P. 👻 Ghosted

Tensor Comprehensions: Framework-Agnostic High-Performance Machine Learning Abstractions

Nicolas Vasilache, Oleksandr Zinenko, ... (+7 more)

cs.PL 🏛 arXiv 📚 472 cites 8 years ago

R.I.P. 👻 Ghosted

Glow: Graph Lowering Compiler Techniques for Neural Networks

Nadav Rotem, Jordan Fix, ... (+16 more)

cs.PL 🏛 arXiv 📚 318 cites 7 years ago

R.I.P. 👻 Ghosted

Learnable Programming: Blocks and Beyond

David Bau, Jeff Gray, ... (+3 more)

cs.PL 🏛 CACM 📚 298 cites 8 years ago

R.I.P. 👻 Ghosted

Scenic: A Language for Scenario Specification and Scene Generation

Daniel J. Fremont, Tommaso Dreossi, ... (+4 more)

cs.PL 🏛 ACM-SIGPLAN Symposium on Programming Language Design and Implementation 📚 297 cites 7 years ago

R.I.P. 👻 Ghosted

Vandal: A Scalable Security Analysis Framework for Smart Contracts

Lexi Brent, Anton Jurisevic, ... (+6 more)

cs.PL 🏛 arXiv 📚 296 cites 7 years ago

Died the same way — 💀 404 Not Found

R.I.P. 💀 404 Not Found

Deep High-Resolution Representation Learning for Visual Recognition

Jingdong Wang, Ke Sun, ... (+10 more)

cs.CV 🏛 IEEE TPAMI 📚 4.4K cites 6 years ago

R.I.P. 💀 404 Not Found

HuggingFace's Transformers: State-of-the-art Natural Language Processing

Thomas Wolf, Lysandre Debut, ... (+20 more)

cs.CL 🏛 arXiv 📚 3.5K cites 6 years ago

R.I.P. 💀 404 Not Found

CCNet: Criss-Cross Attention for Semantic Segmentation

Zilong Huang, Xinggang Wang, ... (+5 more)

cs.CV 🏛 ICCV 📚 2.9K cites 7 years ago

R.I.P. 💀 404 Not Found

Unified Perceptual Parsing for Scene Understanding

Tete Xiao, Yingcheng Liu, ... (+3 more)

cs.CV 🏛 ECCV 📚 2.3K cites 7 years ago