None
EN
Codex (and GPT-4) can’t beat humans on smart contract audits
['Artem Dinaburg', 'Josselin Feist', 'Riccardo Schirone']
The Trail of Bits Blog
Our multi-functional team, consisting of auditors, developers, and machine learning (ML) experts, put serious work into prompt engineering and developed a custom prompting framework that worked around some frustrations and limitations of current large language model (LLM) tooling, such as working with incorrect and inconsistent results, handling rate limits, and creating complex, templated chains of prompts. During Toucan’s development, we created a custom prompting framework, a web-based front end, and rudimentary debugging and testing tools to evaluate prompts and to aid in unit and integration tests.