Skip to yearly menu bar Skip to main content


AgentScript-Eval: Benchmarking LLMs and Agents for Code Generation in an Enterprise DSL

Dipin Khati ⋅ Shubham Mehrotra ⋅ Bin Bi ⋅ Zhou Yu ⋅ Denys Poshyvanyk ⋅ James Zhu ⋅ Sitaram Asur ⋅ Phil Mui

Abstract

Chat is not available.