arXiv CS AI
techCenter
DI-Bench: Systematically Generating In-Domain Data Intelligence Benchmarks for Enterprise Agentstranslating…
arXiv:2609.05776v1 Announce Type: new
Abstract: Evaluating enterprise agents on domain-specific benchmarks is critical, yet public benchmarks rarely evaluate whether agents can integrate business knowledge with analytical computation, and constructing such benchmarks manually…
Keywords#Benchmarks#Agents#Enterprise Agents#Data#Data Intelligence