2606.19319v1 Jun 17, 2026 cs.MA

데이터 지능 에이전트: 자율 코딩 에이전트를 활용한 기업 데이터의 해석, 모델링 및 질의

Data Intelligence Agents: Interpreting, Modeling, and Querying Enterprise Data via Autonomous Coding Agents

S. Pakazad
S. Pakazad
Citations: 33
h-index: 2
Henrik Ohlsson
Henrik Ohlsson
Citations: 1
h-index: 1
Anoushka Vyas
Anoushka Vyas
Citations: 26
h-index: 2
Aarushi Dhanuka
Aarushi Dhanuka
Citations: 36
h-index: 2

생산 데이터 통합은 데이터 소유자, 엔지니어 및 분석가 간의 반복적인 정보 전달 과정에서 발생하는 손실로 인해 어려움을 겪습니다. 이들은 협력하여 기업 데이터를 발견하고 구조화하며 질의를 수행해야 합니다. 본 논문에서는 데이터 지능 에이전트(DIA)라는 시스템을 제시합니다. DIA는 세 가지 에이전트(데이터 해석기, 스키마 생성기, 질의 생성기)로 구성되어 있으며, 자율 코딩 에이전트(ACA)를 핵심 요소로 활용하여 이러한 워크플로우를 간소화합니다. 에이전트는 텍스트 대신 구체적인 결과물을 생성, 실행, 검증 및 수정하며, 공유 메모리를 통해 경험을 재사용하고, 각 단계를 도메인 전문가의 검토를 위해 제공합니다. DIA는 실제 기업 고객 환경에 배포되어 사용되고 있습니다. 본 연구에서는 질의 생성기를 심층적으로 분석하고, 네 가지 작업 범주와 네 가지 방언을 포괄하는 일곱 개의 SQL 벤치마크에서 완전 자율 모드로 평가했습니다. 그 결과, 질의 생성기는 모든 벤치마크에서 최고 수준의 성능을 달성하거나 능가했으며, 이는 실행 기반의 아키텍처가 ACA와 공유 메모리를 통해 데이터 지능 작업 전반에 걸쳐 일반화될 수 있으며, 자연어 지침만으로 적응이 가능하다는 것을 보여줍니다.

Original Abstract

Production data integration is bottlenecked by repeated, lossy handoffs between data owners, engineers, and analysts who must collaboratively discover, structure, and query enterprise data. We present Data Intelligence Agents (DIA), a system of three agents (Data Interpreter, Schema Creator, and Query Generator) that compresses this workflow by treating autonomous coding agents (ACAs) as a first-class abstraction: rather than emitting text, the agents generate, execute, validate, and repair concrete artifacts, draw on a shared memory for experience reuse, and surface each for review by domain experts. DIA is deployed in production for enterprise customers. We study the Query Generator in depth and evaluate it in fully autonomous mode across seven SQL benchmarks spanning four task categories and four dialects. It matches or surpasses the best published results on all seven, demonstrating that an architecture grounded in execution, built on ACAs and a shared memory, generalizes across the data intelligence workload with adaptation confined to natural-language instructions.

0 Citations
0 Influential
1 Altmetric
5.0 Score
Original PDF

No Analysis Report Yet

This paper hasn't been analyzed by Gemini yet.

Log in to request an AI analysis.

댓글

댓글을 작성하려면 로그인하세요.

아직 댓글이 없습니다. 첫 번째 댓글을 남겨보세요!