ConDA : state-based data augmentation for context-dependent text-to-SQL

Dingzirui Wang,Longxu Dou,Wanxiang Che, Jiaqi Wang,Jinbo Liu,Lixin Li, Jingan Shang,Lei Tao,Jie Zhang, Cong Fu, Xuri Song

International Journal of Machine Learning and Cybernetics(2024)

引用 0|浏览13
暂无评分
摘要
The context-dependent text-to-SQL task has profound real-world implications, as it facilitates users in extracting knowledge from vast databases, which allows users to acquire the information interactively for better accuracy. Unfortunately, current models struggle to address this task effectively due to the scarcity of data led by the high annotation overhead. The most straightforward method for addressing this problem is data augmentation, which aims at scaling up the parsing corpus. However, the naive methods suffer from the low diversity of the augmented data. To address this limitation, we propose the state-based CON text-dependent text-to-SQL D ata A ugmentation ( ConDA ), which generate and filter augmented data based on the dialogue state, which has higher diversity. Experimental results show that ConDA yields performance improvement on all experimental datasets with an average boosting of 1.6% , proving the effectiveness of our method.
更多
查看译文
关键词
Context-dependent text-to-SQL,Data augmentation,Semantic parsing,Natural language processing
AI 理解论文
溯源树
样例
生成溯源树,研究论文发展脉络
Chat Paper
正在生成论文摘要