2023 · 52 citations · 42 references
Few-shot LearningLlm Fine-tuningEngineeringFew-shot In-context LearningSemantic WebLarge Language ModelCorpus LinguisticsNatural Language ProcessingZero-shot LearningInformation RetrievalData ScienceKnowledge BasesComputational LinguisticsLanguage StudiesMachine TranslationQuestion AnsweringNatural Language InterfaceDiverse Kbqa DatasetsComputer ScienceKnowledge BaseRetrieval Augmented GenerationLinguistics
Question answering over knowledge bases is considered a difficult problem due to the challenge of generalizing to a wide variety of possible natural language questions. Additionally, the heterogeneity of knowledge base schema items between different knowledge bases often necessitates specialized training for different knowledge base question-answering (KBQA) datasets. To handle questions over diverse KBQA datasets with a unified training-free framework, we propose KB-BINDER, which for the first time enables few-shot in-context learning over KBQA tasks. Firstly, KB-BINDER leverages large language models like Codex to generate logical forms as the draft for a specific question by imitating a few demonstrations. Secondly, KB-BINDER grounds on the knowledge base to bind the generated draft to an executable one with BM25 score matching. The experimental results on four public heterogeneous KBQA datasets show that KB-BINDER can achieve a strong performance with only a few in-context demonstrations. Especially on GraphQA and 3-hop MetaQA, KB-BINDER can even outperform the state-of-the-art trained models. On GrailQA and WebQSP, our model is also on par with other fully-trained models. We believe KB-BINDER can serve as an important baseline for future research. We plan to release all the code and data. Our code is available at https://github.com/ltl3A87/KB-BINDER.
42
Kurt Bollacker, Colin Evans, Praveen Paritosh et al. · 2008 · 4.9K citations
Metaweb Query Language, Engineering, Information Retrieval +15
Semantic Parsing on Freebase from Question-Answer Pairs
Jonathan Berant, Andrew Chou, Roy Frostig et al. · 2013 · 1.6K citations · Full text
Evaluating Large Language Models Trained on Code
Mark Chen, Jerry Tworek, Heewoo Jun et al. · arXiv (Cornell University) · 2021 · 1.4K citations · Full text