2007 · 329 citations · 6 references
Natural Language ProcessingEngineeringInformation RetrievalKnowledge ExtractionData ScienceComputational LinguisticsKnowledge DiscoveryKeyword ExtractionTextrunner SystemSemantic WebData ExtractionInformation ExtractionNamed-entity RecognitionCorpus LinguisticsOpen Information ExtractionText MiningMachine Translation
Traditional information extraction systems have focused on satisfying precise, narrow, pre-specified requests from small, homogeneous corpora. In contrast, the TextRunner system demonstrates a new kind of information extraction, called Open Information Extraction (OIE), in which the system makes a single, data-driven pass over the entire corpus and extracts a large set of relational tuples, without requiring any human input. (Banko et al., 2007) TextRunner is a fully-implemented, highly scalable example of OIE. TextRunner's extractions are indexed, allowing a fast query mechanism.
6
Open information extraction from the web
Michele Banko, Michael Cafarella, Stephen Soderland et al. · 2007 · 1.3K citations
Open information extraction from the web
Oren Etzioni, Michele Banko, Stephen Soderland et al. · Communications of the ACM · 2008 · 1K citations