Results 21 to 30 of about 17,789,050 (310)
Extracting Proceedings Data from Court Cases with Machine Learning
France is rolling out an open data program for all court cases, but with few metadata attached. Reusers will have to use named-entity recognition (NER) within the text body of the case to extract any value from it.
Bruno Mathis
doaj +1 more source
An Information Extraction Customizer [PDF]
When an information extraction system is applied to a new task or domain, we must specify the classes of entities and relations to be extracted. This is best done by a subject matter expert, who may have little training in NLP. To meet this need, we have developed a toolset which is able to analyze a corpus and aid the user in building the ...
Ralph Grishman, Yifan He 0006
openaire +2 more sources
Information extraction from chemical patents [PDF]
The automated extraction of semantic chemical data from the existing literature is demonstrated. For reasons of copyright, the work is focused on the patent literature, though the methods are expected to apply equally to other areas of the chemical ...
core +4 more sources
Progress in information extraction [PDF]
This paper provides a quick summary of the following topics: enhancements to the PLUM information extraction engine, what we learned from MUC-6 (the Sixth Message Understanding Conference), the results of an experiment on merging templates from two different information extraction engines, a learning technique for named entity recognition, and towards ...
Ralph M. Weischedel +6 more
openaire +1 more source
Using Web-Based Knowledge Extraction Techniques to Support Cultural Modeling [PDF]
The World Wide Web is a potentially valuable source of information about the cognitive characteristics of cultural groups. However, attempts to use the Web in the context of cultural modeling activities are hampered by the large-scale nature of the Web ...
Sieck, Winston R +2 more
core +2 more sources
The complexity of information extraction [PDF]
How difficult are decision problems based on natural data, such as pattern recognition? To answer this question, decision problems are characterized by introducing four measures defined on a Boolean function f of N variables: the implementation cost C(f), the randomness R(f), the deterministic entropy H(f), and the complexity K(f).
openaire +2 more sources
Adaptive information extraction [PDF]
The growing availability of online textual sources and the potential number of applications of knowledge acquisition from textual data has lead to an increase in Information Extraction (IE) research. Some examples of these applications are the generation of data bases from documents, as well as the acquisition of knowledge useful for emerging ...
Jordi Turmo, Alicia Ageno, Neus Català
openaire +1 more source
Information extraction and evaluation [PDF]
This topic session focussed on a variety of issues in evaluation (three presentations) and extraction (one presentation). For the extraction presentation, Tsuyoshi Kitani, a visiting researcher at the Center for Machine Translation at Carnegie-Mellon University gave a presention entitled "Overview of TEXTRACT Template-Filling Solutions". This talk gave
openaire +1 more source
Social media is now regarded as the most valuable source of data for trend analysis and innovative business process reengineering preferences. Data made accessible through social media can be utilized for a variety of purposes, such as by an entrepreneur
Heru Susanto, Aida Sari, Fang-Yie Leu
doaj +1 more source
Portable extraction of partially structured facts from the web [PDF]
A novel fact extraction task is defined to fill a gap between current information retrieval and information extraction technologies. It is shown that it is possible to extract useful partially structured facts about different kinds of entities in a broad
Inguna Skadiņa +8 more
core +2 more sources

