Automatic content extraction
From Wikipedia, the free encyclopedia
Page Module:Message box/ambox.css has no content.Page Template:Multiple issues/styles.css has no content.
This article has multiple issues. Please help improve it or discuss these issues on the talk page. (Learn how and when to remove these messages)
Page Module:Message box/ambox.css has no content.
|
Automatic content extraction (ACE) is a research program for developing advanced information extraction technologies convened by the NIST from 1999 to 2008, succeeding MUC and preceding Text Analysis Conference.
Topics and exercises
Given a text in natural language, the ACE challenge is to detect:
- entities mentioned in the text, such as: persons, organizations, locations, facilities, weapons, vehicles, and geo-political entities.
- relations between entities, such as: person A is the manager of company B.
- events mentioned in the text, such as: interaction, movement, transfer, creation and destruction.
The program relates to English, Arabic and Chinese texts.
The ACE corpus is one of the standard benchmarks for testing new information extraction algorithms.
References
- George Doddington@NIS T, Alexis Mitchell@LD C, Mark Przybocki@NIS T, Lance Ramshaw@BB N, Stephanie Strassel@LD C, Ralph Weischedel@BB N. The automatic content extraction (ACE) program–tasks, data, and evaluation. 2004