Neil Ireson, Fabio Ciravegna, Marie Elaine Califf, Dayne Freitag, Nicholas Kushmerick, Alberto Lavelli: Evaluating Machine Learning for Information Extraction, 22nd International Conference on Machine Learning (ICML 2005), Bonn, Germany, 7-11 August, 2005
This is the project page for SecondString, an open-source Java-based package of approximate string-matching techniques. This code was developed by researchers at Carnegie Mellon University from the Center for Automated Learning and Discovery, the Department of Statistics, and the Center for Computer and Communications Security.
SecondString is intended primarily for researchers in information integration and other scientists. It does or will include a range of string-matching methods from a variety of communities, including statistics, artificial intelligence, information retrieval, and databases. It also includes tools for systematically evaluating performance on test data. It is not designed for use on very large data sets.
The main task of the GenIELex project is the development of a biochemistry specific lexicon as well as of an annotated corpus for the evaluation of the system. The need for the construction of such a lexicon is illustrated by the following figures, based
P. Kluegl, M. Atzmueller, und F. Puppe. Proc. 4th International Workshop on Knowledge Engineering and Software Engineering (KESE 2008), 31th German Conference on Artificial Intelligence (KI-2008), accepted, (2008)
P. Kluegl, M. Atzmueller, und F. Puppe. Proceedings of the Biennial GSCL Conference 2009, 2nd UIMA@GSCL Workshop, Seite 233-240. Gunter Narr Verlag, (2009)
H. Chieu, und H. Ng. Eighteenth national conference on Artificial intelligence, Seite 786--791. Menlo Park, CA, USA, American Association for Artificial Intelligence, (2002)
S. Huffman. Connectionist, Statistical, And Symbol Approaches to Learning for
Natural Language Processing, volume 1040, Seite 246-260. Springer, (1996)
M. Banko, M. Cafarella, S. Soderland, M. Broadhead, und O. Etzioni. Proceedings of the 20th International Joint Conference on Artifical Intelligence, Seite 2670--2676. San Francisco, CA, USA, Morgan Kaufmann Publishers Inc., (2007)
D. Maynard, Y. Li, und W. Peters. Proceedings of the 2008 conference on Ontology Learning and Population: Bridging the Gap between Text and Knowledge, Seite 107-127. IOS Press, (2008)
R. Gupta, und S. Sarawagi. Proceedings of the fourth ACM international conference on Web search and data mining, Seite 217--226. New York, NY, USA, ACM, (2011)
P. Kluegl, M. Toepfer, F. Lemmerich, A. Hotho, und F. Puppe. Proceedings of 1st International Conference on Pattern Recognition Applications and Methods (ICPRAM), Seite 240-248. Vilamoura, Algarve, Portugal, SciTePress, (6-8 02 2012)
G. Sautter, und K. Böhm. Proceedings of the Second International Conference on Theory and Practice of Digital Libraries, Seite 370--382. Berlin/Heidelberg, Springer, (2012)
M. Califf, und R. Mooney. Proceedings of the sixteenth national conference on Artificial intelligence and the eleventh Innovative applications of artificial intelligence conference innovative applications of artificial intelligence, Seite 328--334. Menlo Park, CA, USA, American Association for Artificial Intelligence, (1999)
P. Kluegl, M. Atzmueller, und F. Puppe. Proc. 4th International Workshop on Knowledge Engineering and Software Engineering (KESE 2008), 31th German Conference on Artificial Intelligence (KI-2008), (2008)