To help researchers investigate relation extraction, we’re releasing a human-judged dataset of two relations about public figures on Wikipedia: nearly 10,000 examples of “place of birth”, and over 40,000 examples of “attended or graduated from an institution”. Each of these was judged by at least 5 raters, and can be used to train or evaluate relation extraction systems. We also plan to release more relations of new types in the coming months.
R. Baeza-Yates, and A. Tiberi. KDD '07: Proceedings of the 13th ACM SIGKDD international conference on Knowledge discovery and data mining, page 76--85. New York, NY, USA, ACM, (2007)
T. Tezuka, R. Lee, Y. Kambayashi, and H. Takakura. Proceedings of the Second International Conference on Web Information Systems Engineering, 2, page 14--21. (December 2001)
Y. Jin, Y. Matsuo, and M. Ishizuka. Proceedings of the European Semantic Web Conference, ESWC2007, volume 4519 of Lecture Notes in Computer Science, Springer-Verlag, (July 2007)