Kernelized Hashcode Representations for Relation Extraction

descriptionPublicationkeyboard_double_arrow_right Article , Preprint , Conference object 17 Jul 2019Embargo end date: 01 Jan 2017Publisher:Association for the Advancement of Artificial Intelligence (AAAI)Journal:Proceedings of the AAAI Conference on Artificial Intelligence, volume 33, pages 6,431-6,440 (issn: 2159-5399, eissn: 2374-3468,

Copyright policy )

Authors: Sahil Garg; Aram Galstyan; Greg Ver Steeg; Irina Rish; Guillermo A. Cecchi; Shuyang Gao;

doi: 10.1609/aaai.v33i01.33016431 , 10.48550/arxiv.1711.04044

arXiv: 1711.04044

Kernelized Hashcode Representations for Relation Extraction

- Summary
- Subjects
- Related research
  (5)
- Metrics

Abstract

Kernel methods have produced state-of-the-art results for a number of NLP tasks such as relation extraction, but suffer from poor scalability due to the high cost of computing kernel similarities between natural language structures. A recently proposed technique, kernelized locality-sensitive hashing (KLSH), can significantly reduce the computational cost, but is only applicable to classifiers operating on kNN graphs. Here we propose to use random subspaces of KLSH codes for efficiently constructing an explicit representation of NLP structures suitable for general classification methods. Further, we propose an approach for optimizing the KLSH model for classification problems by maximizing an approximation of mutual information between the KLSH codes (feature vectors) and the class labels. We evaluate the proposed approach on biomedical relation extraction datasets, and observe significant and robust improvements in accuracy w.r.t. state-ofthe-art classifiers, along with drastic (orders-of-magnitude) speedup compared to conventional kernel methods.

Related Organizations

University of Southern California
United States
University of California System
United States
IBM Research – Thomas J. Watson Research Center
United States
Information Sciences Institute
United States
IBM (United States)
United States

Keywords

FOS: Computer and information sciences, Computer Science - Machine Learning, Computer Science - Computation and Language, Computation and Language (cs.CL), Information Retrieval (cs.IR), Computer Science - Information Retrieval, Machine Learning (cs.LG)

5 Research products, page 1 of 1

Learning with multiple kernels : algorithms and applications
2019IsAmongTopNSimilarDocuments
Kernelized Locality-Sensitive Hashing for Semi-Supervised Agglomerative Clustering
2013IsAmongTopNSimilarDocuments
Binary Embedding with Additive Homogeneous Kernels
2017IsAmongTopNSimilarDocuments
GPU-based kernelized locality-sensitive hashing for satellite image retrieval
2015IsAmongTopNSimilarDocuments
CorEx software on GitHub
IsRelatedTo

Impact byBIP!

	selected citations These citations are derived from selected sources. This is an alternative to the "Influence" indicator, which also reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically).	2
	popularity This indicator reflects the "current" impact/attention (the "hype") of an article in the research community at large, based on the underlying citation network.	Average
	influence This indicator reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically).	Average
	impulse This indicator reflects the initial momentum of an article directly after its publication, based on the underlying citation network.	Average