An Efficient Framework for Constructing Generalized Locally-Induced Text Metrics

Abstract

In this paper, we propose a new framework for constructing text metrics which can be used to compare and support inferences among terms and sets of terms. Our metric is derived from data-driven kernels on graphs that let us capture global relations among terms and sets of terms, regardless of their complexity and size. To compute the metric efficiently for any two subsets of terms, we develop an approximation technique that relies on the precompiled term-term similarities. To scale-up the approach to problems with huge number of terms, we develop and experiment with a solution that subsamples the term space. We demonstrate the benefits of the whole framework on two text inference tasks: prediction of terms in the article from its abstract and query expansion in information retrieval.

Cite

Text

Amizadeh et al. "An Efficient Framework for Constructing Generalized Locally-Induced Text Metrics." International Joint Conference on Artificial Intelligence, 2011. doi:10.5591/978-1-57735-516-8/IJCAI11-198

Markdown

[Amizadeh et al. "An Efficient Framework for Constructing Generalized Locally-Induced Text Metrics." International Joint Conference on Artificial Intelligence, 2011.](https://mlanthology.org/ijcai/2011/amizadeh2011ijcai-efficient/) doi:10.5591/978-1-57735-516-8/IJCAI11-198

BibTeX

@inproceedings{amizadeh2011ijcai-efficient,
  title     = {{An Efficient Framework for Constructing Generalized Locally-Induced Text Metrics}},
  author    = {Amizadeh, Saeed and Wang, Shuguang and Hauskrecht, Milos},
  booktitle = {International Joint Conference on Artificial Intelligence},
  year      = {2011},
  pages     = {1159-1164},
  doi       = {10.5591/978-1-57735-516-8/IJCAI11-198},
  url       = {https://mlanthology.org/ijcai/2011/amizadeh2011ijcai-efficient/}
}