A Survey on Metric Learning for Feature Vectors and Structured Data

Open AccessPosted Content

A Survey on Metric Learning for Feature Vectors and Structured Data

- 28 Jun 2013 -

TLDR

A systematic review of the metric learning literature is proposed, highlighting the pros and cons of each approach and presenting a wide range of methods that have recently emerged as powerful alternatives, including nonlinear metric learning, similarity learning and local metric learning.

Abstract:

The need for appropriate ways to measure the distance or similarity between data is ubiquitous in machine learning, pattern recognition and data mining, but handcrafting such good metrics for specific problems is generally difficult. This has led to the emergence of metric learning, which aims at automatically learning a metric from data and has attracted a lot of interest in machine learning and related fields for the past ten years. This survey paper proposes a systematic review of the metric learning literature, highlighting the pros and cons of each approach. We pay particular attention to Mahalanobis distance metric learning, a well-studied and successful framework, but additionally present a wide range of methods that have recently emerged as powerful alternatives, including nonlinear metric learning, similarity learning and local metric learning. Recent trends and extensions, such as semi-supervised metric learning, metric learning for histogram data and the derivation of generalization guarantees, are also covered. Finally, this survey addresses metric learning for structured data, in particular edit distance learning, and attempts to give an overview of the remaining challenges in metric learning for the years to come.

A Survey on Metric Learning for Feature Vectors and Structured Data

Citations

UMAP: Uniform Manifold Approximation and Projection for Dimension Reduction

Prototypical Networks for Few-shot Learning

Siamese Neural Networks for One-shot Image Recognition

Multi-View Discriminant Analysis

Learning local feature descriptors with triplets and shallow convolutional neural networks.

References

Maximum likelihood from incomplete data via the EM algorithm

Support-Vector Networks

Convex Optimization

Greedy function approximation: A gradient boosting machine.

On Information and Sufficiency

Related Papers (5)

Distance Metric Learning for Large Margin Nearest Neighbor Classification

Distance Metric Learning with Application to Clustering with Side-Information

Information-theoretic metric learning

Neighbourhood Components Analysis

Learning a similarity metric discriminatively, with application to face verification