English
 
Help Privacy Policy Disclaimer
  Advanced SearchBrowse

Item

ITEM ACTIONSEXPORT
 
 
DownloadE-Mail
 PreviousNext  
  Factoring Out Prior Knowledge from Low-dimensional Embeddings

Heiter, E., Fischer, J., & Vreeken, J. (2021). Factoring Out Prior Knowledge from Low-dimensional Embeddings. Retrieved from https://arxiv.org/abs/2103.01828.

Item is

Files

show Files
hide Files
:
arXiv:2103.01828.pdf (Preprint), 13MB
Name:
arXiv:2103.01828.pdf
Description:
File downloaded from arXiv at 2021-03-04 10:38
OA-Status:
Visibility:
Public
MIME-Type / Checksum:
application/pdf / [MD5]
Technical Metadata:
Copyright Date:
-
Copyright Info:
-

Locators

show

Creators

show
hide
 Creators:
Heiter, Edith1, Author
Fischer, Jonas2, Author           
Vreeken, Jilles1, Author           
Affiliations:
1External Organizations, ou_persistent22              
2Databases and Information Systems, MPI for Informatics, Max Planck Society, ou_24018              

Content

show
hide
Free keywords: Computer Science, Learning, cs.LG,Statistics, Machine Learning, stat.ML
 Abstract: Low-dimensional embedding techniques such as tSNE and UMAP allow visualizing
high-dimensional data and therewith facilitate the discovery of interesting
structure. Although they are widely used, they visualize data as is, rather
than in light of the background knowledge we have about the data. What we
already know, however, strongly determines what is novel and hence interesting.
In this paper we propose two methods for factoring out prior knowledge in the
form of distance matrices from low-dimensional embeddings. To factor out prior
knowledge from tSNE embeddings, we propose JEDI that adapts the tSNE objective
in a principled way using Jensen-Shannon divergence. To factor out prior
knowledge from any downstream embedding approach, we propose CONFETTI, in which
we directly operate on the input distance matrices. Extensive experiments on
both synthetic and real world data show that both methods work well, providing
embeddings that exhibit meaningful structure that would otherwise remain
hidden.

Details

show
hide
Language(s): eng - English
 Dates: 2021-03-022021
 Publication Status: Published online
 Pages: 27 p.
 Publishing info: -
 Table of Contents: -
 Rev. Type: -
 Identifiers: arXiv: 2103.01828
URI: https://arxiv.org/abs/2103.01828
BibTex Citekey: heiter:21:factoring
 Degree: -

Event

show

Legal Case

show

Project information

show

Source

show