Contrastive Learning for Tabular Data - SCARF
Key Takeaways
The SCARF algorithm applies self-supervised contrastive learning to tabular data using random feature corruption, leveraging frameworks like Bootstrap Your Own Latent and empirical marginal distributions to form positive pairs and improve predictive performance.
Full Transcript
this video will provide a quick overview of the scarf algorithm bringing self-supervised contrastive learning to tabular data using random feature corruption so this is using frameworks like bootstrap your own layton where you have these two positive pairs one formed with the data augmentation view of the original example and bringing that kind of idea to tabular data so this is tabular data say it's information about a customer and a business things like name age maybe buying preferences and these different features that describe customer data these kind of type of data sets and we'll look through this catalog this openml cc18 is a data set of uh as a collection of tabular datasets if you're interested in just picking one of these to do academic research or getting a better sense of what these data sets look like compared to the usual deep learning data sets of vision language audio and so on even though the type of data are more common in the machine learning real world anyways but so this is a strategy to bring this kind of advancement of self-supervised pre-training contrastive learning into tabular data this quote communicates what i think is the key id of the algorithm which is how we're going to mask out and corrupt the data augmentation to form the positive pair to align the representations of something like a momentum encoding average like the bootstrap your own latent framework so describe we sample we sample some refraction of the features uniformly at random and replace each of those features by a random draw and then particularly this line from that features empirical marginal distribution so each of these features say it's the age of the customers or how many toilet paper rolls they bought if this is like a shopping basket analysis kind of problem you would have that empirical district marginal distribution marginal distribution of each of the random variables uh say you have x1 x2 x3 so and then they each have their own separate uh distributions as well as the joint distribution and so on so you're going to sample how you're going to mask that and corrupt the example from the marginal distribution of each of the features so if it's age and so on it would have a different distribution of how you sample a new age value and different densities like just say it's a different gaussian distribution with different mean variance parameters or whether it's some poisson distribution total other distribution but so that's how you're going to be sampling which features you're going to be using corrupting in order to form the data augmentation view which is the second positive pairing as you align these v v prime representations in the contrastive learning framework this is the openml cc18 collection of these tabular data sets the authors are going to test this scarf pre-training algorithm with we have all these different domains of tabular data we see the number of instances the number of features in each of the tabular records and the number of class labels to classify them in for supervised evaluation and then supervised fine tuning so each of these domains with each of the features has its own empirical marginal distribution which is how you sample the corruption for forming the data augmentation view for doing this self-supervised pre-training representation learning that facilitates with supervised learning for mapping these data sets to the this number of classes in the in this collection of data so here's some high level takeaways from the study they find the scarf pre-training improves predictive performance it improves performance in the presence of label noise and controlled experiments improves performance when labeled data is limited also in controlled experiments and then the other corruption strategies like not sampling from this empirical marginal distribution are less effective and more sensitive to feature scaling scarf isn't sensitive to batch size and is fairly insensitive to corruption rate and temperature tweaks and then the tweaks of the corruption don't work any better and then the alternatives to the info noise contrast of estimation contrastive loss don't really work any better than the tested algorithm thanks for watching this quick overview of the scarf algorithm bringing latest advances in contrast of self-supervised learning to tabular data please stay tuned for the rest of the ai weekly update series on henry ai labs [Music]
Original Description
Notion Link: https://ebony-scissor-725.notion.site/Henry-AI-Labs-Weekly-Update-July-15th-2021-a68f599395e3428c878dc74c5f0e1124
Thanks for watching! Please Subscribe!
Watch on YouTube ↗
(saves to browser)
Sign in to unlock AI tutor explanation · ⚡30
Playlist
Uploads from Connor Shorten · Connor Shorten · 0 of 60
← Previous
Next →
1
2
3
4
5
6
7
8
9
10
11
12
13
14
15
16
17
18
19
20
21
22
23
24
25
26
27
28
29
30
31
32
33
34
35
36
37
38
39
40
41
42
43
44
45
46
47
48
49
50
51
52
53
54
55
56
57
58
59
60
DenseNets
Connor Shorten
DeepWalk Explained
Connor Shorten
Inception Network Explained
Connor Shorten
StackGAN
Connor Shorten
StyleGAN
Connor Shorten
Progressive Growing of GANs Explained
Connor Shorten
Improved Techniques for Training GANs
Connor Shorten
Word2Vec Explained
Connor Shorten
Must Read Papers on GANs
Connor Shorten
Unsupervised Feature Learning
Connor Shorten
Self-Supervised GANs
Connor Shorten
Embedding Graphs with Deep Learning
Connor Shorten
Transfer Learning in GANs
Connor Shorten
ReLU Activation Function
Connor Shorten
AC-GAN Explained
Connor Shorten
SimGAN Explained
Connor Shorten
DC-GAN Explained!
Connor Shorten
ResNet Explained!
Connor Shorten
Graph Convolutional Networks
Connor Shorten
Neural Architecture Search
Connor Shorten
Henry AI Labs
Connor Shorten
Video Classification with Deep Learning
Connor Shorten
BigGANs in Data Augmentation
Connor Shorten
Introduction to Deep Learning
Connor Shorten
EfficientNet Explained!
Connor Shorten
Self-Attention GAN
Connor Shorten
Curriculum Learning in Deep Neural Networks
Connor Shorten
Deep Learning Podcast #1 | Edward Dixon | Stochastic Weight Averaging
Connor Shorten
Deep Compression
Connor Shorten
Skin Cancer Classification with Deep Learning
Connor Shorten
Deep Learning Podcast #2 | Edward Peake | Deep Learning in Medical Imaging
Connor Shorten
The Lottery Ticket Hypothesis Explained!
Connor Shorten
SqueezeNet
Connor Shorten
GauGAN Explained!
Connor Shorten
AutoML with Hyperband
Connor Shorten
DL Podcast #3 | Yannic Kilcher | Population-Based Search
Connor Shorten
Weakly Supervised Pretraining
Connor Shorten
Image Data Augmentation for Deep Learning
Connor Shorten
Unsupervised Data Augmentation
Connor Shorten
Wide ResNet Explained!
Connor Shorten
RevNet: Backpropagation without Storing Activations
Connor Shorten
GANs with Fewer Labels
Connor Shorten
BigBiGAN Unsupervised Learning!
Connor Shorten
Self-Supervised Learning
Connor Shorten
Multi-Task Self-Supervised Learning
Connor Shorten
Self-Supervised GANs
Connor Shorten
Population Based Training
Connor Shorten
Show, Attend and Tell
Connor Shorten
Siamese Neural Networks
Connor Shorten
WaveGAN Explained!
Connor Shorten
VAE-GAN Explained!
Connor Shorten
Evolution in Neural Architecture Search!
Connor Shorten
AI Research Weekly Update August 18th, 2019
Connor Shorten
Weight Agnostic Neural Networks Explained!
Connor Shorten
AI Research Weekly Update August 25th, 2019
Connor Shorten
Neuroevolution of Augmenting Topologies (NEAT)
Connor Shorten
CoDeepNEAT
Connor Shorten
AI Research Weekly Update September 1st, 2019
Connor Shorten
Randomly Wired Neural Networks
Connor Shorten
Genetic CNN
Connor Shorten
More on: LLM Foundations
View skill →Related Reads
📰
📰
📰
📰
How AI Is Transforming the Newsroom: A 2026 Overview
Medium · Machine Learning
Your AI Job Interview Is Failing You — Here’s When To Walk Away
Forbes Innovation
From Ideas to Execution: What I’m Learning as an AI Developer Intern
Medium · Startup
How AI Is Changing the Future of Jobs: Opportunity, Adaptation, and the Skills That Matter Most
Medium · AI
🎓
Tutor Explanation
DeepCamp AI