https://github.com/animesh/git-re-basin

Code release for "Git Re-Basin: Merging Models modulo Permutation Symmetries"

Science Score: 10.0%

This score indicates how likely this project is to be science-related based on various indicators:

○
CITATION.cff file
○
codemeta.json file
○
.zenodo.json file
○
DOI references
✓
Academic publication links
Links to: arxiv.org
○
Academic email domains
○
Institutional organization owner
○
JOSS paper metadata
○
Scientific vocabulary similarity
Low similarity (6.4%) to scientific vocabulary

Last synced: 6 months ago · JSON representation

Repository

Code release for "Git Re-Basin: Merging Models modulo Permutation Symmetries"

Basic Info

Host: GitHub
Owner: animesh
Default Branch: main
Homepage: https://arxiv.org/abs/2209.04836
Size: 1.65 MB

Statistics

Stars: 0
Watchers: 0
Forks: 0
Open Issues: 0
Releases: 0

Fork of samuela/git-re-basin

Created over 3 years ago · Last pushed over 3 years ago

https://github.com/animesh/git-re-basin/blob/main/

# Git Re-Basin: Merging Models modulo Permutation Symmetries

![Video demonstrating the effect of our permutation matching algorithm on the loss landscape throughout training.](mnist_video.gif)

Code for the paper [Git Re-Basin: Merging Models modulo Permutation Symmetries](https://arxiv.org/abs/2209.04836).

Abstract:

> The success of deep learning is thanks to our ability to solve certain massive non-convex optimization problems with relative ease. Despite non-convex optimization being NP-hard, simple algorithms -- often variants of stochastic gradient descent -- exhibit surprising effectiveness in fitting large neural networks in practice. We argue that neural network loss landscapes contain (nearly) a single basin, after accounting for all possible permutation symmetries of hidden units. We introduce three algorithms to permute the units of one model to bring them into alignment with units of a reference model. This transformation produces a functionally equivalent set of weights that lie in an approximately convex basin near the reference model. Experimentally, we demonstrate the single basin phenomenon across a variety of model architectures and datasets, including the first (to our knowledge) demonstration of zero-barrier linear mode connectivity between independently trained ResNet models on CIFAR-10 and CIFAR-100. Additionally, we identify intriguing phenomena relating model width and training time to mode connectivity across a variety of models and datasets. Finally, we discuss shortcomings of a single basin theory, including a counterexample to the linear mode connectivity hypothesis.

Owner

Name: Ani
Login: animesh
Kind: user
Location: Norway
Company: Norwegian University of Science and Technology

Website: https://www.fuzzylife.org
Twitter: animesh1977
Repositories: 749
Profile: https://github.com/animesh

A medical graduate from Delhi University with post-graduation in bioinformatics from Jawaharlal Nehru University, India.

ecosyste.ms

Data

Tools

Indexes

Applications

Experiments

Open Source Science

https://github.com/animesh/git-re-basin

Science Score: 10.0%

Repository

Basic Info

Statistics

https://github.com/animesh/git-re-basin/blob/main/

Owner

GitHub Events

Total

Last Year