Beyond Random Augmentations:
Pretraining with Hard Views (ICLR 2025)

Self-Supervised Learning (SSL) methods typically rely on random image augmentations, or views, to make models invariant to different transformations. We hypothesize that the efficacy of pretraining pipelines based on conventional random view sampling can be enhanced by explicitly selecting views that benefit the learning progress. A simple yet effective approach is to select hard views that yield a higher loss. In this paper, we propose Hard View Pretraining (HVP), a learning-free strategy that extends random view generation by exposing models to more challenging samples during SSL pretraining. HVP encompasses the following iterative steps: 1) randomly sample multiple views and forward each view through the pretrained model, 2) create pairs of two views and compute their loss, 3) adversarially select the pair yielding the highest loss according to the current model state, and 4) perform a backward pass with the selected pair. In contrast to existing hard view literature, we are the first to demonstrate hard view pretraining's effectiveness at scale, particularly training on the full ImageNet-1k dataset, and evaluating across multiple SSL methods, ConvNets, and ViTs. As a result, \MethodAbbr{} sets a new state-of-the-art on DINO ViT-B/16, reaching 78.8% linear evaluation accuracy (a 0.6% improvement) and consistent gains of 1% for both 100 and 300 epoch pretraining, with similar improvements across transfer tasks in DINO, SimSiam, iBOT, and SimCLR.

Other branches are available here:

Setup:

conda env create -f environment.yaml
conda activate hvp
conda install -c conda-forge tensorboard
pip install omegaconf

Download Model Files

(include pretraining, linear evaluation and finetuning checkpoints for both vanilla and hvp models)

Citation

Please acknowledge the usage of this code by citing the following publication:

@inproceedings{ferreira-iclr25a,
  title        = {Beyond Random Augmentations: Pretraining with Hard Views},
  author       = {F. Ferreira and I. Rapant and J. Franke and F. Hutter},
  booktitle    = {The Thirteenth International Conference on Learning Representations},
  year         = {2025},
  URL          = {https://openreview.net/forum?id=AK1C55o4r7}
}

Name		Name	Last commit message	Last commit date
Latest commit History 36 Commits
configs		configs
models		models
testing		testing
utils		utils
.gitignore		.gitignore
README.md		README.md
builder.py		builder.py
custom_transform.py		custom_transform.py
data.py		data.py
environment.yml		environment.yml
eval_knn.py		eval_knn.py
eval_linear.py		eval_linear.py
method.png		method.png
pretrain.py		pretrain.py
run_pipeline.py		run_pipeline.py
select_crops.py		select_crops.py
test.py		test.py

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

Repository files navigation

Beyond Random Augmentations:
Pretraining with Hard Views (ICLR 2025)

Setup:

Download Model Files

Citation

About

Releases

Packages

Languages

automl/hvp

Folders and files

Latest commit

History

Repository files navigation

Beyond Random Augmentations: Pretraining with Hard Views (ICLR 2025)

Setup:

Download Model Files

Citation

About

Resources

Stars

Watchers

Forks

Releases

Packages 0

Languages

Beyond Random Augmentations:
Pretraining with Hard Views (ICLR 2025)

Packages