awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
salesforce avatar

salesforce/ALBEFArchived

0
View on GitHub↗
1,758 stars·219 forks·Python·BSD-3-Clause·9 views

ALBEF

This is the official PyTorch implementation of the ALBEF paper [Blog] . This repository supports pre-training on custom datasets, as well as finetuning on VQA, SNLI-VE, NLVR2, Image-Text Retrieval on MSCOCO and Flickr30k, and visual grounding on RefCOCO+. Pre-trained and finetuned checkpoints…

Features

  • Multimodal Pretraining - Vision and language learning using momentum distillation.
  • Vision Language Models - Momentum distillation for vision and language representation learning.

Star history

Star history chart for salesforce/albefStar history chart for salesforce/albef

How this analysis was created: This summary and feature list were written by an AI model that read the project's README and public documentation pages. Each feature links to the documentation it came from; stars, license and language come straight from the GitHub API. The model does not read the source code, and the analysis is refreshed when the project is re-analysed. Learn more on our About page.

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Start searching with AI

Frequently asked questions

What does salesforce/albef do?

This is the official PyTorch implementation of the ALBEF paper [Blog] . This repository supports pre-training on custom datasets, as well as finetuning on VQA, SNLI-VE, NLVR2, Image-Text Retrieval on MSCOCO and Flickr30k, and visual grounding on RefCOCO+. Pre-trained and finetuned checkpoints…

What are the main features of salesforce/albef?

The main features of salesforce/albef are: Multimodal Pretraining, Vision Language Models.

What are some open-source alternatives to salesforce/albef?

Open-source alternatives to salesforce/albef include: jackroos/vl-bert — Code for ICLR 2020 paper "VL-BERT: Pre-training of Generic Visual-Linguistic Representations". uclanlp/visualbert. airsplay/lxmert — Our servers break again :(. I have updated the links so that they should work fine now. Sorry for the inconvenience.… evolvinglmms-lab/otter — Otter is a framework and toolkit for the pretraining, fine-tuning, and evaluation of vision-language models. It… google-research/big_vision — This project is a research framework and toolkit designed for training large-scale vision transformers and multimodal… apple/ml-aim — This repository provides the code and model checkpoints for AIMv1 and AIMv2 research projects.

Open-source alternatives to ALBEF

Similar open-source projects, ranked by how many features they share with ALBEF.
  • jackroos/vl-bertjackroos avatar

    jackroos/VL-BERT

    745View on GitHub↗

    Code for ICLR 2020 paper "VL-BERT: Pre-training of Generic Visual-Linguistic Representations".

    Jupyter Notebook
    View on GitHub↗745
  • uclanlp/visualbertU

    uclanlp/visualbert

    0View on GitHub↗
    View on GitHub↗0
  • airsplay/lxmertairsplay avatar

    airsplay/lxmert

    967View on GitHub↗

    Our servers break again :(. I have updated the links so that they should work fine now. Sorry for the inconvenience. Please let me for any further issues. Thanks! --Hao, Dec 03

    Python
    View on GitHub↗967
  • google-research/big_visiongoogle-research avatar

    google-research/big_vision

    3,363View on GitHub↗

    This project is a research framework and toolkit designed for training large-scale vision transformers and multimodal language models. It provides a comprehensive suite for vision-language pretraining, enabling the development of models that map images and text into shared latent spaces. The framework is distinguished by its capabilities in high-fidelity image generation and multimodal research, utilizing normalizing flows and variational autoencoders to produce images from text prompts or class labels. It supports the development of both generative and contrastive models, allowing for a wide

    Jupyter Notebook
    View on GitHub↗3,363
See all 30 alternatives to ALBEF→