awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
jackroos avatar

jackroos/VL-BERT

0
View on GitHub↗
745 stars·110 forks·Jupyter Notebook·MIT·8 views

VL BERT

Code for ICLR 2020 paper "VL-BERT: Pre-training of Generic Visual-Linguistic Representations".

Features

  • Multimodal Pretraining - Pre-training generic visual-linguistic representations.
  • Natural Language Processing - Pre-training generic visual-linguistic representations.
  • Vision Language Models - Generic visual-linguistic representation pre-training.

Star history

Star history chart for jackroos/vl-bertStar history chart for jackroos/vl-bert

How this analysis was created: This summary and feature list were written by an AI model that read the project's README and public documentation pages. Each feature links to the documentation it came from; stars, license and language come straight from the GitHub API. The model does not read the source code, and the analysis is refreshed when the project is re-analysed. Learn more on our About page.

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Start searching with AI

Open-source alternatives to VL BERT

Similar open-source projects, ranked by how many features they share with VL BERT.
  • huggingface/transformershuggingface avatar

    huggingface/transformers

    161,630View on GitHub↗

    Transformers is a comprehensive library for machine learning that provides a unified interface for training, fine-tuning, and deploying transformer-based models. It supports a wide range of tasks, including text classification, language modeling, question answering, and sequence-to-sequence translation, while offering specialized architectures for both text and vision processing. The framework includes tools for managing the entire model lifecycle, from data preprocessing and tokenization to distributed training and inference. The library features extensive support for model optimization and

    Pythonaudiodeep-learningdeepseek
    View on GitHub↗161,630
  • salesforce/albefsalesforce avatar

    salesforce/ALBEF

    1,758View on GitHub↗

    This is the official PyTorch implementation of the ALBEF paper Blog . This repository supports pre-training on custom datasets, as well as finetuning on VQA, SNLI-VE, NLVR2, Image-Text Retrieval on MSCOCO and Flickr30k, and visual grounding on RefCOCO+. Pre-trained and finetuned checkpoints…

    Python
    View on GitHub↗1,758
  • airsplay/lxmertairsplay avatar

    airsplay/lxmert

    967View on GitHub↗

    Our servers break again :(. I have updated the links so that they should work fine now. Sorry for the inconvenience. Please let me for any further issues. Thanks! --Hao, Dec 03

    Python
    View on GitHub↗967
  • uclanlp/visualbertU

    uclanlp/visualbert

    0View on GitHub↗
    View on GitHub↗0
See all 30 alternatives to VL BERT→

Frequently asked questions

What does jackroos/vl-bert do?

Code for ICLR 2020 paper "VL-BERT: Pre-training of Generic Visual-Linguistic Representations".

What are the main features of jackroos/vl-bert?

The main features of jackroos/vl-bert are: Multimodal Pretraining, Natural Language Processing, Vision Language Models.

What are some open-source alternatives to jackroos/vl-bert?

Open-source alternatives to jackroos/vl-bert include: huggingface/transformers — Transformers is a comprehensive library for machine learning that provides a unified interface for training,… salesforce/albef — This is the official PyTorch implementation of the ALBEF paper [Blog] . This repository supports pre-training on… airsplay/lxmert — Our servers break again :(. I have updated the links so that they should work fine now. Sorry for the inconvenience.… uclanlp/visualbert. evolvinglmms-lab/otter — Otter is a framework and toolkit for the pretraining, fine-tuning, and evaluation of vision-language models. It… google-research/big_vision — This project is a research framework and toolkit designed for training large-scale vision transformers and multimodal…