How this analysis was created: This summary and feature list were written by an AI model that read the project's README and public documentation pages. Each feature links to the documentation it came from; stars, license and language come straight from the GitHub API. The model does not read the source code, and the analysis is refreshed when the project is re-analysed. Learn more on our About page.
Our servers break again :(. I have updated the links so that they should work fine now. Sorry for the inconvenience. Please let me for any further issues. Thanks! --Hao, Dec 03
ConvLLaVA: Hierarchical Backbones as Visual Encoder for Large Multimodal Models
An Open-source Toolkit for LLM Development
Welcome to the official repository for LLM2CLIP! This project leverages large language models (LLMs) as powerful textual teachers for CLIP's visual encoder, enabling more nuanced and comprehensive multimodal learning.
The main features of microsoft/llm2clip are: Vision Language Models.
Open-source alternatives to microsoft/llm2clip include: airsplay/lxmert — Our servers break again :(. I have updated the links so that they should work fine now. Sorry for the inconvenience.… alibaba/conv-llava — ConvLLaVA: Hierarchical Backbones as Visual Encoder for Large Multimodal Models. alpha-vllm/llama2-accessory — An Open-source Toolkit for LLM Development. apple/ml-aim — This repository provides the code and model checkpoints for AIMv1 and AIMv2 research projects. baaivision/eve — 2024/05: Unveiling Encoder-Free Vision-Language Models (NeurIPS 2024, spotlight). aidc-ai/parrot — 📍Quick Start • 👨🏫Acknowledgement • 🤗Contact.