Efficient computing methods developed by Huawei Noah's Ark Lab
🧀 Code and models for the ICML 2023 paper "Grounding Language Models to Images for Multimodal Inputs and Outputs".
🐟 Code and models for the NeurIPS 2023 paper "Generating Images with Multimodal Language Models".
CVPR 2024 Prompt Highlighter: Interactive Control for Multi-Modal LLMs
Official implementation of SEED-LLaMA (ICLR 2024).
The main features of ailab-cvc/seed are: Model Utilities.
Open-source alternatives to ailab-cvc/seed include: dvlab-research/prompt-highlighter — [CVPR 2024] Prompt Highlighter: Interactive Control for Multi-Modal LLMs. huawei-noah/efficient-computing — Efficient computing methods developed by Huawei Noah's Ark Lab. kohjingyu/fromage — 🧀 Code and models for the ICML 2023 paper "Grounding Language Models to Images for Multimodal Inputs and Outputs". kohjingyu/gill — 🐟 Code and models for the NeurIPS 2023 paper "Generating Images with Multimodal Language Models". shi-labs/vcoder — [CVPR 2024] VCoder: Versatile Vision Encoders for Multimodal Large Language Models. tingyu215/ts-llava — TS-LLaVA: Constructing Visual Tokens through Thumbnail-and-Sampling for Training-Free Video Large Language Models.