CVPR 2024 Prompt Highlighter: Interactive Control for Multi-Modal LLMs
Efficient computing methods developed by Huawei Noah's Ark Lab
🐟 Code and models for the NeurIPS 2023 paper "Generating Images with Multimodal Language Models".
🧀 Code and models for the ICML 2023 paper "Grounding Language Models to Images for Multimodal Inputs and Outputs".
The main features of kohjingyu/fromage are: Model Utilities.
Open-source alternatives to kohjingyu/fromage include: ailab-cvc/seed — Official implementation of SEED-LLaMA (ICLR 2024). dvlab-research/prompt-highlighter — [CVPR 2024] Prompt Highlighter: Interactive Control for Multi-Modal LLMs. huawei-noah/efficient-computing — Efficient computing methods developed by Huawei Noah's Ark Lab. kohjingyu/gill — 🐟 Code and models for the NeurIPS 2023 paper "Generating Images with Multimodal Language Models". shi-labs/vcoder — [CVPR 2024] VCoder: Versatile Vision Encoders for Multimodal Large Language Models. tingyu215/ts-llava — TS-LLaVA: Constructing Visual Tokens through Thumbnail-and-Sampling for Training-Free Video Large Language Models.