awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
RupertLuo avatar

RupertLuo/Valley

0
View on GitHub↗
231 stars·15 forks·Python·8 views

Valley

Understanding Complex Videos Relying on Large Language and Vision Models [Project Page] [Paper] The online demo is no longer available, because we released the code for offline demo deployment

Features

  • Multimodal Learning - Enhances video assistant capabilities with language models.

Star history

Star history chart for rupertluo/valleyStar history chart for rupertluo/valley

How this analysis was created: This summary and feature list were written by an AI model that read the project's README and public documentation pages. Each feature links to the documentation it came from; stars, license and language come straight from the GitHub API. The model does not read the source code, and the analysis is refreshed when the project is re-analysed. Learn more on our About page.

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Start searching with AI

Frequently asked questions

What does rupertluo/valley do?

Understanding Complex Videos Relying on Large Language and Vision Models [Project Page] [Paper] The online demo is no longer available, because we released the code for offline demo deployment

What are the main features of rupertluo/valley?

The main features of rupertluo/valley are: Multimodal Learning.

What are some open-source alternatives to rupertluo/valley?

Open-source alternatives to rupertluo/valley include: deepmind/deepmind-research — This project is an AI research implementation library and machine learning research repository. It provides a… deepseek-ai/janus — Janus is a multimodal large language model and unified framework that integrates visual understanding and image… fuxiaoliu/lrv-instruction — [ICLR'24] Mitigating Hallucination in Large Multi-Modal Models via Robust Instruction Tuning. hanzhanggit/stackgan — Pytorch implementation. haotian-liu/llava — LLaVA is a multimodal large language model architecture designed to process and interpret both image and text inputs… baai-dcai/visual-instruction-tuning — Scale up visual instruction tuning to millions by GPT-4.

Open-source alternatives to Valley

Similar open-source projects, ranked by how many features they share with Valley.
  • deepmind/deepmind-researchdeepmind avatar

    deepmind/deepmind-research

    15,024View on GitHub↗

    This project is an AI research implementation library and machine learning research repository. It provides a collection of reference code, illustrative implementations, and open-source research datasets used to verify hypotheses and build upon existing models in artificial intelligence. The repository focuses on scientific research reproduction by translating theoretical findings from published papers into executable code. It includes specialized scientific simulation environments designed to test the behavior of autonomous agents and models within controlled settings. The project covers AI

    Jupyter Notebook
    View on GitHub↗15,024
  • deepseek-ai/janusdeepseek-ai avatar

    deepseek-ai/Janus

    17,746View on GitHub↗

    Janus is a multimodal large language model and unified framework that integrates visual understanding and image generation within a single neural network. It functions as both a visual understanding model for analyzing images and a text-to-image generator. The system uses a unified transformer backbone and a multimodal latent space to bridge the gap between text and visual data. This architecture employs decoupled visual encoding and cross-modal tokenization to separate the paths for discriminative understanding and generative tasks, representing images as grids of discrete codes. The projec

    Pythonany-to-anyfoundation-modelsllm
    View on GitHub↗17,746
  • fuxiaoliu/lrv-instructionFuxiaoLiu avatar

    FuxiaoLiu/LRV-Instruction

    297View on GitHub↗

    ICLR'24 Mitigating Hallucination in Large Multi-Modal Models via Robust Instruction Tuning

    Pythonchatgptevaluationevaluation-metrics
    View on GitHub↗297
  • baai-dcai/visual-instruction-tuningBAAI-DCAI avatar

    BAAI-DCAI/Visual-Instruction-Tuning

    168View on GitHub↗

    Scale up visual instruction tuning to millions by GPT-4.

    Python
    View on GitHub↗168
See all 18 alternatives to Valley→