awesome-repositories.com
المدونة
MCP
awesome-repositories.com

اكتشف أفضل مستودعات المصادر المفتوحة باستخدام بحث مدعوم بالذكاء الاصطناعي.

استكشفعمليات بحث منسقةبدائل مفتوحة المصدربرمجيات ذاتية الاستضافةالمدونةخريطة الموقع
المشروعحولكيفية ترتيب النتائجالصحافةخادم MCP
قانونيالخصوصيةالشروط
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
airsplay avatar

airsplay/vokenization

0
View on GitHub↗
191 نجوم·21 تفرعات·Python·MIT·2 مشاهدات

Vokenization

PyTorch code for EMNLP 2020 Paper "Vokenization: Improving Language Understanding with Visual Supervision"

Features

  • Multimodal Pretraining - Improving language understanding with visually-grounded supervision.

سجل النجوم

مخطط تاريخ النجوم لـ airsplay/vokenizationمخطط تاريخ النجوم لـ airsplay/vokenization

بحث بالذكاء الاصطناعي

استكشف المزيد من المستودعات الرائعة

صف ما تحتاجه بلغة بسيطة — وسيقوم الذكاء الاصطناعي بترتيب آلاف المشاريع مفتوحة المصدر المنسقة حسب الصلة.

Start searching with AI

بدائل مفتوحة المصدر لـ Vokenization

مشاريع مفتوحة المصدر مشابهة، مرتبة حسب عدد الميزات المشتركة مع Vokenization.
  • google-research/big_visionالصورة الرمزية لـ google-research

    google-research/big_vision

    3,363عرض على GitHub↗

    This project is a research framework and toolkit designed for training large-scale vision transformers and multimodal language models. It provides a comprehensive suite for vision-language pretraining, enabling the development of models that map images and text into shared latent spaces. The framework is distinguished by its capabilities in high-fidelity image generation and multimodal research, utilizing normalizing flows and variational autoencoders to produce images from text prompts or class labels. It supports the development of both generative and contrastive models, allowing for a wide

    Jupyter Notebook
    عرض على GitHub↗3,363
  • evolvinglmms-lab/otterالصورة الرمزية لـ EvolvingLMMs-Lab

    EvolvingLMMs-Lab/Otter

    3,331عرض على GitHub↗

    Otter is a framework and toolkit for the pretraining, fine-tuning, and evaluation of vision-language models. It provides a pipeline for training large language models to process high-resolution images and video frames, integrating visual encoders with textual token spaces. The system is designed for multi-visual input processing, allowing models to interpret multiple images or video sequences within a single prompt. It supports multi-round conversation management to maintain context across interactions for detailed scene comprehension and visual reasoning. The framework covers a full develop

    Pythonartificial-inteligencechatgptdeep-learning
    عرض على GitHub↗3,331
  • jayleicn/clipbertالصورة الرمزية لـ jayleicn

    jayleicn/ClipBERT

    730عرض على GitHub↗

    Less is More: ClipBERT for Video-and-Language Learning via Sparse Sampling

    Python
    عرض على GitHub↗730
  • salesforce/albefالصورة الرمزية لـ salesforce

    salesforce/ALBEF

    1,758عرض على GitHub↗

    This is the official PyTorch implementation of the ALBEF paper Blog . This repository supports pre-training on custom datasets, as well as finetuning on VQA, SNLI-VE, NLVR2, Image-Text Retrieval on MSCOCO and Flickr30k, and visual grounding on RefCOCO+. Pre-trained and finetuned checkpoints…

    Python
    عرض على GitHub↗1,758
عرض جميع البدائل الـ 8 لـ Vokenization→

الأسئلة الشائعة

ما هي وظيفة airsplay/vokenization؟

PyTorch code for EMNLP 2020 Paper "Vokenization: Improving Language Understanding with Visual Supervision"

ما هي الميزات الرئيسية لـ airsplay/vokenization؟

الميزات الرئيسية لـ airsplay/vokenization هي: Multimodal Pretraining.

ما هي البدائل مفتوحة المصدر لـ airsplay/vokenization؟

تشمل البدائل مفتوحة المصدر لـ airsplay/vokenization: google-research/big_vision — This project is a research framework and toolkit designed for training large-scale vision transformers and multimodal… evolvinglmms-lab/otter — Otter is a framework and toolkit for the pretraining, fine-tuning, and evaluation of vision-language models. It… jayleicn/clipbert — Less is More: ClipBERT for Video-and-Language Learning via Sparse Sampling. salesforce/albef — This is the official PyTorch implementation of the ALBEF paper [Blog] . This repository supports pre-training on… uclanlp/visualbert. airsplay/lxmert — Our servers break again :(. I have updated the links so that they should work fine now. Sorry for the inconvenience.…