awesome-repositories.com
博客
MCP
awesome-repositories.com

通过 AI 驱动的搜索,发现最优秀的开源仓库。

探索精选搜索开源替代品自托管软件博客网站地图
项目MCP 服务器关于排名机制媒体报道
法律隐私政策服务条款
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
Back to microsoft/i-code

Open-source alternatives to I Code

30 open-source projects similar to microsoft/i-code, ranked by how many features they have in common. Compare stars, activity and what each one does to find the best I Code alternative.

  • simular-ai/agent-ssimular-ai 的头像

    simular-ai/Agent-S

    11,855在 GitHub 上查看↗

    Agent-S is a multimodal AI agent and LLM desktop automation framework designed to control operating systems through graphical user interface interactions. It functions as a computer use interface, utilizing vision-language grounding to translate natural language goals into precise screen coordinates and system actions. The project differentiates itself by combining structured accessibility tree inspection with vision-based element localization. It manages cross-application workflows by mapping conceptual descriptions to physical pixels and simulating low-level keyboard and mouse events to mov

    Pythonagent-computer-interfaceai-agentscomputer-automation
    在 GitHub 上查看↗11,855
  • qwenlm/qwen3-omniQwenLM 的头像

    QwenLM/Qwen3-Omni

    3,843在 GitHub 上查看↗

    Qwen3-Omni is an omni-modal large language model designed to process and generate text, audio, images, and video within a single unified neural architecture. It functions as a real-time voice assistant and multimodal AI agent capable of reasoning across different media types and executing external tool-calling functions via APIs. The system supports low-latency conversational AI through autoregressive token streaming and natural turn-taking. It enables multilingual speech translation and generation across dozens of languages, featuring customizable speaker profiles and tones. The model's cap

    Jupyter Notebook
    在 GitHub 上查看↗3,843
  • othersideai/self-operating-computerOthersideAI 的头像

    OthersideAI/self-operating-computer

    10,153在 GitHub 上查看↗

    This project is a computer control framework that uses multimodal vision models to simulate mouse and keyboard inputs for automating desktop tasks. It functions as an autonomous agent and vision-based orchestrator that interprets screen visuals to interact with user interfaces. The system employs vision language models and object detection to locate and click interface elements. It utilizes visual grounding to overlay numerical markers on UI components and uses optical character recognition to map on-screen text to precise pixel coordinates. The framework supports voice-controlled computing

    Pythonautomationopenaipyautogui
    在 GitHub 上查看↗10,153

AI 搜索

探索更多 awesome 仓库

用简单的语言描述您的需求 —— AI 将根据相关性为您从数千个精选开源项目中进行排序。

Find more with AI search
  • 11cafe/jaaz11cafe 的头像

    11cafe/jaaz

    6,384在 GitHub 上查看↗

    Jaaz is a self-hosted AI design suite and multimodal workspace used for generating and editing images and videos. It functions as a design workspace where users can produce visual content and assets through a combination of local and cloud-based AI models. The project features a hybrid model orchestrator that routes requests between local model runners and remote APIs to balance data privacy with processing performance. It utilizes an infinite canvas collaborative tool for organizing storyboards and assets, and includes an image prompt optimizer to translate rough ideas into detailed generati

    TypeScript
    在 GitHub 上查看↗6,384
  • aim-uofa/gendefA

    aim-uofa/GenDeF

    0在 GitHub 上查看↗
    在 GitHub 上查看↗0
  • ali-vilab/vaceali-vilab 的头像

    ali-vilab/VACE

    3,645在 GitHub 上查看↗

    VACE is a set of software tools and frameworks for reference-guided video generation, diffusion-based editing, and video-to-video translation. It provides utilities to produce new video content and modify existing sequences by using reference materials to guide visual style, subject matter, and composition. The framework enables video-to-video translation and synthesis, allowing for the update of visual styles and depth. It also functions as a video editor for modifying properties and content through reference-guided transformations. The system covers localized video editing and inpainting,

    Pythonvideo-editingvideo-generation
    在 GitHub 上查看↗3,645
  • ai-chef/hugginggptAI-Chef 的头像

    AI-Chef/HuggingGPT

    24在 GitHub 上查看↗

    The mission of JARVIS is to explore artificial general intelligence (AGI) and deliver cutting-edge research to the whole community.

    Python
    在 GitHub 上查看↗24
  • alpha-vllm/lumina-t2xAlpha-VLLM 的头像

    Alpha-VLLM/Lumina-T2X

    2,250在 GitHub 上查看↗

    Lumina-T2X is a unified framework for Text to Any Modality Generation

    Pythonaigcdiffusiondiffusion-model
    在 GitHub 上查看↗2,250
  • ailab-cvc/videocrafterailab-cvc 的头像

    ailab-cvc/videocrafter

    5,063在 GitHub 上查看↗

    Videocrafter is a latent diffusion model designed for AI video synthesis. It functions as both a text-to-video and image-to-video generation system, synthesizing high-quality video sequences from descriptive text prompts or static image inputs. The model utilizes a diffusion-based neural network to transform inputs into animated content, ensuring visual consistency and temporal coherence throughout the generated sequences. This allows for the creation of custom video clips and the animation of static images into fluid motion.

    Python
    在 GitHub 上查看↗5,063
  • alpha-vllm/lumina-videoA

    Alpha-VLLM/Lumina-Video

    0在 GitHub 上查看↗
    在 GitHub 上查看↗0
  • aniki-ly/flowzeroA

    aniki-ly/FlowZero

    0在 GitHub 上查看↗
    在 GitHub 上查看↗0
  • anonymous0769/dreamvideoA

    anonymous0769/DreamVideo

    0在 GitHub 上查看↗
    在 GitHub 上查看↗0
  • anonymous0x233/reuseanddiffuseA

    anonymous0x233/ReuseAndDiffuse

    0在 GitHub 上查看↗
    在 GitHub 上查看↗0
  • araachie/riverA

    araachie/river

    0在 GitHub 上查看↗
    在 GitHub 上查看↗0
  • baaivision/emubaaivision 的头像

    baaivision/Emu

    1,775在 GitHub 上查看↗

    Emu Series: Generative Multimodal Models from BAAI

    Python
    在 GitHub 上查看↗1,775
  • allenai/unified-io-2allenai 的头像

    allenai/unified-io-2

    649在 GitHub 上查看↗

    This repo contains code for Unified-IO 2, including code to run a demo, do training, and do inference. This codebase is modified from T5X.

    Python
    在 GitHub 上查看↗649
  • arthur-qiu/longercrafterA

    arthur-qiu/LongerCrafter

    0在 GitHub 上查看↗
    在 GitHub 上查看↗0
  • baaivision/novaB

    baaivision/NOVA

    0在 GitHub 上查看↗
    在 GitHub 上查看↗0
  • boese0601/magicdanceB

    Boese0601/MagicDance

    0在 GitHub 上查看↗
    在 GitHub 上查看↗0
  • buggyyang/rvdbuggyyang 的头像

    buggyyang/rvd

    63在 GitHub 上查看↗

    Diffusion Probabilistic Modeling for Video Generation

    Python
    在 GitHub 上查看↗63
  • chenhsing/simdaC

    ChenHsing/SimDA

    0在 GitHub 上查看↗
    在 GitHub 上查看↗0
  • cliport/cliportcliport 的头像

    cliport/cliport

    545在 GitHub 上查看↗

    CLIPort: What and Where Pathways for Robotic Manipulation Mohit Shridhar, Lucas Manuelli, Dieter Fox CoRL 2021

    Jupyter Notebook
    在 GitHub 上查看↗545
  • cvlab-columbia/vipercvlab-columbia 的头像

    cvlab-columbia/viper

    1,717在 GitHub 上查看↗

    Code for the paper "ViperGPT: Visual Inference via Python Execution for Reasoning"

    Jupyter Notebook
    在 GitHub 上查看↗1,717
  • da-group-pku/magic-1-for-1D

    DA-Group-PKU/Magic-1-For-1

    0在 GitHub 上查看↗
    在 GitHub 上查看↗0
  • daeunni/videorepairD

    daeunni/VideoRepair

    0在 GitHub 上查看↗
    在 GitHub 上查看↗0
  • damo-vilab/i2vgen-xlD

    damo-vilab/i2vgen-xl

    0在 GitHub 上查看↗
    在 GitHub 上查看↗0
  • damo-vilab/videocomposerD

    damo-vilab/videocomposer

    0在 GitHub 上查看↗
    在 GitHub 上查看↗0
  • dlyuangod/tinygpt-vDLYuanGod 的头像

    DLYuanGod/TinyGPT-V

    1,315在 GitHub 上查看↗

    TinyGPT-V: Efficient Multimodal Large Language Model via Small Backbones

    Python
    在 GitHub 上查看↗1,315
  • doubiiu/dynamicrafterDoubiiu 的头像

    Doubiiu/DynamiCrafter

    3,001在 GitHub 上查看↗

    ECCV 2024, Oral DynamiCrafter: Animating Open-domain Images with Video Diffusion Priors

    Pythonimage-animationimage-to-videovideo-generation
    在 GitHub 上查看↗3,001
  • araachie/yodaA

    araachie/yoda

    0在 GitHub 上查看↗
    在 GitHub 上查看↗0