awesome-repositories.com
المدونة
MCP
awesome-repositories.com

اكتشف أفضل مستودعات المصادر المفتوحة باستخدام بحث مدعوم بالذكاء الاصطناعي.

استكشفعمليات بحث منسقةبدائل مفتوحة المصدربرمجيات ذاتية الاستضافةالمدونةخريطة الموقع
المشروعخادم MCPحولكيفية ترتيب النتائجالصحافة
قانونيالخصوصيةالشروط
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·

5 مستودعات

Awesome GitHub Repositories3D

Locating and classifying objects within three-dimensional space using neural networks.

Distinct from Object Detection: Extends standard 2D object detection into 3D spatial coordinates.

Explore 5 awesome GitHub repositories matching artificial intelligence & ml · 3D. Refine with filters or upvote what's useful.

Awesome 3D GitHub Repositories

اعثر على أفضل المستودعات باستخدام الذكاء الاصطناعي.سنبحث عن أفضل المستودعات المطابقة باستخدام الذكاء الاصطناعي.
  • facebookresearch/detectron2الصورة الرمزية لـ facebookresearch

    facebookresearch/detectron2

    34,548عرض على GitHub↗

    Detectron2 is a PyTorch computer vision framework and visual recognition platform designed for training and deploying models for object detection, image segmentation, and visual recognition. It provides a research-oriented environment for training complex vision models with multi-GPU acceleration. The project includes a specialized object detection library for identifying and locating multiple objects via bounding boxes, as well as an image segmentation toolkit for creating pixel-level masks through instance, semantic, and panoptic segmentation. Additionally, it features a human pose estimati

    Locates and classifies objects within three-dimensional space using fully convolutional networks.

    Python
    عرض على GitHub↗34,548
  • idea-research/grounded-segment-anythingالصورة الرمزية لـ IDEA-Research

    IDEA-Research/Grounded-Segment-Anything

    17,633عرض على GitHub↗

    Grounded-Segment-Anything is a suite of specialized tools for multimodal visual analysis, text-based segmentation, and generative image editing. It integrates text-to-bounding-box detection and high-precision image segmentation masks to function as a text-based image segmenter and an automated visual labeling tool. The project enables text-driven image editing by identifying objects through natural language to perform inpainting and element replacement. It further extends visual analysis into three dimensions, allowing for 3D human reconstruction and the generation of 3D bounding boxes from t

    Extends two-dimensional segmentation masks into three-dimensional bounding boxes by projecting image coordinates.

    Jupyter Notebook3d-whole-body-pose-estimationautomatic-labeling-systemcaption
    عرض على GitHub↗17,633
  • xingyizhou/centernetالصورة الرمزية لـ xingyizhou

    xingyizhou/CenterNet

    7,565عرض على GitHub↗

    CenterNet هو إطار عمل لاكتشاف الكائنات بنقطة المركز وخط أنابيب رؤية حاسوبية في الوقت الفعلي. يحدد الكائنات والوضعيات من خلال التنبؤ بنقاط المركز بدلاً من استخدام صناديق التثبيت (Anchor boxes). يعمل النظام كمقدر لصناديق الإحاطة ثلاثية الأبعاد، ونموذج لتقدير وضعية الإنسان، وأداة لاكتشاف الكائنات في الوقت الفعلي. يعامل وضع المفاصل ومواقع الكائنات كمشكلات اكتشاف نقطة المركز لتحديد الكيانات في الصور والفضاء ثلاثي الأبعاد. تغطي القدرات اكتشاف الكائنات ثلاثية الأبعاد، وتقدير النقاط الرئيسية للإنسان، وتحليل الفيديو المباشر. يستخدم خط الأنابيب عملية استدلال تغذية أمامية أحادية المرحلة لإجراء تحليل مستمر على كاميرات الويب أو ملفات الفيديو.

    Locates and classifies objects within three-dimensional space using center point coordinates.

    Python
    عرض على GitHub↗7,565
  • francescopace/espectreالصورة الرمزية لـ francescopace

    francescopace/espectre

    6,472عرض على GitHub↗

    Espectre is an edge machine learning framework and motion detection platform that uses Wi-Fi Channel State Information to identify human presence and movement. It functions as a sensing toolkit for ESP32 microcontrollers, enabling the detection of motion through walls without the use of cameras or wearables. The project distinguishes itself by executing compact neural network classifiers and mathematical detection algorithms directly on the microcontroller. It utilizes a MicroPython runtime to allow for the prototyping and deployment of sensing logic and wireless signal processing algorithms

    Estimates the 3D position of people or objects using an array of phase-coherent wireless nodes.

    Pythoncsidiyesp-32
    عرض على GitHub↗6,472
  • fundamentalvision/bevformerالصورة الرمزية لـ fundamentalvision

    fundamentalvision/BEVFormer

    4,519عرض على GitHub↗

    BEVFormer هو إطار عمل للإدراك يحول صور الكاميرات المتعددة إلى تمثيلات منظور عين الطائر (Bird's-eye-view) للقيادة الذاتية. يعمل كخط أنابيب رؤية متعدد الكاميرات يدمج تدفقات كاميرات متعددة في منظور مكاني موحد لتسهيل فهم البيئة. ينفذ النظام بنية قائمة على المحولات (Transformer) تستخدم استخراج الميزات القائم على الاستعلام والشبكات الزمانية المكانية لتجميع ميزات الصورة المكانية والبيانات التاريخية الزمنية. يستخدم التراكم الزمني المتكرر للحفاظ على ذاكرة مستمرة للمشهد عبر الإطارات المتتالية. يوفر إطار العمل قدرات لاكتشاف الكائنات ثلاثية الأبعاد وتقسيم الخريطة الدلالية. يجمع بين دمج الصور متعددة العروض ورأس اكتشاف تلافيفي لتحديد الكائنات ثلاثية الأبعاد وتقسيم البيانات البيئية إلى مناطق دلالية ذات معنى.

    Locates and identifies three-dimensional objects in a scene by converting camera images into a bird's-eye-view perspective.

    Pythonautonomous-drivingcomputer-visiondeep-learning
    عرض على GitHub↗4,519
  1. Home
  2. Artificial Intelligence & ML
  3. Computer Vision Systems
  4. Computer Vision
  5. Object Detection and Tracking
  6. Object Detection
  7. 3D