This is the code page for people-scene interaction modeling, and generating 3D people in scenes without people.
The main features of yz-cnsdqz/psi-release are: Body Reshaping And Generation, Human Reconstruction.
Open-source alternatives to yz-cnsdqz/psi-release include: facebookresearch/pifuhd — pifuhd is a 3D human reconstruction framework that generates high-resolution 3D meshes of people from a single 2D… idea-research/grounded-segment-anything — Grounded-Segment-Anything is a suite of specialized tools for multimodal visual analysis, text-based segmentation, and… arthur151/romp — | ROMP | BEV | TRACE | | :---: | :---: | :---: | | Monocular, One-stage, Regression of Multiple 3D People (ICCV21) |… facebookresearch/d3d-hoi — Xiang Xu, Hanbyul Joo, Greg Mori, Manolis Savva. brjathu/t3dp — Code repository for the paper "Tracking People with 3D Representations". \ Jathushan Rajasegaran, Georgios Pavlakos,… andreiburov/dsfn — Andrei Burov, Matthias Nießner, Justus Thies.
pifuhd is a 3D human reconstruction framework that generates high-resolution 3D meshes of people from a single 2D image. It utilizes pixel-aligned implicit functions to map image pixels to 3D space, predicting surface occupancy and distance to create detailed geometry. The system includes a pipeline for creating digital human assets, moving from 2D image feature projection to the extraction of discrete triangular meshes. It features specialized tools for refining these models, including a post-processor that removes geometric artifacts by isolating the largest connected component of the mesh.
Grounded-Segment-Anything is a suite of specialized tools for multimodal visual analysis, text-based segmentation, and generative image editing. It integrates text-to-bounding-box detection and high-precision image segmentation masks to function as a text-based image segmenter and an automated visual labeling tool. The project enables text-driven image editing by identifying objects through natural language to perform inpainting and element replacement. It further extends visual analysis into three dimensions, allowing for 3D human reconstruction and the generation of 3D bounding boxes from t
| ROMP | BEV | TRACE | | :---: | :---: | :---: | | Monocular, One-stage, Regression of Multiple 3D People (ICCV21) | Putting People in their Place: Monocular Regression of 3D People in Depth (CVPR2022) | TRACE: 5D Temporal Regression of Avatars with Dynamic Cameras in 3D Environments (CVPR2023)…
Code repository for the paper "Tracking People with 3D Representations". \ Jathushan Rajasegaran, Georgios Pavlakos, Angjoo Kanazawa, Jitendra Malik.\ Neural Information Processing Systems (NeurIPS), 2021. \