{"id":14,"date":"2025-05-10T02:55:53","date_gmt":"2025-05-10T02:55:53","guid":{"rendered":"https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/?page_id=14"},"modified":"2025-05-10T03:22:08","modified_gmt":"2025-05-10T03:22:08","slug":"related-work","status":"publish","type":"page","link":"https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/related-work\/","title":{"rendered":"Related Work"},"content":{"rendered":"\n<h2 class=\"wp-block-heading\"><strong>1. Co-evolution of pose and mesh for 3d human body estimation from video.<\/strong><\/h2>\n\n\n\n<ol class=\"wp-block-list\"><\/ol>\n\n\n\n<p><strong>Key Insight<\/strong>: Human mesh recovery is not just about using 3D keypoints (Pose) or 3D shapes (Mesh) alone\u2014it\u2019s about joint collaboration between them.<\/p>\n\n\n\n<figure class=\"wp-block-image size-large is-resized\"><img loading=\"lazy\" decoding=\"async\" width=\"1024\" height=\"515\" src=\"https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/wp-content\/uploads\/sites\/121\/2025\/05\/dmdmd-1024x515.jpg\" alt=\"\" class=\"wp-image-40\" style=\"width:678px;height:auto\" srcset=\"https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/wp-content\/uploads\/sites\/121\/2025\/05\/dmdmd-1024x515.jpg 1024w, https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/wp-content\/uploads\/sites\/121\/2025\/05\/dmdmd-300x151.jpg 300w, https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/wp-content\/uploads\/sites\/121\/2025\/05\/dmdmd-768x386.jpg 768w, https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/wp-content\/uploads\/sites\/121\/2025\/05\/dmdmd-1536x772.jpg 1536w, https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/wp-content\/uploads\/sites\/121\/2025\/05\/dmdmd.jpg 1552w\" sizes=\"auto, (max-width: 767px) 89vw, (max-width: 1000px) 54vw, (max-width: 1071px) 543px, 580px\" \/><\/figure>\n\n\n\n<ul class=\"wp-block-list\">\n<li><strong>Decoupling<\/strong> and <strong>co-evolution<\/strong> of pose estimation and mesh prediction:<\/li>\n<\/ul>\n\n\n\n<p>Poses provide information about human motion; meshes provide information about body shape<\/p>\n\n\n\n<h2 class=\"wp-block-heading\"><strong>2. World-Grounded Human Motion Recovery via Gravity-View Coordinates.<\/strong><\/h2>\n\n\n\n<p><strong>Key Insight<\/strong>: Use a Gravity-View (GV) Coordinate system to infer per-frame human motion, enabling robust result of world-grounded HMR from video.<\/p>\n\n\n\n<figure class=\"wp-block-image size-large\"><img loading=\"lazy\" decoding=\"async\" width=\"1024\" height=\"276\" src=\"https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/wp-content\/uploads\/sites\/121\/2025\/05\/Snipaste_2025-05-09_23-18-20-1024x276.jpg\" alt=\"\" class=\"wp-image-42\" srcset=\"https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/wp-content\/uploads\/sites\/121\/2025\/05\/Snipaste_2025-05-09_23-18-20-1024x276.jpg 1024w, https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/wp-content\/uploads\/sites\/121\/2025\/05\/Snipaste_2025-05-09_23-18-20-300x81.jpg 300w, https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/wp-content\/uploads\/sites\/121\/2025\/05\/Snipaste_2025-05-09_23-18-20-768x207.jpg 768w, https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/wp-content\/uploads\/sites\/121\/2025\/05\/Snipaste_2025-05-09_23-18-20-1536x414.jpg 1536w, https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/wp-content\/uploads\/sites\/121\/2025\/05\/Snipaste_2025-05-09_23-18-20-2048x551.jpg 2048w\" sizes=\"auto, (max-width: 767px) 89vw, (max-width: 1000px) 54vw, (max-width: 1071px) 543px, 580px\" \/><\/figure>\n\n\n\n<ul class=\"wp-block-list\">\n<li>Aligning with gravity and camera view direction<\/li>\n\n\n\n<li>GV Coordinate system eliminates inconsistencies between coordinate systems of different frames.<\/li>\n<\/ul>\n\n\n\n<figure class=\"wp-block-image size-large\"><img loading=\"lazy\" decoding=\"async\" width=\"1024\" height=\"276\" src=\"https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/wp-content\/uploads\/sites\/121\/2025\/05\/Snipaste_2025-05-09_23-18-43-1024x276.jpg\" alt=\"\" class=\"wp-image-44\" srcset=\"https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/wp-content\/uploads\/sites\/121\/2025\/05\/Snipaste_2025-05-09_23-18-43-1024x276.jpg 1024w, https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/wp-content\/uploads\/sites\/121\/2025\/05\/Snipaste_2025-05-09_23-18-43-300x81.jpg 300w, https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/wp-content\/uploads\/sites\/121\/2025\/05\/Snipaste_2025-05-09_23-18-43-768x207.jpg 768w, https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/wp-content\/uploads\/sites\/121\/2025\/05\/Snipaste_2025-05-09_23-18-43-1536x414.jpg 1536w, https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/wp-content\/uploads\/sites\/121\/2025\/05\/Snipaste_2025-05-09_23-18-43.jpg 1758w\" sizes=\"auto, (max-width: 767px) 89vw, (max-width: 1000px) 54vw, (max-width: 1071px) 543px, 580px\" \/><\/figure>\n\n\n\n<h2 class=\"wp-block-heading\"><strong>3. CameraHMR: Aligning People with Perspective.<\/strong><\/h2>\n\n\n\n<p><strong>Key Insight<\/strong>: Integrate the predicted camera field of view (FoV) into the reconstruction pipeline, improves HMR in monocular images with severe perspective distortion.<\/p>\n\n\n\n<figure class=\"wp-block-image size-large\"><img loading=\"lazy\" decoding=\"async\" width=\"1024\" height=\"632\" src=\"https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/wp-content\/uploads\/sites\/121\/2025\/05\/\u622a\u5c4f2025-05-09-17.23.06-1024x632.png\" alt=\"\" class=\"wp-image-47\" srcset=\"https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/wp-content\/uploads\/sites\/121\/2025\/05\/\u622a\u5c4f2025-05-09-17.23.06-1024x632.png 1024w, https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/wp-content\/uploads\/sites\/121\/2025\/05\/\u622a\u5c4f2025-05-09-17.23.06-300x185.png 300w, https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/wp-content\/uploads\/sites\/121\/2025\/05\/\u622a\u5c4f2025-05-09-17.23.06-768x474.png 768w, https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/wp-content\/uploads\/sites\/121\/2025\/05\/\u622a\u5c4f2025-05-09-17.23.06.png 1265w\" sizes=\"auto, (max-width: 767px) 89vw, (max-width: 1000px) 54vw, (max-width: 1071px) 543px, 580px\" \/><\/figure>\n\n\n\n<ul class=\"wp-block-list\">\n<li><strong>HumanFoV<\/strong>: Predicts the FoV directly from the input image.<\/li>\n\n\n\n<li><strong>CamSMPLify<\/strong>: Incorporates the predicted FoV into a full perspective camera model, replacing the traditional weak-perspective assumption.<\/li>\n\n\n\n<li><strong>CameraHMR<\/strong>: Improves the original HMR2.0 architecture by integrating the camera intrinsics predicted by HumanFoV.<\/li>\n<\/ul>\n\n\n\n<p class=\"has-medium-font-size\">[References]<\/p>\n\n\n\n<ol class=\"wp-block-list\">\n<li class=\"has-small-font-size\">You, Yingxuan, et al. &#8220;Co-evolution of pose and mesh for 3d human body estimation from video.&#8221; Proceedings of the IEEE\/CVF International Conference on Computer Vision. 2023.<\/li>\n\n\n\n<li class=\"has-small-font-size\">Shen, Zehong, et al. &#8220;World-Grounded Human Motion Recovery via Gravity-View Coordinates.&#8221; SIGGRAPH Asia 2024 Conference Papers. 2024.<\/li>\n\n\n\n<li class=\"has-small-font-size\">Patel, Priyanka, and Michael J. Black. &#8220;CameraHMR: Aligning People with Perspective.&#8221; arXiv preprint arXiv:2411.08128 (2024).<\/li>\n<\/ol>\n","protected":false},"excerpt":{"rendered":"<p>1. Co-evolution of pose and mesh for 3d human body estimation from video. Key Insight: Human mesh recovery is not just about using 3D keypoints (Pose) or 3D shapes (Mesh) alone\u2014it\u2019s about joint collaboration between them. Poses provide information about human motion; meshes provide information about body shape 2. World-Grounded Human Motion Recovery via Gravity-View &hellip; <\/p>\n<p class=\"link-more\"><a href=\"https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/related-work\/\" class=\"more-link\">Continue reading<span class=\"screen-reader-text\"> &#8220;Related Work&#8221;<\/span><\/a><\/p>\n","protected":false},"author":235,"featured_media":0,"parent":0,"menu_order":0,"comment_status":"closed","ping_status":"closed","template":"","meta":{"footnotes":""},"class_list":["post-14","page","type-page","status-publish","hentry"],"yoast_head":"<!-- This site is optimized with the Yoast SEO plugin v28.1 - https:\/\/yoast.com\/product\/yoast-seo-wordpress\/ -->\n<title>Related Work - World-Grounded Human Mesh Recovery from Video<\/title>\n<meta name=\"robots\" content=\"index, follow, max-snippet:-1, max-image-preview:large, max-video-preview:-1\" \/>\n<link rel=\"canonical\" href=\"https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/related-work\/\" \/>\n<meta property=\"og:locale\" content=\"en_US\" \/>\n<meta property=\"og:type\" content=\"article\" \/>\n<meta property=\"og:title\" content=\"Related Work - World-Grounded Human Mesh Recovery from Video\" \/>\n<meta property=\"og:description\" content=\"1. Co-evolution of pose and mesh for 3d human body estimation from video. Key Insight: Human mesh recovery is not just about using 3D keypoints (Pose) or 3D shapes (Mesh) alone\u2014it\u2019s about joint collaboration between them. Poses provide information about human motion; meshes provide information about body shape 2. World-Grounded Human Motion Recovery via Gravity-View &hellip; Continue reading &quot;Related Work&quot;\" \/>\n<meta property=\"og:url\" content=\"https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/related-work\/\" \/>\n<meta property=\"og:site_name\" content=\"World-Grounded Human Mesh Recovery from Video\" \/>\n<meta property=\"article:modified_time\" content=\"2025-05-10T03:22:08+00:00\" \/>\n<meta property=\"og:image\" content=\"https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/wp-content\/uploads\/sites\/121\/2025\/05\/dmdmd-1024x515.jpg\" \/>\n<meta name=\"twitter:card\" content=\"summary_large_image\" \/>\n<meta name=\"twitter:label1\" content=\"Est. reading time\" \/>\n\t<meta name=\"twitter:data1\" content=\"1 minute\" \/>\n<script type=\"application\/ld+json\" class=\"yoast-schema-graph\">{\"@context\":\"https:\\\/\\\/schema.org\",\"@graph\":[{\"@type\":\"WebPage\",\"@id\":\"https:\\\/\\\/mscvprojects.ri.cmu.edu\\\/2025team9-1\\\/related-work\\\/\",\"url\":\"https:\\\/\\\/mscvprojects.ri.cmu.edu\\\/2025team9-1\\\/related-work\\\/\",\"name\":\"Related Work - World-Grounded Human Mesh Recovery from Video\",\"isPartOf\":{\"@id\":\"https:\\\/\\\/mscvprojects.ri.cmu.edu\\\/2025team9-1\\\/#website\"},\"primaryImageOfPage\":{\"@id\":\"https:\\\/\\\/mscvprojects.ri.cmu.edu\\\/2025team9-1\\\/related-work\\\/#primaryimage\"},\"image\":{\"@id\":\"https:\\\/\\\/mscvprojects.ri.cmu.edu\\\/2025team9-1\\\/related-work\\\/#primaryimage\"},\"thumbnailUrl\":\"https:\\\/\\\/mscvprojects.ri.cmu.edu\\\/2025team9-1\\\/wp-content\\\/uploads\\\/sites\\\/121\\\/2025\\\/05\\\/dmdmd-1024x515.jpg\",\"datePublished\":\"2025-05-10T02:55:53+00:00\",\"dateModified\":\"2025-05-10T03:22:08+00:00\",\"breadcrumb\":{\"@id\":\"https:\\\/\\\/mscvprojects.ri.cmu.edu\\\/2025team9-1\\\/related-work\\\/#breadcrumb\"},\"inLanguage\":\"en-US\",\"potentialAction\":[{\"@type\":\"ReadAction\",\"target\":[\"https:\\\/\\\/mscvprojects.ri.cmu.edu\\\/2025team9-1\\\/related-work\\\/\"]}]},{\"@type\":\"ImageObject\",\"inLanguage\":\"en-US\",\"@id\":\"https:\\\/\\\/mscvprojects.ri.cmu.edu\\\/2025team9-1\\\/related-work\\\/#primaryimage\",\"url\":\"https:\\\/\\\/mscvprojects.ri.cmu.edu\\\/2025team9-1\\\/wp-content\\\/uploads\\\/sites\\\/121\\\/2025\\\/05\\\/dmdmd.jpg\",\"contentUrl\":\"https:\\\/\\\/mscvprojects.ri.cmu.edu\\\/2025team9-1\\\/wp-content\\\/uploads\\\/sites\\\/121\\\/2025\\\/05\\\/dmdmd.jpg\",\"width\":1552,\"height\":780},{\"@type\":\"BreadcrumbList\",\"@id\":\"https:\\\/\\\/mscvprojects.ri.cmu.edu\\\/2025team9-1\\\/related-work\\\/#breadcrumb\",\"itemListElement\":[{\"@type\":\"ListItem\",\"position\":1,\"name\":\"Home\",\"item\":\"https:\\\/\\\/mscvprojects.ri.cmu.edu\\\/2025team9-1\\\/\"},{\"@type\":\"ListItem\",\"position\":2,\"name\":\"Related Work\"}]},{\"@type\":\"WebSite\",\"@id\":\"https:\\\/\\\/mscvprojects.ri.cmu.edu\\\/2025team9-1\\\/#website\",\"url\":\"https:\\\/\\\/mscvprojects.ri.cmu.edu\\\/2025team9-1\\\/\",\"name\":\"World-Grounded Human Mesh Recovery from Video\",\"description\":\"Liting Wen, Yiwen Zhao, Ce Zheng, L\u00e1szl\u00f3 A. Jeni \",\"potentialAction\":[{\"@type\":\"SearchAction\",\"target\":{\"@type\":\"EntryPoint\",\"urlTemplate\":\"https:\\\/\\\/mscvprojects.ri.cmu.edu\\\/2025team9-1\\\/?s={search_term_string}\"},\"query-input\":{\"@type\":\"PropertyValueSpecification\",\"valueRequired\":true,\"valueName\":\"search_term_string\"}}],\"inLanguage\":\"en-US\"}]}<\/script>\n<!-- \/ Yoast SEO plugin. -->","yoast_head_json":{"title":"Related Work - World-Grounded Human Mesh Recovery from Video","robots":{"index":"index","follow":"follow","max-snippet":"max-snippet:-1","max-image-preview":"max-image-preview:large","max-video-preview":"max-video-preview:-1"},"canonical":"https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/related-work\/","og_locale":"en_US","og_type":"article","og_title":"Related Work - World-Grounded Human Mesh Recovery from Video","og_description":"1. Co-evolution of pose and mesh for 3d human body estimation from video. Key Insight: Human mesh recovery is not just about using 3D keypoints (Pose) or 3D shapes (Mesh) alone\u2014it\u2019s about joint collaboration between them. Poses provide information about human motion; meshes provide information about body shape 2. World-Grounded Human Motion Recovery via Gravity-View &hellip; Continue reading \"Related Work\"","og_url":"https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/related-work\/","og_site_name":"World-Grounded Human Mesh Recovery from Video","article_modified_time":"2025-05-10T03:22:08+00:00","og_image":[{"url":"https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/wp-content\/uploads\/sites\/121\/2025\/05\/dmdmd-1024x515.jpg","type":"","width":"","height":""}],"twitter_card":"summary_large_image","twitter_misc":{"Est. reading time":"1 minute"},"schema":{"@context":"https:\/\/schema.org","@graph":[{"@type":"WebPage","@id":"https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/related-work\/","url":"https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/related-work\/","name":"Related Work - World-Grounded Human Mesh Recovery from Video","isPartOf":{"@id":"https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/#website"},"primaryImageOfPage":{"@id":"https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/related-work\/#primaryimage"},"image":{"@id":"https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/related-work\/#primaryimage"},"thumbnailUrl":"https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/wp-content\/uploads\/sites\/121\/2025\/05\/dmdmd-1024x515.jpg","datePublished":"2025-05-10T02:55:53+00:00","dateModified":"2025-05-10T03:22:08+00:00","breadcrumb":{"@id":"https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/related-work\/#breadcrumb"},"inLanguage":"en-US","potentialAction":[{"@type":"ReadAction","target":["https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/related-work\/"]}]},{"@type":"ImageObject","inLanguage":"en-US","@id":"https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/related-work\/#primaryimage","url":"https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/wp-content\/uploads\/sites\/121\/2025\/05\/dmdmd.jpg","contentUrl":"https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/wp-content\/uploads\/sites\/121\/2025\/05\/dmdmd.jpg","width":1552,"height":780},{"@type":"BreadcrumbList","@id":"https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/related-work\/#breadcrumb","itemListElement":[{"@type":"ListItem","position":1,"name":"Home","item":"https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/"},{"@type":"ListItem","position":2,"name":"Related Work"}]},{"@type":"WebSite","@id":"https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/#website","url":"https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/","name":"World-Grounded Human Mesh Recovery from Video","description":"Liting Wen, Yiwen Zhao, Ce Zheng, L\u00e1szl\u00f3 A. Jeni ","potentialAction":[{"@type":"SearchAction","target":{"@type":"EntryPoint","urlTemplate":"https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/?s={search_term_string}"},"query-input":{"@type":"PropertyValueSpecification","valueRequired":true,"valueName":"search_term_string"}}],"inLanguage":"en-US"}]}},"_links":{"self":[{"href":"https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/wp-json\/wp\/v2\/pages\/14","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/wp-json\/wp\/v2\/pages"}],"about":[{"href":"https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/wp-json\/wp\/v2\/types\/page"}],"author":[{"embeddable":true,"href":"https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/wp-json\/wp\/v2\/users\/235"}],"replies":[{"embeddable":true,"href":"https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/wp-json\/wp\/v2\/comments?post=14"}],"version-history":[{"count":7,"href":"https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/wp-json\/wp\/v2\/pages\/14\/revisions"}],"predecessor-version":[{"id":50,"href":"https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/wp-json\/wp\/v2\/pages\/14\/revisions\/50"}],"wp:attachment":[{"href":"https:\/\/mscvprojects.ri.cmu.edu\/2025team9-1\/wp-json\/wp\/v2\/media?parent=14"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}