{"id":30150,"date":"2026-08-11T11:28:51","date_gmt":"2026-08-11T11:28:51","guid":{"rendered":"https:\/\/framos.com\/?post_type=articles&#038;p=30150"},"modified":"2026-08-11T14:50:48","modified_gmt":"2026-08-11T14:50:48","slug":"object-centric-perception-for-robotics","status":"publish","type":"articles","link":"https:\/\/framos.com\/de\/fachartikel\/object-centric-perception-for-robotics\/","title":{"rendered":"When Detection Isn&#8217;t Enough: Object-centric Perception for Robotic Manipulation"},"content":{"rendered":"\n<p class=\"wp-block-paragraph\">In a benchmark, perception ends with a label and a bounding box. On a robot, that is where the problem starts. To manipulate an object, a system needs to know not just what something is, but where it is in space, how it is oriented, and whether it is partially hidden &#8211; and it needs to keep that estimate stable while the robot, the object, and the scene all move.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">At ImagingNext 2026, <strong>Maximilian Durner<\/strong>, Research Group Leader in the Perception and Cognition Department at the <a href=\"https:\/\/www.dlr.de\/en\/rm\" target=\"_blank\" rel=\"noopener\"><strong>German Aerospace Center (DLR<\/strong>)<\/a>, presents \u201c<strong>Object-centric Perception for Robotic Manipulation<\/strong>\u201d &#8211; a session on what it takes for robotic perception to hold up outside structured environments.<\/p>\n\n\n\n<div class=\"wp-block-media-text is-stacked-on-mobile\"><figure class=\"wp-block-media-text__media\"><img loading=\"lazy\" decoding=\"async\" width=\"1080\" height=\"1080\" src=\"https:\/\/framos.com\/wp-content\/uploads\/2026\/03\/durner.jpg\" alt=\"Maximilian Durner\" class=\"wp-image-28161 size-full\" srcset=\"https:\/\/framos.com\/wp-content\/uploads\/2026\/03\/durner.jpg 1080w, https:\/\/framos.com\/wp-content\/uploads\/2026\/03\/durner-300x300.jpg 300w, https:\/\/framos.com\/wp-content\/uploads\/2026\/03\/durner-150x150.jpg 150w, https:\/\/framos.com\/wp-content\/uploads\/2026\/03\/durner-64x64.jpg 64w\" sizes=\"auto, (max-width: 1080px) 100vw, 1080px\" \/><\/figure><div class=\"wp-block-media-text__content\">\n<h2 class=\"wp-block-heading\">About Maximilian Durner<\/h2>\n\n\n\n<p class=\"has-text-align-left wp-block-paragraph\"><a href=\"https:\/\/www.linkedin.com\/in\/maximilian-durner-a44489112\/?locale=en\" target=\"_blank\" rel=\"noopener\"><strong>Maximilian Durner<\/strong><\/a> is a Research Fellow and Research Group Leader in the Perception and Cognition Department at the Institute of Robotics and Mechatronics of the German Aerospace Center (DLR), where he has been working since 2016. He received his Ph.D. from the Technical University of Munich (TUM). His research group develops methods for robust robotic perception and semantic scene understanding for complex, real-world environments.<\/p>\n<\/div><\/div>\n\n\n\n<h3 class=\"wp-block-heading has-large-font-size\"><strong>Why this matters now<\/strong><\/h3>\n\n\n\n<p class=\"wp-block-paragraph\">Robotic manipulation is moving out of the structured cell and into environments that were never arranged for robots &#8211; logistics, service, and lab settings with open object sets, clutter, occlusion, and changing conditions. In those environments, perception is usually the limiting factor. The object-centric approach treats the object, rather than the image frame, as the unit of understanding: detection, recognition of novel objects, pose estimation, and tracking feed one coherent representation that a manipulation planner can actually act on.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">Durner&#8217;s group at DLR&#8217;s Institute of Robotics and Mechatronics develops methods for robust robotic perception and semantic scene understanding, enabling robots to operate reliably in complex, real-world environments. His research focuses on object-centric vision systems for mobile manipulation, with an emphasis on robustness, continual adaptation, and long-term autonomy &#8211; integrating physical priors, world knowledge, and multimodal sensory information to improve capabilities such as object detection, recognition of novel objects, pose estimation, and tracking.<\/p>\n\n\n\n<h3 class=\"wp-block-heading has-large-font-size\"><strong>What you&#8217;ll take away<\/strong><\/h3>\n\n\n\n<ul class=\"wp-block-list\">\n<li>What \u201cobject-centric\u201d perception means in practice, and how it differs from frame-level detection pipelines.<\/li>\n\n\n\n<li>How physical priors, world knowledge, and multimodal sensing are combined to make detection, pose estimation, and tracking robust.<\/li>\n\n\n\n<li>Where perception still breaks in real-world manipulation &#8211; and what robustness and continual adaptation demand from the vision system.<\/li>\n<\/ul>\n\n\n<div class=\"wp-block-framos-cta\">\n    <div class=\"mx-auto max-w-6xl py-12 px-4 md:px-6 frontend\">\n        <div class=\"relative isolate overflow-hidden px-6 py-24 text-center shadow-2xl rounded-3xl sm:px-16 bg-framosDarkBlue\">\n                            <h2 class=\"font-framos-light text-3xl lg:text-4xl leading-tight mb-6 text-white mx-auto max-w-2xl\">\n                    When Detection Isn&#8217;t Enough: Object-centric Perception for Robotic Manipulation                <\/h2>\n            \n                            <p class=\"text-base leading-7 mb-4 text-gray-100 mx-auto mt-6 max-w-xl\">\n                    Maximilian&#8217;s session is one of the talks at ImagingNext 2026 &#8211; two days on end-to-end Vision AI systems, edge deployment, and honest engineering exchange. October 14\u201315, smartvillage Bogenhausen, Munich                <\/p>\n            \n            <div class=\"mt-10 flex flex-col items-center justify-center gap-y-8 sm:flex-row sm:gap-x-6\">\n                <a href=\"https:\/\/framos.com\/events\/imaging-next-2026\/#tickets\" target=\"_blank\" rel=\"noopener noreferrer\" class=\"btn-round-arrow group visited:text-white\" data-cta=\"true\" data-cta-id=\"framos_cta.1.primary_button\" data-cta-label=\"get your ticket\" data-cta-location=\"framos_cta@object-centric-perception-for-robotics\" data-cta-destination=\"https:\/\/framos.com\/events\/imaging-next-2026\/#tickets\" data-cta-type=\"primary\">\n                    <span>Get Your Ticket<\/span>\n                    <div class=\"ml-1 -rotate-45 transition-all duration-200 group-hover:rotate-0\">\n                        <svg\n                            width=\"15\"\n                            height=\"15\"\n                            viewBox=\"0 0 15 15\"\n                            fill=\"none\"\n                            xmlns=\"http:\/\/www.w3.org\/2000\/svg\"\n                            class=\"h-5 w-5\"\n                        >\n                            <path\n                                d=\"M8.14645 3.14645C8.34171 2.95118 8.65829 2.95118 8.85355 3.14645L12.8536 7.14645C13.0488 7.34171 13.0488 7.65829 12.8536 7.85355L8.85355 11.8536C8.65829 12.0488 8.34171 12.0488 8.14645 11.8536C7.95118 11.6583 7.95118 11.3417 8.14645 11.1464L11.2929 8H2.5C2.22386 8 2 7.77614 2 7.5C2 7.22386 2.22386 7 2.5 7H11.2929L8.14645 3.85355C7.95118 3.65829 7.95118 3.34171 8.14645 3.14645Z\"\n                                fill=\"currentColor\"\n                                fill-rule=\"evenodd\"\n                                clip-rule=\"evenodd\"\n                            \/>\n                        <\/svg>\n                    <\/div>\n                <\/a>\n\n                                    <a href=\"https:\/\/framos.com\/events\/imaging-next-2026\/#agenda\" target=\"_blank\" rel=\"noopener noreferrer\" class=\"text-sm font-semibold leading-6 text-white hover:text-primary visited:text-white\" data-cta=\"true\" data-cta-id=\"framos_cta.2.secondary_button\" data-cta-label=\"see the full speaker lineup\" data-cta-location=\"framos_cta@object-centric-perception-for-robotics\" data-cta-destination=\"https:\/\/framos.com\/events\/imaging-next-2026\/#agenda\" data-cta-type=\"secondary\">\n                        <span>See the Full Speaker Lineup<\/span>\n                        <span aria-hidden=\"true\">\u2192<\/span>\n                    <\/a>\n                \n            <\/div>\n\n            <svg\n                viewBox=\"0 0 1024 1024\"\n                class=\"absolute left-1\/2 top-1\/2 -z-10 h-[64rem] w-[64rem] -translate-x-1\/2 [mask-image:radial-gradient(closest-side,white,transparent)]\"\n                aria-hidden=\"true\"\n            >\n                                <circle cx=\"512\" cy=\"512\" r=\"512\" fill=\"url(#framos-cta-gradient-1)\" fill-opacity=\"0.7\" \/>\n                <defs>\n                    <radialGradient id=\"framos-cta-gradient-1\">\n                        <stop stop-color=\"#00c4d4\" \/>\n                        <stop offset=\"1\" stop-color=\"#0098b0\" \/>\n                    <\/radialGradient>\n                <\/defs>\n            <\/svg>\n        <\/div>\n    <\/div>\n<\/div>\n","protected":false},"excerpt":{"rendered":"<p>In a benchmark, perception ends with a label and a bounding box. On a robot, that is where the problem starts. To manipulate an object, a system needs to know [&hellip;]<\/p>\n","protected":false},"author":2,"featured_media":30159,"comment_status":"closed","ping_status":"closed","template":"","format":"standard","meta":{"_acf_changed":false,"inline_featured_image":false},"article-category":[],"article-tag":[],"class_list":["post-30150","articles","type-articles","status-publish","format-standard","has-post-thumbnail","hentry"],"acf":[],"_links":{"self":[{"href":"https:\/\/framos.com\/de\/wp-json\/wp\/v2\/articles\/30150","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/framos.com\/de\/wp-json\/wp\/v2\/articles"}],"about":[{"href":"https:\/\/framos.com\/de\/wp-json\/wp\/v2\/types\/articles"}],"author":[{"embeddable":true,"href":"https:\/\/framos.com\/de\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/framos.com\/de\/wp-json\/wp\/v2\/comments?post=30150"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/framos.com\/de\/wp-json\/wp\/v2\/media\/30159"}],"wp:attachment":[{"href":"https:\/\/framos.com\/de\/wp-json\/wp\/v2\/media?parent=30150"}],"wp:term":[{"taxonomy":"article-category","embeddable":true,"href":"https:\/\/framos.com\/de\/wp-json\/wp\/v2\/article-category?post=30150"},{"taxonomy":"article-tag","embeddable":true,"href":"https:\/\/framos.com\/de\/wp-json\/wp\/v2\/article-tag?post=30150"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}