{
    "componentChunkName": "component---src-templates-blogpost-js",
    "path": "/blog/boost-real-time-3d-object-detection-accuracy-with-smarter-ai-sequence-analysis-for-robotics-and-automotive-216523",
    "result": {"data":{"OTHER_POSTS":{"nodes":[{"title":"Universal Mobile Payment System Enables Secure Transactions Across Any App or Platform","slug":"universal-mobile-payment-system-enables-secure-transactions-across-any-app-or-platform-918414","date":"2025-12-11T06:10:00","seo":{"metaDesc":""},"tags":{"nodes":[{"name":"Alphabet"},{"name":"Patent Review"}]},"featuredImage":{"node":{"mediaItemUrl":"https://wp.inventiv.org/wp-content/uploads/2025/12/generated_image_20251211_130602_c0df7ae0.png","localFile":{"childImageSharp":{"gatsbyImageData":{"layout":"constrained","backgroundColor":"#5838a8","images":{"fallback":{"src":"/static/66ed8abed7803034c70f55827491b36f/b5658/generated_image_20251211_130602_c0df7ae0.png","srcSet":"/static/66ed8abed7803034c70f55827491b36f/acb7c/generated_image_20251211_130602_c0df7ae0.png 256w,\n/static/66ed8abed7803034c70f55827491b36f/ccc41/generated_image_20251211_130602_c0df7ae0.png 512w,\n/static/66ed8abed7803034c70f55827491b36f/b5658/generated_image_20251211_130602_c0df7ae0.png 1024w","sizes":"(min-width: 1024px) 1024px, 100vw"},"sources":[{"srcSet":"/static/66ed8abed7803034c70f55827491b36f/22bfc/generated_image_20251211_130602_c0df7ae0.webp 256w,\n/static/66ed8abed7803034c70f55827491b36f/d689f/generated_image_20251211_130602_c0df7ae0.webp 512w,\n/static/66ed8abed7803034c70f55827491b36f/67ded/generated_image_20251211_130602_c0df7ae0.webp 1024w","type":"image/webp","sizes":"(min-width: 1024px) 1024px, 100vw"}]},"width":1024,"height":1024}}}}}},{"title":"Unlock Instant Device Control with Ultra-Fast Communication for Smarter Industrial and IoT Systems","slug":"unlock-instant-device-control-with-ultra-fast-communication-for-smarter-industrial-and-iot-systems-651401","date":"2025-12-01T04:46:30","seo":{"metaDesc":""},"tags":{"nodes":[{"name":"Amazon"},{"name":"Patent Review"}]},"featuredImage":{"node":{"mediaItemUrl":"https://wp.inventiv.org/wp-content/uploads/2025/12/generated_image_20251201_114346_dbfb2624.png","localFile":{"childImageSharp":{"gatsbyImageData":{"layout":"constrained","backgroundColor":"#081858","images":{"fallback":{"src":"/static/28ca2c688620d650900bab71d678b91a/b5658/generated_image_20251201_114346_dbfb2624.png","srcSet":"/static/28ca2c688620d650900bab71d678b91a/acb7c/generated_image_20251201_114346_dbfb2624.png 256w,\n/static/28ca2c688620d650900bab71d678b91a/ccc41/generated_image_20251201_114346_dbfb2624.png 512w,\n/static/28ca2c688620d650900bab71d678b91a/b5658/generated_image_20251201_114346_dbfb2624.png 1024w","sizes":"(min-width: 1024px) 1024px, 100vw"},"sources":[{"srcSet":"/static/28ca2c688620d650900bab71d678b91a/22bfc/generated_image_20251201_114346_dbfb2624.webp 256w,\n/static/28ca2c688620d650900bab71d678b91a/d689f/generated_image_20251201_114346_dbfb2624.webp 512w,\n/static/28ca2c688620d650900bab71d678b91a/67ded/generated_image_20251201_114346_dbfb2624.webp 1024w","type":"image/webp","sizes":"(min-width: 1024px) 1024px, 100vw"}]},"width":1024,"height":1024}}}}}},{"title":"Centralized Platform Streamlines AI Agent Registration and Management for Enterprises","slug":"centralized-platform-streamlines-ai-agent-registration-and-management-for-enterprises-216165","date":"2025-12-28T10:44:03","seo":{"metaDesc":""},"tags":{"nodes":[{"name":"Amazon"},{"name":"Patent Review"}]},"featuredImage":{"node":{"mediaItemUrl":"https://wp.inventiv.org/wp-content/uploads/2025/12/generated_image_20251228_174003_f9761009.png","localFile":{"childImageSharp":{"gatsbyImageData":{"layout":"constrained","backgroundColor":"#082858","images":{"fallback":{"src":"/static/72d0b6000657868892318cf34420c4b4/b5658/generated_image_20251228_174003_f9761009.png","srcSet":"/static/72d0b6000657868892318cf34420c4b4/acb7c/generated_image_20251228_174003_f9761009.png 256w,\n/static/72d0b6000657868892318cf34420c4b4/ccc41/generated_image_20251228_174003_f9761009.png 512w,\n/static/72d0b6000657868892318cf34420c4b4/b5658/generated_image_20251228_174003_f9761009.png 1024w","sizes":"(min-width: 1024px) 1024px, 100vw"},"sources":[{"srcSet":"/static/72d0b6000657868892318cf34420c4b4/22bfc/generated_image_20251228_174003_f9761009.webp 256w,\n/static/72d0b6000657868892318cf34420c4b4/d689f/generated_image_20251228_174003_f9761009.webp 512w,\n/static/72d0b6000657868892318cf34420c4b4/67ded/generated_image_20251228_174003_f9761009.webp 1024w","type":"image/webp","sizes":"(min-width: 1024px) 1024px, 100vw"}]},"width":1024,"height":1024}}}}}},{"title":"Unlock Targeted Crop Enhancement with Fruit-Specific Gene Promoters for Superior Produce Quality","slug":"unlock-targeted-crop-enhancement-with-fruit-specific-gene-promoters-for-superior-produce-quality-209251","date":"2026-01-09T07:04:56","seo":{"metaDesc":""},"tags":{"nodes":[{"name":"Patent Review"}]},"featuredImage":{"node":{"mediaItemUrl":"https://wp.inventiv.org/wp-content/uploads/2026/01/generated_image_20260109_140158_ea04ad35.png","localFile":{"childImageSharp":{"gatsbyImageData":{"layout":"constrained","backgroundColor":"#d85818","images":{"fallback":{"src":"/static/5a33b72200eced76c64e2767617f1676/b5658/generated_image_20260109_140158_ea04ad35.png","srcSet":"/static/5a33b72200eced76c64e2767617f1676/acb7c/generated_image_20260109_140158_ea04ad35.png 256w,\n/static/5a33b72200eced76c64e2767617f1676/ccc41/generated_image_20260109_140158_ea04ad35.png 512w,\n/static/5a33b72200eced76c64e2767617f1676/b5658/generated_image_20260109_140158_ea04ad35.png 1024w","sizes":"(min-width: 1024px) 1024px, 100vw"},"sources":[{"srcSet":"/static/5a33b72200eced76c64e2767617f1676/22bfc/generated_image_20260109_140158_ea04ad35.webp 256w,\n/static/5a33b72200eced76c64e2767617f1676/d689f/generated_image_20260109_140158_ea04ad35.webp 512w,\n/static/5a33b72200eced76c64e2767617f1676/67ded/generated_image_20260109_140158_ea04ad35.webp 1024w","type":"image/webp","sizes":"(min-width: 1024px) 1024px, 100vw"}]},"width":1024,"height":1024}}}}}},{"title":"Unlocking Advanced Battery Performance with Next-Generation Composite Substrates for Electrode Manufacturing","slug":"unlocking-advanced-battery-performance-with-next-generation-composite-substrates-for-electrode-manufacturing-943519","date":"2026-01-20T08:45:01","seo":{"metaDesc":""},"tags":{"nodes":[{"name":"Patent Review"},{"name":"Samsung"}]},"featuredImage":{"node":{"mediaItemUrl":"https://wp.inventiv.org/wp-content/uploads/2026/01/generated_image_20260120_154207_b9a4de4f.png","localFile":{"childImageSharp":{"gatsbyImageData":{"layout":"constrained","backgroundColor":"#185888","images":{"fallback":{"src":"/static/39ca38ef88b8253d6cf6e4b7eaacfec8/b5658/generated_image_20260120_154207_b9a4de4f.png","srcSet":"/static/39ca38ef88b8253d6cf6e4b7eaacfec8/acb7c/generated_image_20260120_154207_b9a4de4f.png 256w,\n/static/39ca38ef88b8253d6cf6e4b7eaacfec8/ccc41/generated_image_20260120_154207_b9a4de4f.png 512w,\n/static/39ca38ef88b8253d6cf6e4b7eaacfec8/b5658/generated_image_20260120_154207_b9a4de4f.png 1024w","sizes":"(min-width: 1024px) 1024px, 100vw"},"sources":[{"srcSet":"/static/39ca38ef88b8253d6cf6e4b7eaacfec8/22bfc/generated_image_20260120_154207_b9a4de4f.webp 256w,\n/static/39ca38ef88b8253d6cf6e4b7eaacfec8/d689f/generated_image_20260120_154207_b9a4de4f.webp 512w,\n/static/39ca38ef88b8253d6cf6e4b7eaacfec8/67ded/generated_image_20260120_154207_b9a4de4f.webp 1024w","type":"image/webp","sizes":"(min-width: 1024px) 1024px, 100vw"}]},"width":1024,"height":1024}}}}}},{"title":"Unlocking Smarter AI: A Unified Model for Fast, Efficient Multimodal Learning Across Industries","slug":"unlocking-smarter-ai-a-unified-model-for-fast-efficient-multimodal-learning-across-industries-914257","date":"2026-02-22T05:13:25","seo":{"metaDesc":""},"tags":{"nodes":[{"name":"Facebook/Meta"},{"name":"Patent Review"}]},"featuredImage":{"node":{"mediaItemUrl":"https://wp.inventiv.org/wp-content/uploads/2026/02/generated_image_20260222_121057_050755f8.png","localFile":{"childImageSharp":{"gatsbyImageData":{"layout":"constrained","backgroundColor":"#081838","images":{"fallback":{"src":"/static/f7b9939c8e987b504acbc2ab10554aac/b5658/generated_image_20260222_121057_050755f8.png","srcSet":"/static/f7b9939c8e987b504acbc2ab10554aac/acb7c/generated_image_20260222_121057_050755f8.png 256w,\n/static/f7b9939c8e987b504acbc2ab10554aac/ccc41/generated_image_20260222_121057_050755f8.png 512w,\n/static/f7b9939c8e987b504acbc2ab10554aac/b5658/generated_image_20260222_121057_050755f8.png 1024w","sizes":"(min-width: 1024px) 1024px, 100vw"},"sources":[{"srcSet":"/static/f7b9939c8e987b504acbc2ab10554aac/22bfc/generated_image_20260222_121057_050755f8.webp 256w,\n/static/f7b9939c8e987b504acbc2ab10554aac/d689f/generated_image_20260222_121057_050755f8.webp 512w,\n/static/f7b9939c8e987b504acbc2ab10554aac/67ded/generated_image_20260222_121057_050755f8.webp 1024w","type":"image/webp","sizes":"(min-width: 1024px) 1024px, 100vw"}]},"width":1024,"height":1024}}}}}},{"title":"Unlocking Ultra-Efficient 3D Data Compression for Faster, Smarter AR and Digital Twins","slug":"unlocking-ultra-efficient-3d-data-compression-for-faster-smarter-ar-and-digital-twins-647469","date":"2025-11-09T10:43:36","seo":{"metaDesc":""},"tags":{"nodes":[{"name":"Patent Review"}]},"featuredImage":{"node":{"mediaItemUrl":"https://wp.inventiv.org/wp-content/uploads/2025/11/generated_image_20251109_174053_02379c93.png","localFile":{"childImageSharp":{"gatsbyImageData":{"layout":"constrained","backgroundColor":"#081818","images":{"fallback":{"src":"/static/459152c7527917b800f35393e3362e8f/b5658/generated_image_20251109_174053_02379c93.png","srcSet":"/static/459152c7527917b800f35393e3362e8f/acb7c/generated_image_20251109_174053_02379c93.png 256w,\n/static/459152c7527917b800f35393e3362e8f/ccc41/generated_image_20251109_174053_02379c93.png 512w,\n/static/459152c7527917b800f35393e3362e8f/b5658/generated_image_20251109_174053_02379c93.png 1024w","sizes":"(min-width: 1024px) 1024px, 100vw"},"sources":[{"srcSet":"/static/459152c7527917b800f35393e3362e8f/22bfc/generated_image_20251109_174053_02379c93.webp 256w,\n/static/459152c7527917b800f35393e3362e8f/d689f/generated_image_20251109_174053_02379c93.webp 512w,\n/static/459152c7527917b800f35393e3362e8f/67ded/generated_image_20251109_174053_02379c93.webp 1024w","type":"image/webp","sizes":"(min-width: 1024px) 1024px, 100vw"}]},"width":1024,"height":1024}}}}}},{"title":"UPLINK CONFIGURED GRANT OF FULL DUPLEX CAPABLE USER EQUIPMENT","slug":"uplink-configured-grant-of-full-duplex-capable-user-equipment-398996","date":"2025-07-08T02:27:17","seo":{"metaDesc":""},"tags":{"nodes":[{"name":"Patent Review"}]},"featuredImage":{"node":{"mediaItemUrl":"https://wp.inventiv.org/wp-content/uploads/2025/07/generated_image_20250708_092436_5de55f18.png","localFile":{"childImageSharp":{"gatsbyImageData":{"layout":"constrained","backgroundColor":"#f8f8e8","images":{"fallback":{"src":"/static/ee5946da998724be5b36e573721983f4/b5658/generated_image_20250708_092436_5de55f18.png","srcSet":"/static/ee5946da998724be5b36e573721983f4/acb7c/generated_image_20250708_092436_5de55f18.png 256w,\n/static/ee5946da998724be5b36e573721983f4/ccc41/generated_image_20250708_092436_5de55f18.png 512w,\n/static/ee5946da998724be5b36e573721983f4/b5658/generated_image_20250708_092436_5de55f18.png 1024w","sizes":"(min-width: 1024px) 1024px, 100vw"},"sources":[{"srcSet":"/static/ee5946da998724be5b36e573721983f4/22bfc/generated_image_20250708_092436_5de55f18.webp 256w,\n/static/ee5946da998724be5b36e573721983f4/d689f/generated_image_20250708_092436_5de55f18.webp 512w,\n/static/ee5946da998724be5b36e573721983f4/67ded/generated_image_20250708_092436_5de55f18.webp 1024w","type":"image/webp","sizes":"(min-width: 1024px) 1024px, 100vw"}]},"width":1024,"height":1024}}}}}},{"title":"VALIDATING VEHICLE SENSOR CALIBRATION","slug":"validating-vehicle-sensor-calibration-943049","date":"2025-08-18T07:22:54","seo":{"metaDesc":""},"tags":{"nodes":[{"name":"Amazon"},{"name":"Patent Review"}]},"featuredImage":{"node":{"mediaItemUrl":"https://wp.inventiv.org/wp-content/uploads/2025/08/generated_image_20250818_141939_2a50b42f.png","localFile":{"childImageSharp":{"gatsbyImageData":{"layout":"constrained","backgroundColor":"#283838","images":{"fallback":{"src":"/static/3c25007218b33c9b67379ddc374ba5f2/b5658/generated_image_20250818_141939_2a50b42f.png","srcSet":"/static/3c25007218b33c9b67379ddc374ba5f2/acb7c/generated_image_20250818_141939_2a50b42f.png 256w,\n/static/3c25007218b33c9b67379ddc374ba5f2/ccc41/generated_image_20250818_141939_2a50b42f.png 512w,\n/static/3c25007218b33c9b67379ddc374ba5f2/b5658/generated_image_20250818_141939_2a50b42f.png 1024w","sizes":"(min-width: 1024px) 1024px, 100vw"},"sources":[{"srcSet":"/static/3c25007218b33c9b67379ddc374ba5f2/22bfc/generated_image_20250818_141939_2a50b42f.webp 256w,\n/static/3c25007218b33c9b67379ddc374ba5f2/d689f/generated_image_20250818_141939_2a50b42f.webp 512w,\n/static/3c25007218b33c9b67379ddc374ba5f2/67ded/generated_image_20250818_141939_2a50b42f.webp 1024w","type":"image/webp","sizes":"(min-width: 1024px) 1024px, 100vw"}]},"width":1024,"height":1024}}}}}}]},"RANDOM_POSTS":{"nodes":[]},"THE_POST":{"seo":{"canonical":"","focuskw":"","metaDesc":"","metaKeywords":"","title":"Boost Real-Time 3D Object Detection Accuracy with Smarter AI Sequence Analysis for Robotics and Automotive - Inventiv.org","twitterTitle":"","twitterDescription":"","opengraphDescription":"Invented by CAHU; Arthur, MARCUSANU; Ana, Dassault Systèmes In this blog post, we’ll explore a new way to teach computers…","opengraphPublishedTime":"2025-12-28T11:02:20+00:00","opengraphModifiedTime":"2025-12-28T11:05:47+00:00","opengraphTitle":"Boost Real-Time 3D Object Detection Accuracy with Smarter AI Sequence Analysis for Robotics and Automotive - Inventiv.org","opengraphType":"article","opengraphImage":{"sourceUrl":"https://wp.inventiv.org/wp-content/uploads/2025/12/generated_image_20251228_175947_be0506f7.png"}},"date":"December 28, 2025","content":"<h2 class=\"abpost-category\" style=\"padding-left:50px;padding-top:15px\">Invented by CAHU; Arthur, MARCUSANU; Ana, Dassault Systèmes</h2>\n<p><p><img loading=\"lazy\" class=\"alignnone size-medium wp-image-56050 aligncenter\" src=\"https://wp.inventiv.org/wp-content/uploads/2025/12/generated_image_20251228_180106_8207ba26-768x768.png\" alt=\"\" width=\"768\" height=\"768\" data-mce-src=\"https://wp.inventiv.org/wp-content/uploads/2025/12/generated_image_20251228_180106_8207ba26-768x768.png\" data-mce-selected=\"1\" srcset=\"https://wp.inventiv.org/wp-content/uploads/2025/12/generated_image_20251228_180106_8207ba26-768x768.png 768w, https://wp.inventiv.org/wp-content/uploads/2025/12/generated_image_20251228_180106_8207ba26-300x300.png 300w, https://wp.inventiv.org/wp-content/uploads/2025/12/generated_image_20251228_180106_8207ba26-150x150.png 150w, https://wp.inventiv.org/wp-content/uploads/2025/12/generated_image_20251228_180106_8207ba26-800x800.png 800w, https://wp.inventiv.org/wp-content/uploads/2025/12/generated_image_20251228_180106_8207ba26.png 1024w\" sizes=\"(max-width: 768px) 100vw, 768px\" /></p>\n<p>In this blog post, we’ll explore a new way to teach computers to find objects in real-world 3D scenes. We’ll break down a recent patent application that introduces a smarter machine learning method for working with 3D point clouds. You’ll learn how this invention fits into today’s technology, how it builds on past ideas, and what makes it so special.\n</p>\n<div style=\"display: flex; justify-content: center; align-items: center; margin: 10px 0; width: 100%;\">\n<video width=\"640\" height=\"360\" controls style=\"max-width: 100%; height: auto; display: block; margin: 0px auto;\"><source src=\"https://wp.inventiv.org/wp-content/uploads/2025/12/video_20251228_180532_99668f92.mp4\" type=\"video/mp4\">Your browser does not support the video tag.</video>\n</div>\n<h2>Background and Market Context</h2>\n<p>\nImagine you walk into a room with your phone and use it to slowly scan the space. What happens inside your phone is amazing—tiny points are collected by the camera, each one marking a spot in the room. These points, together, form a “point cloud.” But while humans can look at the dots and know where the chair, bed, or table is, computers need help. That’s where machine learning comes in.\n</p>\n<p>\nRight now, technology companies and researchers are racing to make computers smarter at understanding 3D scenes. This is important for things like home design apps, robots that move indoors, and even augmented reality games. The better these systems get at recognizing objects in 3D, the more useful and friendly our devices become.\n</p>\n<p>\nThe market for indoor 3D scene understanding is growing fast. Companies like Apple, Google, and many startups are working on ways to let users scan rooms and get instant feedback. For example, Apple’s RoomPlan helps people design their homes by detecting big objects like sofas or beds in real time as they move their phones. But there are limits: RoomPlan mainly finds large items, and struggles with smaller things, like books or lamps.\n</p>\n<p>\nMost current methods are “offline.” This means they take all the scan data at once, ignore the order in which points were collected, and then try to recognize objects. While this works if you have already captured the whole room, it’s not very helpful if you want feedback while you’re still scanning. Imagine you’re scanning a room corner by corner and want to know right away if you missed something—today’s offline methods can’t help much.\n</p>\n<p>\nAnother problem is that offline methods need a lot of computer memory and time. Each time you add new points, you may need to start over or re-calculate a lot of things. This slows down the process and drains your phone’s battery.\n</p>\n<p>\nSome new methods, borrowed from outdoor uses like self-driving cars, try to track both time and space, but they are often designed for big outdoor scenes and miss the fine details of indoor objects. Indoor scenes are trickier because objects can be stacked, hidden, or very close together. Plus, you may have both big objects and tiny items in the same space.</p>\n<p><img loading=\"lazy\" class=\"alignnone size-medium wp-image-56052 aligncenter\" src=\"https://wp.inventiv.org/wp-content/uploads/2025/12/generated_image_20251228_180028_df82587c-768x768.png\" alt=\"\" width=\"768\" height=\"768\" data-mce-src=\"https://wp.inventiv.org/wp-content/uploads/2025/12/generated_image_20251228_180028_df82587c-768x768.png\" data-mce-selected=\"1\" srcset=\"https://wp.inventiv.org/wp-content/uploads/2025/12/generated_image_20251228_180028_df82587c-768x768.png 768w, https://wp.inventiv.org/wp-content/uploads/2025/12/generated_image_20251228_180028_df82587c-300x300.png 300w, https://wp.inventiv.org/wp-content/uploads/2025/12/generated_image_20251228_180028_df82587c-150x150.png 150w, https://wp.inventiv.org/wp-content/uploads/2025/12/generated_image_20251228_180028_df82587c-800x800.png 800w, https://wp.inventiv.org/wp-content/uploads/2025/12/generated_image_20251228_180028_df82587c.png 1024w\" sizes=\"(max-width: 768px) 100vw, 768px\" /></p>\n</p>\n<p>\nSo, the need is clear: we want a way to scan a room, see what objects are already found, and get feedback in real time—no matter if you’re looking at a sofa or a book. The patent we’re discussing today sets out to solve exactly these problems, offering a new way for computers to learn and make sense of 3D scenes as the scan happens.\n</p>\n<div class=\"contact-us-holder\"></div>\n<h2>Scientific Rationale and Prior Art</h2>\n<p>\nTo understand what’s new, let’s look at how things have been done so far.\n</p>\n<p>\nMost modern 3D object detection systems use a kind of machine learning called neural networks. These networks are trained with lots of data—scans of rooms where every object is labeled. Famous datasets like SUN RGB-D and ScanNet are used for training. The main models, like FCAF3D and TR3D, use convolutional neural networks (CNNs) to process 3D data.\n</p>\n<p>\nHere’s how a typical model works:\n</p>\n<p>\nFirst, the raw point cloud is turned into a “voxel grid.” Imagine dividing the room into small cubes (like 3D pixels). If a cube contains any points, it’s marked as “occupied.” The grid size matters—a finer grid gives better detail, but uses more memory and takes longer to process.\n</p>\n<p>\nNext, a CNN processes the voxel grid. The network tries to find patterns that match known objects, like the corners of a chair or the top of a table. But these models often make too many guesses, sometimes marking parts of the same chair as different objects. To fix this, a rule called “non-maximum suppression” (NMS) is used. NMS keeps only the best guess for each object, based on how much predictions overlap.\n</p>\n<p>\nBut there’s a big catch. These models treat the whole scan as a single snapshot. They don’t care about the order in which the room was scanned. This is called an “offline” method. It’s good for accuracy if the whole room is already scanned, but not if you want to see results as you go.</p>\n<p><img loading=\"lazy\" class=\"alignnone size-medium wp-image-56054 aligncenter\" src=\"https://wp.inventiv.org/wp-content/uploads/2025/12/460df7a9f6fbe3ce46c48d7764dc4b99-794x1024.png\" alt=\"\" width=\"794\" height=\"1024\" data-mce-src=\"https://wp.inventiv.org/wp-content/uploads/2025/12/460df7a9f6fbe3ce46c48d7764dc4b99-794x1024.png\" data-mce-selected=\"1\" srcset=\"https://wp.inventiv.org/wp-content/uploads/2025/12/460df7a9f6fbe3ce46c48d7764dc4b99-794x1024.png 794w, https://wp.inventiv.org/wp-content/uploads/2025/12/460df7a9f6fbe3ce46c48d7764dc4b99-233x300.png 233w, https://wp.inventiv.org/wp-content/uploads/2025/12/460df7a9f6fbe3ce46c48d7764dc4b99-768x990.png 768w, https://wp.inventiv.org/wp-content/uploads/2025/12/460df7a9f6fbe3ce46c48d7764dc4b99-1192x1536.png 1192w, https://wp.inventiv.org/wp-content/uploads/2025/12/460df7a9f6fbe3ce46c48d7764dc4b99-1589x2048.png 1589w\" sizes=\"(max-width: 794px) 100vw, 794px\" /></p>\n</p>\n<p>\nApple’s RoomPlan tries to offer real-time feedback. But to do this quickly, it uses a coarse grid—meaning it only finds big objects, missing smaller items. Plus, it breaks the scan into chunks and runs detection separately on each chunk, then combines the results. This approach is fast but can miss things or get confused if objects are partly scanned.\n</p>\n<p>\nOutdoor 3D detection models, such as PointPillars or BEVFusion, focus on moving objects in open spaces, like cars or people on the street. They often flatten the 3D problem down to a 2D “bird’s-eye view.” This doesn’t work so well for indoor scenes, where objects can be anywhere—on the floor, on a table, or hanging from the ceiling.\n</p>\n<p>\nSo, the main problems with prior art are:\n</p>\n<p>\n&#8211; Offline models need the whole scene scanned before they can help.<br />\n&#8211; Online models are fast but miss small details or use tricks that only work for big objects.<br />\n&#8211; No system so far really uses the order of scanning to get smarter as new data comes in.\n</p>\n<p>\nThis is the gap that the new patent aims to fill. It combines the best of both worlds: using information about both space and time, so the computer can learn from how the scan unfolds and give feedback instantly.\n</p>\n<h2>Invention Description and Key Innovations</h2>\n<p>\nNow, let’s look closely at what this new patent brings to the table.\n</p>\n<p><img loading=\"lazy\" class=\"alignnone size-medium wp-image-56056 aligncenter\" src=\"https://wp.inventiv.org/wp-content/uploads/2025/12/c6870027c8c16e0a2decf407bd6ac6d0-794x1024.png\" alt=\"\" width=\"794\" height=\"1024\" data-mce-src=\"https://wp.inventiv.org/wp-content/uploads/2025/12/c6870027c8c16e0a2decf407bd6ac6d0-794x1024.png\" data-mce-selected=\"1\" srcset=\"https://wp.inventiv.org/wp-content/uploads/2025/12/c6870027c8c16e0a2decf407bd6ac6d0-794x1024.png 794w, https://wp.inventiv.org/wp-content/uploads/2025/12/c6870027c8c16e0a2decf407bd6ac6d0-233x300.png 233w, https://wp.inventiv.org/wp-content/uploads/2025/12/c6870027c8c16e0a2decf407bd6ac6d0-768x990.png 768w, https://wp.inventiv.org/wp-content/uploads/2025/12/c6870027c8c16e0a2decf407bd6ac6d0-1192x1536.png 1192w, https://wp.inventiv.org/wp-content/uploads/2025/12/c6870027c8c16e0a2decf407bd6ac6d0-1589x2048.png 1589w\" sizes=\"(max-width: 794px) 100vw, 794px\" /></p>\n<p>\nThe heart of the invention is a computer-implemented method for training a function (think of a clever program) that takes a 3D point cloud frame and outputs “localized representations”—that is, for each object in the scene, it gives you its position and what kind of object it is (like “sofa” or “lamp”).\n</p>\n<p>\nWhat’s special is how the training works. Instead of just looking at each frame alone, or lumping everything together, the function is trained to use both the current frame and at least the previous frame. This way, it learns how the scene grows as you scan, remembering what has already been seen and using that to make better predictions about new data.\n</p>\n<p>\nThe process starts with a dataset of sequences—each sequence is a list of point cloud frames, captured over time as someone scans a room. Each frame shows a part of the scene and comes with labels for every object found. The method trains the function so that, for each frame in the sequence, it predicts the objects in that frame using both the frame itself and the one(s) just before it.\n</p>\n<p>\nInside the function, there are two main neural networks:\n</p>\n<p>\n1. The first is a “local” network, like a CNN. It looks at each frame and creates a feature vector—a compact summary of what’s in the frame.<br />\n2. The second is an “aggregation” network (sometimes using transformers), which takes the feature vectors from the current and previous frames and combines them. This step helps the system understand how the scene is changing and what objects might be present based on what has already been seen.\n</p>\n<p>\nIf you imagine scanning a table with some books on it, the local network will spot parts of the books in each frame. The aggregation network will “remember” what was seen in the last frame, helping to piece together that it’s the same book, even if you only see a new corner each time.\n</p>\n<p>\nAnother advance is that the function can handle different resolution levels. For example, it can output feature vectors at both fine and coarse detail, which helps it spot both big and small objects. The aggregation network performs matching at each resolution, so no object slips through the cracks.\n</p>\n<p>\nTo make training efficient, the method uses “batch training,” processing several frames at once. But—and this is important—it always respects time order. The function never uses information from future frames when predicting the current one. This keeps it realistic for real-time use, where you can’t know the future.\n</p>\n<p>\nThe data for training can come from real scans (physical rooms) or virtual scans (simulated rooms). The method can work with any kind of 3D point cloud data, as long as you know the order in which frames were captured.\n</p>\n<p>\nWhen the trained function is used, it takes a new sequence of frames as you scan. For each frame, it creates a feature vector, then combines it with previous vectors to figure out which objects are present, where they are, and what kind they are. The output is a set of localized representations—think of them as little boxes drawn around every object, with labels.\n</p>\n<p>\nA key benefit is that the aggregation network is much lighter (uses less memory) than the local network. So, once the local network has processed a frame, the aggregation can quickly update predictions as you scan new areas. You don’t need to keep all past frames in memory—just the feature vectors from the last few frames. This makes the method fast and practical for phones or tablets.\n</p>\n<p>\nThe system can be implemented as a computer program, run on a device, or even as a cloud service. It can give users real-time feedback as they scan, telling them instantly if they missed an object or if more scanning is needed in a certain spot.\n</p>\n<p>\nTests show that this new method is much more accurate for online detection than simply concatenating frames and running an offline detector. It can spot both big and small objects, works in real time, and is smart about using past information to improve its guesses.\n</p>\n<p>\nIn short, the invention teaches computers to think not just about what they see now, but also about what they saw a moment ago. This “memory” helps make sense of tricky scenes and gives much better feedback to users as they scan.\n</p>\n<h2>Conclusion</h2>\n<p>\nThis new method for machine learning on 3D point clouds changes how computers see and understand real-world scenes. By using both the current view and what was seen before, it brings real-time, accurate object detection to mobile devices and other platforms. It solves big problems in today’s technology, letting you scan a room and get instant, smart feedback—no matter what you’re looking for. As more devices use 3D scanning, inventions like this will become key to making our digital experiences richer, smoother, and more helpful.\n</p>\n<p>Click here <a href=\"https://ppubs.uspto.gov/pubwebapp/\" target=\"_blank\" rel=\"noopener\">https://ppubs.uspto.gov/pubwebapp/</a> and search 20250363660.</p>\n","id":"cG9zdDo1NjA1OA==","uri":"/boost-real-time-3d-object-detection-accuracy-with-smarter-ai-sequence-analysis-for-robotics-and-automotive-216523/","title":"Boost Real-Time 3D Object Detection Accuracy with Smarter AI Sequence Analysis for Robotics and Automotive","tags":{"nodes":[{"name":"Patent Review"}]},"categories":{"nodes":[{"name":"Apple"}]},"featuredImage":{"node":{"mediaItemUrl":"https://wp.inventiv.org/wp-content/uploads/2025/12/generated_image_20251228_175947_be0506f7.png"}}}},"pageContext":{"slug":"boost-real-time-3d-object-detection-accuracy-with-smarter-ai-sequence-analysis-for-robotics-and-automotive-216523","neId":"56058","categories":["3D Printing","Animal Health","Apple","Artificial Intelligence","Autonomous Vehicles","Biopharmaceuticals","Blockchain Fintech","Cannabis Patents and Trademarks","Chemical Products","Communications","Consumer Products","Digital Healthcare","Electronics","Food Science Packaging","Industrial Products","Infectious Diseases","Internet","Medical Device","Metaverse","Nanotechnology","Prescription &amp; OTC Drugs","Software","Start Up","Therapeutic Antibodies","uncategorized"],"tags":["Alphabet","Amazon","Facebook/Meta","Microsoft","Patent Review","Samsung"],"siteUrl":"https://inventivs.net"}},
    "staticQueryHashes": ["764694655"]}