IP Library › Granted Patent US 12,205,195
Granted Patent B2
US 12,205,195 · App. 17/810,714 · Granted Jan 21, 2025

Mobile AR prototyping for proxemic and gestural interactions with real-world IoT enhanced spaces

Inventors: Hongbo Fu (Sha Tin, HK); Hui Ye (Sha Tin, HK)
Assignee: CITY UNIVERSITY OF HONG KONG
G06T11/00G06F3/017
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,205,195
App. No.
17/810,714
Filed
Jul 5, 2022
Granted
Jan 21, 2025
Kind
B2
Examiner
WANG, YI
Art Unit
2611
USPC
345/633
Abstract

One or more devices, systems, methods and/or non-transitory, machine-readable mediums are described herein for specifying one or more events in an augmented reality (AR) environment relative to a real-world (RW) environment. A system can comprise a memory that stores executable components, and a processor, coupled to the memory, that executes or facilitates execution of the executable components. The executable components can comprise a visual component that analyzes captured visual content of a RW environment, an interface component that integrates the visual content of the RW environment with AR content of an AR environment overlaying the RW environment, and a design component that facilitates in-situ placement of the AR content in the AR environment based on the visual content being overlayed, wherein the AR content comprises a specification of an event to be executed in the RW environment that triggers a virtual asset in the AR environment.

Claims (50)

1. A system, comprising:

at least one memory that stores executable components; and

at least one processor, coupled to the at least one memory, that executes or facilitates execution of the executable components, the executable components comprising:

a visual component that analyzes captured visual content of a real-world environment;

an interface component that renders an interface on a mobile device that displays the captured visual content of the real-world environment overlayed with augmented reality (AR) content of an AR environment and one or more menus; and

a design component that,

in response to receipt, via a first interaction with the interface, of a selection of a virtual asset from the one or more menus, adds the virtual asset to the AR content overlayed on the captured visual content,

places the virtual asset at a first location on the captured visual content indicated via a second interaction with the interface,

in response to receipt, via a third interaction with the interface, of an indication of a second location within the real-world environment to be associated with an event, adds an event proxy graphic representing the event to the AR content overlayed on the captured visual content at the second location, and

configures, based on interactions with the virtual asset and the event proxy graphic received via the interface, a mapping between the event and the virtual asset, wherein the mapping configures the virtual asset to perform an effect in response to detection of the event.

2. The system of claim 1 , wherein the event is a person being at a defined location within the real-world environment, the person being within a defined distance from a device represented by the virtual asset, a defined gesture performed by the person, or a defined orientation of the person.

3. The system of claim 2 , wherein the executable components further comprise a gesture recognition component that identifies the gesture absent a view of a full body of the person within the captured visual content.

4. The system of claim 1 , wherein at least one of the first interaction or the second interaction is received via a first-person view rendered via the interface or a third-person view rendered by the interface.

5. The system of claim 1 , wherein the effect performed by the virtual asset represents at least one of a visual response or an audio response to be triggered in the AR environment in response to detection of the event in the real-world environment.

6. The system of claim 1 , wherein

the executable components further comprise a detection component that detects an execution of the event in the real-world environment, and

detection of the event by the detection component causes the virtual asset to perform the effect in the AR environment.

7. The system of claim 6 , wherein,

while in a first-person testing mode, the detection component tests performance of the effect in response to triggering of the event via detection of a three-dimensional pose of the mobile device, and

while in a third-person testing mode, the detection component tests performance of the effect in response to detection of a test subject within the real-world environment performing the event.

8. The system of claim 1 , wherein the interface component further displays, at a periphery of the interface, a portion of the AR content that is not in a current view of the AR environment being displayed via the interface.

9. The system of claim 1 , wherein

the interface component further recognizes a horizontal plane and a vertical plane of the real-world environment, and visualizes, via the interface, an AR horizontal plane and an AR vertical plane corresponding to the horizontal plane and the vertical plane of the real-world environment, and

the interface component further recognizes the captured visual content of the real-world environment relative to the AR horizontal plane and the AR vertical plane.

10. A non-transitory machine-readable medium comprising executable instructions that, in response to execution by a processor, facilitate performance of operations, the operations comprising:

capturing visual content of a real-world environment;

rendering, via a mobile device, an interface that displays the visual content with an augmented reality (AR) environment and one or more menus overlaying the visual content;

in response to receiving, via a first interaction with the interface, a selection of a virtual asset from the one or more menus, adding the virtual asset to the AR environment overlayed on the visual content;

in response to receiving, via a second interaction with the interface, a first indication of a first location on the visual content, placing the virtual asset at the first location on the visual content;

in response to receiving, via a third interaction with the interface, a second indication of a second location on the visual content to be associated with an event, adding an event proxy graphic representing the event to the AR environment overlayed on the visual content at the second location; and

configuring, based on interactions with the virtual asset and the event proxy graphic received via the interface, a mapping between the event and an effect of the virtual asset, wherein the mapping configures the virtual asset to trigger the effect in response to detection of the event.

11. The non-transitory machine-readable medium of claim 10 , wherein at least one of the first interaction, the second interaction, or the third interaction is received via a first-person view or a third-person view displayed via the interface.

12. The non-transitory machine-readable medium of claim 10 , wherein

the event is a gesture performed by a person within the real-world environment, and

the operations further comprise recognizing the gesture absent a view of a full body of the person within the visual content.

13. The non-transitory machine-readable medium of claim 10 , wherein

the event is at least one of an orientation of a person within the real-world environment or a gesture of the person at a location within the real-world environment or within a distance from a device represented by the virtual asset.

14. The non-transitory machine-readable medium of claim 10 , wherein the operations further comprise displaying, at a periphery of the interface, AR content that is not in a view of the AR environment presently being displayed via the interface.

15. The non-transitory machine-readable medium of claim 10 , wherein the operations further comprise detecting the event independent of a location of a person in the real-world environment.

16. A method, comprising:

integrating, by a system operatively coupled to a processor, visual content of a real-world environment with an augmented reality (AR) environment overlaying the real-world environment on an interface rendered via a mobile device;

in response to receiving, via a first interaction with the interface, a selection of a virtual asset from a menu of virtual assets rendered on the interface, adding the virtual asset to the AR environment overlayed on the visual content;

in response to receiving, via a second interaction with the interface, a first indication of a first location within the real-world environment, overlaying the virtual asset at the first location on the visual content;

in response to receiving, via a third interaction with the interface, a second indication of a second location within the real-world environment to be associated with an event, adding an event proxy graphic corresponding to the event at the second location on the visual content; and

in response to receiving interactions with the virtual asset and the event proxy graphic defining a mapping between the event and the virtual asset, setting the mapping between the event and the virtual asset, wherein the mapping configures the virtual asset to trigger an effect in response to detection of the event.

17. The method of claim 16 , wherein the effect is at least one of a visual response or an audible response in the AR environment.

18. The method of claim 16 , wherein

at least one of the first interaction or the second interaction is received via a first-person view or a third-person rendered via the interface.

19. The method of claim 18 , further comprising specifying, by the system, the AR content to be placed on both of a pair of mobile devices.

20. The method of claim 16 , wherein the event is at least one of a person being at a specified location within the real-world environment, the person being within a specified distance from the first location of the virtual asset, a specified gesture performed by the person, or a specified orientation of the person.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jul 5, 2022
From: FU, HONGBO; YE, HUI
To: CITY UNIVERSITY OF HONG KONG
Reel/Frame 060574/0857 →
Continuity (1)
Related Publication 20240013450A1 · Jan 11, 2024
References Cited (70)
US 10679062B1 · Donnelly · 2020 [cited by examiner]
US 11087561B2 · Fu et al. · 2021 [cited by applicant]
US 20170061490A1 · Ghahremani · 2017 [cited by examiner]
Apple. 2019. Reality Composer. https://apps.apple.com/us/app/reality-composer/id1462358802. [cited by applicant]
Narges Ashtari, Andrea Bunt, Joanna McGrenere, Michael Nebeling, and Parmit K Chilana. 2020. Creating augmented and virtual reality applications: Current practices, challenges, and opportunities. In Proceedings of the 2… [cited by applicant]
Till Ballendat, Nicolai Marquardt, and Saul Greenberg. 2010. Proxemic interaction: designing for a proximity and prientation-aware environment. In ACM International Conference on Interactive Tabletops and Surfaces. 121-… [cited by applicant]
Andrea Bellucci, Aneesh P Tarun, Ahmed Sabbir Arif, and Ali Mazalek. 2016. Developing Responsive and Interactive Environments with the ROSS Toolkit. In Proceedings of the TEI'16: Tenth International Conference on Tangib… [cited by applicant]
Yuanzhi Cao, Tianyi Wang, Xun Qian, Pawan S Rao, Manav Wadhawan, Ke Huo, and Karthik Ramani. 2019. GhostAR: A time-space editor for embodied authoring of human-robot collaborative task with augmented reality. In Proceed… [cited by applicant]
Zhe Cao, Gines Hidalgo, Tomas Simon, Shih-En Wei, and Yaser Sheikh. 2018. OpenPose: realtime multi-person 2D pose estimation using Part Affinity Fields. arXiv preprint arXiv:1812.08008 (2018). [cited by applicant]
Kathy Charmaz. 2008. Constructionism and the grounded theory method. Handbook of constructionist research 1, 1 (2008), 397-412. [cited by applicant]
Kathy Charmaz. 2014. Constructing grounded theory. sage. [cited by applicant]
Nils Dahlb ck, Arne J nsson, and Lars Ahrenberg. 1993. Wizard of Oz studies—why and how. Knowledge-based systems 6, 4 (1993), 258-266. [cited by applicant]
Mark Fiala. 2005. ARTag, a fiducial marker system using digital techniques. In 2005 IEEE Computer Society Conference on Computer Vision and Pattern Recognition (CVPR'05), vol. 2. IEEE, 590-596. [cited by applicant]
James Fogarty, Jodi Forlizzi, and Scott E Hudson. 2002. Specifying behavior and semantic meaning in an unmodified layered drawing package. In Proceedings of the 15th annual ACM symposium on User interface software and t… [cited by applicant]
Maxime Garcia, R mi Ronfard, and Marie-Paule Cani. 2019. Spatial Motion Doodles: Sketching Animation in VR Using Hand Gestures and Laban Motion Analysis. In Motion, Interaction and Games. 1-10. [cited by applicant]
Danilo Gasques, Janet G Johnson, Tommy Sharkey, and Nadir Weibel. 2019. What you sketch is what you get: Quick and easy augmented reality prototyping with pintar. In Extended Abstracts of the 2019 CHI Conference on Huma… [cited by applicant]
Saul Greenberg, Nicolai Marquardt, Till Ballendat, Rob Diaz-Marino, and MiaosenWang. 2011. Proxemic interactions: the new ubicomp? interactions 18, 1 (2011), 42-50. [cited by applicant]
Sinem Güven and Steven Feiner. 2003. Authoring 3D hypermedia for wearable augmented and virtual reality. In Proceedings of IEEE International Symposium on Wearable Computers (ISWC'03). 21-23. [cited by applicant]
Edward Twitchell Hall. 1966. The hidden dimension. vol. 609. Garden City, NY: Doubleday. [cited by applicant]
Bj rn Hartmann, Scott R Klemmer, Michael Bernstein, Leith Abdulla, Brandon Burr, Avi Robinson-Mosher, and Jennifer Gee. 2006. Reflective physical prototyping through integrated design, test, and analysis. In Proceedings… [cited by applicant]
Rex Hartson and Pardha Pyla. 2012. The UX Book: Process and guidelines for ensuring a quality user experience. Elsevier. [cited by applicant]
Ke Huo, Yuanzhi Cao, Sang Ho Yoon, Zhuangying Xu, Guiming Chen, and Karthik Ramani. 2018. Scenariot: spatially mapping smart things within augmented reality scenes. In Proceedings of the 2018 CHI Conference on Human Fac… [cited by applicant]
Adobe Inc. 2019. Adobe Aero. https://apps.apple.com/us/app/adobe-aero/id1401748913. [cited by applicant]
Sujin Jang, Niklas Elmqvist, and Karthik Ramani. 2014. GestureAnalyzer: visual analytics for pattern analysis of mid-air hand gestures. In Proceedings of the 2nd ACM symposium on Spatial user interaction. 30-39. [cited by applicant]
Runchang Kang, Anhong Guo, Gierad Laput, Yang Li, and Xiang‘Anthony’ Chen. 2019. Minuet: Multimodal interaction with an Internet of Things. In Symposium on Spatial User Interaction. 1-10. [cited by applicant]
Hirokazu Kato and Mark Billinghurst. 1999. Marker tracking and hmd calibration for a video-based augmented reality conferencing system. In Proceedings 2nd IEEE and ACM International Workshop on Augmented Reality (IWAR'9… [cited by applicant]
Annie Kelly, R Benjamin Shapiro, Jonathan de Halleux, and Thomas Ball. 2018. ARcadia: A rapid prototyping platform for real-time tangible interfaces. In Proceedings of the 2018 CHI Conference on Human Factors in Computi… [cited by applicant]
Daehwan Kim and Daijin Kim. 2006. An intelligent smart home control using body gestures. In 2006 International Conference on Hybrid Information Technology, vol. 2. IEEE, 439-446. [cited by applicant]
Han-Jong Kim, Chang Min Kim, and Tek-Jin Nam. 2018. Sketchstudio: Experience prototyping with 2.5-dimensional animated design scenarios. In Proceedings of the 2018 Designing Interactive Systems Conference. 831-843. [cited by applicant]
Han-Jong Kim, Ju-Whan Kim, and Tek-Jin Nam. 2016. miniStudio: Designers' Tool for Prototyping Ubicomp Space with Interactive Miniature. In Proceedings of the 2016 CHI Conference on Human Factors in Computing Systems. 21… [cited by applicant]
Scott R Klemmer, Jack Li, James Lin, and James A Landay. 2004. Papier-Mache: toolkit support for tangible input. In Proceedings of the SIGCHI conference on Human factors in computing systems. 399-406. [cited by applicant]
Barry Kollee, Sven Kratz, and Anthony Dunnigan. 2014. Exploring gestural interaction in smart spaces using head mounted devices with ego-centric sensing. In Proceedings of the 2nd ACM symposium on Spatial user interacti… [cited by applicant]
Veronika Krau , Alexander Boden, Leif Oppermann, and Ren Reiners. 2021. Current practices, challenges, and design implications for collaborative AR/VR application development. In Proceedings of the 2021 CHI Conference o… [cited by applicant]
Kin Chung Kwan and Hongbo Fu. 2019. Mobi3DSketch: 3D Sketching in Mobile AR. In Proceedings of the 2019 CHI Conference on Human Factors in Computing Systems. ACM. [cited by applicant]
David Ledo, Saul Greenberg, Nicolai Marquardt, and Sebastian Boring. 2015. Proxemic-aware controls: Designing remote controls for ubiquitous computing ecologies. In Proceedings of the 17th International Conference on Hu… [cited by applicant]
David Ledo, Steven Houben, Jo Vermeulen, Nicolai Marquardt, Lora Oehlberg, and Saul Greenberg. 2018. Evaluation strategies for HCI toolkit research. In Proceedings of the 2018 CHI Conference on Human Factors in Computin… [cited by applicant]
David Ledo, Jo Vermeulen, Sheelagh Carpendale, Saul Greenberg, Lora Oehlberg, and Sebastian Boring. 2019. Astral: Prototyping Mobile and Smart Object Interactive Behaviours Using Familiar Applications. In Proceedings of… [cited by applicant]
Sang-Su Lee, Jeonghun Chae, Hyunjeong Kim, Youn-kyung Lim, and Kun-pyo Lee. 2013. Towards more natural digital content manipulation via user freehand gestural interaction in a living room. In Proceedings of the 2013 ACM… [cited by applicant]
Wei-Po Lee, Che Kaoli, and Jhih-Yuan Huang. 2014. A smart TV system with body-gesture control, tag-based rating and context-aware recommendation. Knowledge-Based Systems 56 (2014), 167-178. [cited by applicant]
Germ n Leiva and Michel Beaudouin-Lafon. 2018. Montage: A video prototyping system to reduce re-shooting and increase re-usability. In Proceedings of the 31st Annual ACM Symposium on User Interface Software and Technolo… [cited by applicant]
Germ n Leiva, Jens Emil Gr nb k, Clemens Nylandsted Klokmose, Cuong Nguyen, Rubaiat Habib Kazi, and Paul Asente. 2021. Rapido: Prototyping Interactive AR Experiences through Programming by Demonstration. In The 34th Ann… [cited by applicant]
Germ n Leiva, Cuong Nguyen, Rubaiat Habib Kazi, and Paul Asente. 2020. Pronto: Rapid Augmented Reality Video Prototyping Using Sketches and Enaction. In Proceedings of the 2020 CHI Conference on Human Factors in Computi… [cited by applicant]
Yang Li, Jason I Hong, and James A Landay. 2004. Topiary: a tool for prototyping location-enhanced applications. In Proceedings of the 17th annual ACM symposium on User interface software and technology. 217-226. [cited by applicant]
Hao Lu and Yang Li. 2012. Gesture coder: a tool for programming multi-touch gestures by demonstration. In Proceedings of the SIGCHI Conference on Human Factors in Computing Systems. 2875-2884. [cited by applicant]
Hao Lü and Yang Li. 2013. Gesture studio: authoring multi-touch interactions through demonstration and declaration. In Proceedings of the SIGCHI Conference on Human Factors in Computing Systems. 257-266. [cited by applicant]
Blair MacIntyre, Maribeth Gandy, Steven Dow, and Jay David Bolter. 2004. DART: a toolkit for rapid design exploration of augmented reality experiences. In Proceedings of the 17th annual ACM symposium on User interface s… [cited by applicant]
Nicolai Marquardt, Robert Diaz-Marino, Sebastian Boring, and Saul Greenberg. 2011. The proximity toolkit: prototyping proxemic interactions in ubiquitous computing ecologies. In Proceedings of the 24th annual ACM sympos… [cited by applicant]
Nicolai Marquardt and Saul Greenberg. 2012. Informing the design of proxemic interactions. IEEE Pervasive Computing 11, 2 (2012), 14-23. [cited by applicant]
Nolwenn Maudet, Germ n Leiva, Michel Beaudouin-Lafon, andWendy Mackay. 2017. Design Breakdowns: Designer-Developer Gaps in Representing and Interpreting Interactive Systems. In Proceedings of the 2017 ACM Conference on … [cited by applicant]
Brad Myers, Sun Young Park, Yoko Nakano, Greg Mueller, and Andrew Ko. 2008. How designers design and program interactive behaviors. In 2008 EEE Symposium on Visual Languages and Human-Centric Computing. IEEE, 177-184. [cited by applicant]
Yasuto Nakanishi. 2012. Virtual prototyping using miniature model and visualization for interactive public displays. In Proceedings of the Designing Interactive Systems Conference. 458-467. [cited by applicant]
Tek-Jin Nam. 2005. Sketch-based rapid prototyping platform for hardware-software integrated interactive products. In CHI'05 extended abstracts on Human factors in computing systems. 1689-1692. [cited by applicant]
Michael Nebeling and Katy Madier. 2019. 360proto: Making Interactive Virtual Reality & Augmented Reality Prototypes from Paper. In Proceedings of the 2019 CHI Conference on Human Factors in Computing Systems. 1-13. [cited by applicant]
Michael Nebeling, Janet Nebeling, Ao Yu, and Rob Rumble. 2018. Protoar: Rapid physical-digital prototyping of mobile augmented reality applications. In Proceedings of the 2018 CHI Conference on Human Factors in Computin… [cited by applicant]
Dan R Olsen Jr. 2007. Evaluating user interface systems research. In Proceedings of the 20th annual ACM symposium on User interface software and technology. 251-258. [cited by applicant]
Hyungjun Park, Hee-Cheol Moon, and Jae Yeol Lee. 2009. Tangible augmented prototyping of digital handheld products. Computers in Industry 60, 2 (2009), 114-125. [cited by applicant]
Sarah Prange and Florian Alt. 2020. I Wish You Were Smart (er): Investigating Users' Desires and Needs Towards Home Appliances. In Extended Abstracts of the 2020 CHI Conference on Human Factors in Computing Systems. 1-8. [cited by applicant]
Joseph Redmon, Santosh Divvala, Ross Girshick, and Ali Farhadi. 2016. You only look once: Unified, real-time object detection. In Proceedings of the EEE conference on computer vision and pattern recognition. 779-788. [cited by applicant]
Gang Ren, Wenbin Li, and Eamonn O'Neill. 2016. Towards the design of effective freehand gestural interaction for interactive tv. Journal of Intelligent & Fuzzy Systems 31, 5 (2016), 2659-2674. [cited by applicant]
Nazmus Saquib, Rubaiat Habib Kazi, Li-Yi Wei, and Wilmot Li. 2019. Interactive body-driven graphics for augmented video performance. In Proceedings of the 2019 CHI Conference on Human Factors in Computing Systems. 1-12. [cited by applicant]
Hartmut Seichter, Julian Looser, and Mark Billinghurst. 2008. ComposAR: An intuitive tool for authoring AR applications. In 2008 7th IEEE/ACM International Symposium on Mixed and Augmented Reality. IEEE, 177-178. [cited by applicant]
Teddy Seyed, Alaa Azazi, Edwin Chan, Yuxi Wang, and Frank Maurer. 2015. Sod-toolkit: A toolkit for interactively prototyping and developing multi-sensor, multi-device environments. In Proceedings of the 2015 Internation… [cited by applicant]
Kihoon Son, Hwiwon Chun, Sojin Park, and Kyung Hoon Hyun. 2020. C-Space: An Interactive Prototyping Platform for Collaborative Spatial Design Exploration. In Proceedings of the 2020 CHI Conference on Human Factors in Co… [cited by applicant]
Maximilian Speicher and Michael Nebeling. 2018. GestureWiz: A human-powered gesture design environment for user interface prototypes. In Proceedings of the 2018 CHI Conference on Human Factors in Computing Systems. 1-11. [cited by applicant]
Ryo Suzuki, Rubaiat Habib Kazi, Li-Yi Wei, Stephen DiVerdi, Wilmot Li, and Daniel Leithinger. 2020. RealitySketch: Embedding Responsive Graphics and Visualizations in AR through Dynamic Sketching. In Proceedings of the … [cited by applicant]
John Underkoffler and Hiroshi Ishii. 1999. Urp: a luminous-tangible workbench for urban planning and design. In Proceedings of the SIGCHI conference on Human Factors in Computing Systems. 386-393. [cited by applicant]
Tianyi Wang, Xun Qian, Fengming He, Xiyun Hu, Yuanzhi Cao, and Karthik Ramani. 2021. GesturAR: An Authoring System for Creating Freehand Interactive Augmented Reality Applications. In The 34th Annual ACM Symposium on Us… [cited by applicant]
Tianyi Wang, Xun Qian, Fengming He, Xiyun Hu, Ke Huo, Yuanzhi Cao, and Karthik Ramani. 2020. CAPturAR: An Augmented Reality Tool for Authoring Human-Involved Context-Aware Applications. In Proceedings of the 33rd Annual… [cited by applicant]
Zeyu Wang, Cuong Nguyen, Paul Asente, and Julie Dorsey. 2021. DistanciAR: Authoring Site-Specific Augmented Reality Experiences for Remote Environments. In Proceedings of the 2021 CHI Conference on Human Factors in Comp… [cited by applicant]
Xiabing Liu, Wei Liang, Yumeng Wang, Shuyang Li, and Mingtao Pei. 2016. 3D head pose estimation with convolutional neural network trained on synthetic images. In 2016 IEEE International Conference on Image Processing (I… [cited by applicant]