WO2023035829A1 - 一种确定和呈现目标标记信息的方法与设备 - Google Patents
一种确定和呈现目标标记信息的方法与设备 Download PDFInfo
- Publication number
- WO2023035829A1 WO2023035829A1 PCT/CN2022/110472 CN2022110472W WO2023035829A1 WO 2023035829 A1 WO2023035829 A1 WO 2023035829A1 CN 2022110472 W CN2022110472 W CN 2022110472W WO 2023035829 A1 WO2023035829 A1 WO 2023035829A1
- Authority
- WO
- WIPO (PCT)
- Prior art keywords
- information
- target
- scene
- real
- image
- Prior art date
- Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
- Ceased
Links
Images
Classifications
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06T—IMAGE DATA PROCESSING OR GENERATION, IN GENERAL
- G06T7/00—Image analysis
- G06T7/70—Determining position or orientation of objects or cameras
- G06T7/73—Determining position or orientation of objects or cameras using feature-based methods
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06F—ELECTRIC DIGITAL DATA PROCESSING
- G06F18/00—Pattern recognition
- G06F18/20—Analysing
- G06F18/22—Matching criteria, e.g. proximity measures
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06F—ELECTRIC DIGITAL DATA PROCESSING
- G06F3/00—Input arrangements for transferring data to be processed into a form capable of being handled by the computer; Output arrangements for transferring data from processing unit to output unit, e.g. interface arrangements
- G06F3/01—Input arrangements or combined input and output arrangements for interaction between user and computer
- G06F3/011—Arrangements for interaction with the human body, e.g. for user immersion in virtual reality
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06T—IMAGE DATA PROCESSING OR GENERATION, IN GENERAL
- G06T19/00—Manipulating three-dimensional [3D] models or images for computer graphics
- G06T19/006—Mixed reality
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06T—IMAGE DATA PROCESSING OR GENERATION, IN GENERAL
- G06T7/00—Image analysis
- G06T7/20—Analysis of motion
- G06T7/246—Analysis of motion using feature-based methods, e.g. the tracking of corners or segments
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06V—IMAGE OR VIDEO RECOGNITION OR UNDERSTANDING
- G06V20/00—Scenes; Scene-specific elements
- G06V20/20—Scenes; Scene-specific elements in augmented reality scenes
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06T—IMAGE DATA PROCESSING OR GENERATION, IN GENERAL
- G06T2207/00—Indexing scheme for image analysis or image enhancement
- G06T2207/10—Image acquisition modality
- G06T2207/10016—Video; Image sequence
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06T—IMAGE DATA PROCESSING OR GENERATION, IN GENERAL
- G06T2207/00—Indexing scheme for image analysis or image enhancement
- G06T2207/10—Image acquisition modality
- G06T2207/10028—Range image; Depth image; 3D point clouds
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06T—IMAGE DATA PROCESSING OR GENERATION, IN GENERAL
- G06T2207/00—Indexing scheme for image analysis or image enhancement
- G06T2207/30—Subject of image; Context of image processing
- G06T2207/30244—Camera pose
Definitions
- the present application relates to the communication field, and in particular to a technology for determining and presenting target marking information.
- Augmented Reality (AR) technology is a technology that ingeniously integrates virtual information with the real world. It uses multimedia, 3D modeling, real-time tracking and registration, intelligent interaction, sensing and other technical means to integrate computer
- the generated text, images, 3D models, music, video and other virtual information are simulated and applied to the real world, and the two kinds of information complement each other, thereby realizing the "enhancement" of the real world.
- AR augmented reality technology is a relatively new technical content that promotes the integration of real world information and virtual world information.
- implement simulation processing superimpose virtual information content in the real world for effective application, and in this process can be perceived by human senses, so as to achieve a sensory experience beyond reality.
- it is usually necessary to mark a specific location or a specific target to make it easier for the user to obtain interactive content and the like.
- An object of the present application is to provide a method and device for determining and presenting target marker information.
- a method for determining target tag information is provided, which is applied to a first user equipment, and the method includes:
- the real-time scene image about the current scene is captured by the camera device, the real-time pose information of the camera device and the 3D point cloud information corresponding to the current scene are obtained through three-dimensional tracking, and the real-time image of the user in the current scene is obtained.
- a method for presenting target tag information which is applied to a second user equipment.
- the method includes:
- Target scene information includes corresponding target tag information
- target tag information includes corresponding tag information and target spatial position
- the corresponding target image position is determined according to the current pose information of the camera device and the target spatial position, and the marker information is superimposed on the target image position of the current scene image.
- a method for determining and presenting target marker information comprising:
- the first user equipment captures an initial scene image about the current scene by a camera device, and performs three-dimensional tracking initialization on the current scene according to the initial scene image;
- the first user equipment captures real-time scene images of the current scene through the camera device, obtains real-time pose information of the camera device and 3D point cloud information corresponding to the current scene through three-dimensional tracking, and obtains The tag information input in the real-time scene image of the current scene and the target image position corresponding to the tag information;
- the first user equipment determines a corresponding target spatial position according to the real-time pose information and the target image position, and generates corresponding target marker information according to the target spatial position and the marker information, wherein the target marker The information is used to superimpose and present the marker information at the target space position of the 3D point cloud information;
- the second user equipment acquires the target scene information that matches the current scene, where the target scene information includes corresponding target tag information, and the target tag information includes corresponding tag information and a target spatial position;
- the second user equipment captures a current scene image related to the current scene by the camera device
- the second user equipment determines a corresponding target image position according to the current pose information of the camera device and the target spatial position, and superimposes and presents the marker information on the target image position of the current scene image.
- a first user equipment for determining target tag information includes:
- a module used for taking an initial scene image about the current scene by a camera, and carrying out three-dimensional tracking initialization for the current scene according to the initial scene image;
- the first and second modules are used to shoot real-time scene images about the current scene through the camera device, obtain real-time pose information of the camera device and 3D point cloud information corresponding to the current scene through three-dimensional tracking, and obtain the user's
- the tag information input in the real-time scene image of the current scene and the target image position corresponding to the tag information;
- a third module configured to determine the corresponding target spatial position according to the real-time pose information and the target image position, and generate corresponding target marker information according to the target spatial position and the marker information, wherein the target marker The information is used to superimpose and present the marker information at the target space position of the 3D point cloud information.
- a second user device for presenting target marking information wherein the device includes:
- a two-one module configured to acquire target scene information matching the current scene, wherein the target scene information includes corresponding target tag information, and the target tag information includes corresponding tag information and target spatial position;
- a two-two module configured to capture a current scene image about the current scene through the camera device
- the second and third modules are configured to determine the corresponding target image position according to the current pose information of the camera device and the target spatial position, and superimpose and present the marker information on the target image position of the current scene image.
- a computer device wherein the device includes:
- a memory arranged to store computer-executable instructions which, when executed, cause the processor to perform the steps of any one of the methods described above.
- a computer-readable storage medium on which computer programs/instructions are stored, wherein the computer program/instructions, when executed, cause the system to perform any of the methods described above. step.
- a computer program product including computer programs/instructions, which is characterized in that, when the computer program/instructions are executed by a processor, the steps of any one of the methods described above are implemented.
- this application obtains the 3D point cloud information corresponding to the current scene and the real-time pose of the camera device by performing three-dimensional tracking on the current scene, and is used to determine the corresponding marker information input by the user in the real-time scene image of the current scene.
- the spatial location can facilitate the user to edit the target object in the scene to provide relevant interactive information.
- it is also conducive to the scene collaboration of multiple users, etc., which improves the user experience.
- FIG. 1 shows a flowchart of a method for determining target tag information according to an embodiment of the present application
- FIG. 2 shows a flow chart of a method for presenting target tag information according to another embodiment of the present application
- Fig. 3 shows a flow chart of a system method for determining and presenting target marker information according to an embodiment of the present application
- FIG. 4 shows functional modules of a first user equipment according to an embodiment of the present application
- FIG. 5 shows functional modules of a second user equipment according to an embodiment of the present application
- FIG. 6 illustrates an exemplary system that may be used to implement various embodiments described in this application.
- the terminal, the device serving the network, and the trusted party all include one or more processors (for example, a central processing unit (Central Processing Unit, CPU)), an input/output interface, a network interface and Memory.
- processors for example, a central processing unit (Central Processing Unit, CPU)
- CPU Central Processing Unit
- Memory may include non-permanent memory in computer-readable media, random access memory (Random Access Memory, RAM) and/or non-volatile memory, such as read-only memory (Read Only Memory, ROM) or flash memory (Flash Memory). Memory is an example of computer readable media.
- RAM Random Access Memory
- ROM Read Only Memory
- Flash Memory Flash Memory
- Computer-readable media including both permanent and non-permanent, removable and non-removable media, can be implemented by any method or technology for storage of information.
- Information may be computer readable instructions, data structures, modules of a program, or other data.
- the example of the storage medium of computer includes, but not limited to Phase-Change Memory (Phase-Change Memory, PCM), Programmable Random Access Memory (Programmable Random Access Memory, PRAM), Static Random-Access Memory (Static Random-Access Memory, SRAM), Dynamic Random Access Memory (Dynamic Random Access Memory, DRAM), other types of Random Access Memory (RAM), Read Only Memory (ROM), Electrically Erasable Programmable Read-Only Memory (Electrically-Erasable Programmable Read- Only Memory, EEPROM), flash memory or other memory technology, CD-ROM (Compact Disc Read-Only Memory, CD-ROM), Digital Versatile Disc (Digital Versatile Disc, DVD) or other optical storage, Magnetic tape cartridge, tape disk storage or other magnetic storage device or any other
- the equipment referred to in this application includes, but is not limited to, user equipment, network equipment, or equipment formed by integrating user equipment and network equipment through a network.
- the user equipment includes but is not limited to any mobile electronic product that can perform human-computer interaction with the user (for example, human-computer interaction through a touchpad), such as a smart phone, a tablet computer, smart glasses, etc., and the mobile electronic product can Use any operating system, such as Android operating system, iOS operating system, etc.
- the network device includes an electronic device that can automatically perform numerical calculation and information processing according to pre-set or stored instructions, and its hardware includes but is not limited to a microprocessor, an application specific integrated circuit (Application Specific Integrated Circuit, ASIC) , Programmable Logic Device (PLD), Field Programmable Gate Array (Field Programmable Gate Array, FPGA), Digital Signal Processor (Digital Signal Processor, DSP), embedded devices, etc.
- ASIC Application Specific Integrated Circuit
- PLD Programmable Logic Device
- FPGA Field Programmable Gate Array
- DSP Digital Signal Processor
- the network equipment includes but is not limited to a computer, a network host, a single network server, a plurality of network server sets or a cloud formed by multiple servers; here, the cloud is composed of a large number of computers or network servers based on cloud computing (Cloud Computing), wherein , cloud computing is a kind of distributed computing, a virtual supercomputer composed of a group of loosely coupled computer sets.
- the network includes, but is not limited to, the Internet, a wide area network, a metropolitan area network, a local area network, a VPN network, a wireless self-organizing network (Ad Hoc network) and the like.
- the device may also be a program running on the user device, network device, or a device formed by integrating user device and network device, network device, touch terminal or network device and touch terminal through a network.
- Fig. 1 shows a method for determining target tag information according to one aspect of the present application, which is applied to a first user equipment, wherein the method includes step S101, step S102 and step S103.
- step S101 an initial scene image about the current scene is taken by the camera, and the three-dimensional tracking initialization of the current scene is performed according to the initial scene image
- step S102 an image about the current scene is taken by the camera Real-time scene image, obtain the real-time pose information of the camera device and the 3D point cloud information corresponding to the current scene through three-dimensional tracking, and obtain the mark information and the mark information input by the user in the real-time scene image of the current scene Corresponding target image position
- step S103 determine the corresponding target spatial position according to the real-time pose information and the target image position, and generate corresponding target tag information according to the target spatial position and the tag information, wherein , the target mark information is used to superimpose and present the mark information on the target space position of the 3D point cloud information.
- the first user equipment includes but is not limited to a human-computer interaction device with a camera, such as a personal computer, a smart phone, a tablet computer, a projector, smart glasses, or a smart helmet, etc., wherein the first user equipment It can contain a camera itself, or it can be an external camera.
- the first user equipment further includes a display device (such as a display screen, a projector, etc.), configured to superimpose and display corresponding marker information and the like in the presented real-time scene image.
- an initial scene image of the current scene is captured by a camera device, and three-dimensional tracking initialization is performed on the current scene according to the initial scene image.
- the user holds the first user equipment, and collects an initial scene image about the current scene through a camera of the first user equipment, and the number of the initial scene images may be one or more.
- the first user equipment may perform three-dimensional tracking initialization on the current scene based on the one or more initial scene images, such as SLAM (simultaneous localization and mapping, simultaneous localization and mapping) initialization.
- the specific initialization methods include, but are not limited to, double-frame initialization, single-frame initialization, 2D recognition initialization, 3D model initialization, and the like.
- the first user equipment may complete the 3D tracking initialization locally, or may upload the collected initial scene image to other devices (such as a network server), and other devices may perform 3D tracking initialization, which is not limited here.
- the initial pose information of the camera device can be obtained by performing three-dimensional tracking initialization.
- the initial map point information corresponding to the current scene can also be obtained.
- the first user equipment may present a prompt message indicating that the initialization is complete to the user, so as to remind the user to add marker information (such as AR marker information) and the like.
- the initialization method includes but is not limited to: completing initialization by identifying preset 2D markers in the current scene; single-frame initialization; double-frame initialization; 3D model initialization, etc. Specific examples are as follows:
- the first user equipment obtains the image with the 2D recognition map through the camera device, it extracts image features and performs matching recognition with the stored feature library, determines the pose of the camera device relative to the current scene, and uses the information obtained by the recognition algorithm (camera device pose) is sent to the SLAM algorithm to complete the initialization of the tracking.
- the recognition algorithm camera device pose
- Single-frame initialization uses the homography matrix to obtain the corresponding rotation and translation matrices under the condition that the image sensor (that is, the camera device) acquires an approximate planar scene, thereby initializing the map points and the pose of the camera device.
- the 3D model of the tracking target it is necessary to obtain the 3D model of the tracking target, use the 3D model to obtain 3D edge features, and then obtain the initial pose of the tracking target. Render the 3D edge features in the initial pose on the application interface, shoot a video containing the tracking target, and read the image frame of the video. The user aligns the target in the scene with the 3D model edge feature, performs tracking and matching, and after tracking the target, obtains The current pose and feature point information are passed to the SLAM algorithm for initialization.
- step S102 the real-time scene image of the current scene is captured by the camera device, the real-time pose information of the camera device and the 3D point cloud information corresponding to the current scene are obtained through three-dimensional tracking, and the user's position in the current scene is obtained.
- the first user equipment performs three-dimensional tracking initialization, it continues to scan the current scene through the camera device to obtain a corresponding real-time scene image, and processes the real-time scene image through the tracking thread of the three-dimensional tracking algorithm, the local mapping thread, etc., and obtains the real-time scene image.
- the 3D point cloud information includes 3D points.
- the 3D point cloud information includes 3D map points determined after image matching and depth information acquisition. Data type, the scan data is recorded in the form of points, each point contains three-dimensional coordinates, and some may contain color information (R, G, B) or the intensity of the reflective surface of the object.
- the 3D point cloud information in addition to corresponding 3D map points, also includes, but is not limited to: key frames corresponding to point cloud information, common view information corresponding to point cloud information, and growth trees corresponding to point cloud information information etc.
- the real-time pose information includes the real-time position and attitude of the camera device in space, through which the image position and spatial position can be converted, etc.
- the pose information includes the world coordinates of the camera device relative to the current scene
- the pose information includes the external parameters of the camera device relative to the world coordinate system of the current scene and the internal parameters of the camera coordinate system and the image/pixel coordinate system of the camera device, which are not limited here.
- the first user device may obtain relevant operations of the user in the real-time scene image, such as selection of a target point determined by touch, click, voice, gesture, or head movement.
- the first user equipment can determine the target image position of the target point based on the user's selection operation in the real-time scene image, and the target image position is used to represent the two-dimensional coordinates of the target point in the pixel/image coordinate system corresponding to the real-time scene image information etc.
- the first user equipment may also acquire the marking information input by the user, where the marking information includes human-computer interaction content input by the user, and is used to be superimposed on the spatial position corresponding to the target image position in the current scene, and the like. For example, the user selects a mark information on the first user equipment, such as a 3D arrow, and then clicks on a certain position in the real-time scene image displayed on the display screen. scene.
- the tag information includes, but is not limited to: identification information; file information; form information; application call information; real-time sensor information.
- marking information may include identification information such as arrows, brushes-scribbles on the screen, circles, geometric shapes, and the like.
- the marking information may also include corresponding multimedia file information, such as pictures, videos, 3D models, PDF files, office documents, and other types of files.
- the marking information may also include form information, for example, a form is generated at a position corresponding to the target image for the user to view or input content.
- the tag information may also include application invocation information, related instructions for executing the application, etc., such as opening the application, invoking specific functions of the application, such as making a call, opening a link, and the like.
- the tag information may also include real-time sensing information, which is used to connect a sensing device (such as a sensor, etc.) and acquire sensing data of a target object.
- the mark information includes any of the following: mark identification, such as mark icon, name, etc.; mark content, such as PDF file content, mark color, size, etc.; mark type, such as file information, application call information etc.
- mark identification such as mark icon, name, etc.
- mark content such as PDF file content, mark color, size, etc.
- mark type such as file information, application call information etc.
- the application calling information is used to call the first target application installed in the current device.
- the first user equipment currently has multiple applications installed, and the application calling information is used to call one of the applications currently installed on the first user equipment, for example, after starting the corresponding application and executing related shortcut instructions, such as starting a phone application, and launching Call to Zhang XX, etc.
- the application calling information is used to call the application in the first user equipment, and if the tag information is presented on other user equipment (such as the second user equipment, etc.), the application The calling information is used to call the corresponding application in the second user equipment, such as starting the phone application of the second user equipment, and initiating a call to Zhang XX.
- the application invocation information is used to prompt the user to install the corresponding application and perform invocation after the installation is completed.
- the application invocation information is used to invoke a second target application installed in a corresponding third user equipment, where there is a communication connection between the third user equipment and the first user equipment.
- the third user equipment is in the current scene, and the third user equipment is connected to the first user equipment through wired, wireless or network equipment.
- the first user equipment may send an instruction corresponding to the application invocation information to the third user equipment, so that the third user equipment invokes a relevant application to perform a corresponding operation and the like.
- the first user device is augmented reality glasses
- the third user device is an operating device on the workbench in the current scene
- the third user device is installed with an operation application for operating workpieces
- the first user device may be based on user The operation adds tag information in the real-time scene image, and the tag information is the application invocation information corresponding to the operation application of the third user equipment. If the trigger operation of the relevant user on the application invocation information is acquired subsequently, the third user equipment will call The corresponding operation application processes the workpiece, etc.
- the user equipment when the user equipment (such as the first user equipment or the second user equipment, etc.) acquires the user's trigger operation on the displayed marker information in the scene image, it sends an instruction corresponding to the application calling information to the third user. equipment, so that the third user equipment invokes related applications to execute related operations or instructions.
- the trigger operation includes but not limited to click, touch, gesture instruction, voice instruction, button, head movement instruction and so on.
- the corresponding target spatial position is determined according to the real-time pose information and the target image position, and the corresponding target marker information is generated according to the target spatial position and the marker information, wherein the target marker information It is used to overlay and present the marker information at the target space position of the 3D point cloud information.
- the spatial three-dimensional coordinates corresponding to the two-dimensional coordinates on the real-time scene image captured by the camera device can be estimated, so as to determine the corresponding target spatial position based on the target image position. For example, after SLAM initialization is completed, the 3D point cloud and the pose of the camera device in the environment are calculated in real time based on the real-time scene image.
- the algorithm uses the 3D point cloud in the current scene to fit a world coordinate
- the plane under the system is obtained the plane expression.
- a ray based on the camera coordinate system is constructed, and then the ray is transformed into the world coordinate system, which is calculated by the ray expression and the plane expression in the world coordinate system
- the intersection point of the ray and the plane, this intersection point is the 3D space point corresponding to the 2D click point in the scene image captured by the camera device, the coordinate position corresponding to the 3D space point is determined as the target space position corresponding to the marker information, the target space position It is used for placing corresponding marker information in space, so that the marker information is superimposed and displayed at a corresponding position in the real-time scene image captured by the camera device, and the marker information is rendered on the real-time scene image.
- the first user equipment generates corresponding target tag information according to the corresponding target spatial position and tag information, such as generating a corresponding tag configuration file, the configuration file includes spatial coordinate information and tag information, etc.
- the corresponding tag configuration file can be related to
- the corresponding scene 3D point cloud information is stored in the same file, which facilitates management and storage.
- the tag information includes real-time sensing information
- the real-time sensing information is used to indicate the real-time sensing data of the corresponding sensing device, and there is a communication connection.
- real-time sensing data includes real-time collection of any object or process that needs to be monitored, connected, and interacted with through various devices and technologies such as information sensors, radio frequency identification technology, global positioning systems, infrared sensors, and laser scanners. Sound, light, heat, electricity, mechanics, chemistry, biology, location and other necessary information.
- more common sensing devices include temperature sensors, humidity sensors, brightness sensors, etc., and the above sensing devices are just examples and not limited.
- a communication connection is established between the corresponding sensing device and the first user equipment through wired, wireless or network equipment, so that based on the communication connection, the first user equipment can obtain real-time sensing information of the sensing device, etc., and the sensing device sets Since the target object is used to collect real-time sensing data of the target object (such as other equipment or objects, etc.), the real-time sensing information can be updated based on the currently collected real-time image, or the real-time sensing information can be updated based on a predetermined time interval renew.
- the first user equipment may present one or more selectable objects contained in the scene image, where each selectable object has been provided with a corresponding Select the mark information of a real-time sensing information, such as a temperature sensor, and then click on an optional object of the temperature sensor displayed on the display screen, and the real-time temperature sensing mark information is superimposed on the selectable object displayed in the real-time scene
- a real-time sensing information such as a temperature sensor
- the data of the subsequent temperature sensor is updated every 0.5 seconds; for another example, the user determines the sensor device set by a certain selectable object, and then the user selects the real-time sensor device corresponding to the sensor device on the first user equipment.
- the selected real-time sensing mark information will be superimposed and displayed at the position of the selectable object in the real-time scene, and the subsequent data of the sensor will be displayed according to each real-time scene
- the image is updated.
- the method further includes step S104 (not shown), in step S104, receiving the real-time sensing information returned by the network device and acquired by the sensing device; wherein, the Generating corresponding target marker information from the target spatial position and the marker information includes: generating corresponding target marker information according to the target spatial position and the real-time sensing information, wherein the target marker information is used in the The target spatial position of the 3D point cloud information is superimposed to present the real-time sensing information.
- the first user equipment determines to add corresponding real-time sensing information to the real-time scene image based on the user's operation, and determines the target spatial position corresponding to the real-time sensing information, and establishes communication between the first user equipment and the sensing device through network equipment connection, the network device can obtain real-time sensing data of the sensing device in real time, and send the real-time sensing data to the first user equipment.
- the first user equipment receives the corresponding real-time sensing data, and superimposes and presents the real-time sensing data at the corresponding spatial position of the real-time scene image (for example, calculates the corresponding real-time image position in real time according to the target spatial position, real-time pose information, etc.).
- the determining the corresponding target spatial position according to the real-time pose information and the target image position includes: mapping the target image position to space according to the real-time pose information to determine the corresponding 3D straight line: determine the target spatial position corresponding to the target point according to the 3D point cloud information and the 3D straight line.
- the first user device first determines the corresponding 3D straight line according to the real-time pose information and the target image position of the target point in the real-time scene image, and then determines the corresponding target point according to the spatial position relationship between the 3D point cloud information and the 3D straight line
- the target space position of the position such as according to the intersection point of the 3D straight line and a certain plane, the distance between the 3D straight line and the feature points in the 3D point cloud information, etc., wherein the 3D straight line can be a straight line in the camera coordinate system of the camera device, or it can be A line in the world coordinate system.
- the target image position that is, the 2D marker point (target point) P2d is mapped to a straight line L3dC in the camera coordinate system of the camera device, and the 3D point cloud obtained by the SLAM algorithm is mapped to the camera coordinate system to obtain the 3D point, in some embodiments, find the point P3d'C closest to the vertical distance from the straight line L3dC in the 3D point cloud under the camera coordinate system, and use the depth value of P3d'C to obtain the corresponding value of the depth value on the straight line L3dC Point P3dC; in some other embodiments, the distance weighted average between the 3D point cloud and L3dC in the camera coordinate system is used as the depth value, and according to this depth value, the point P3dC corresponding to the depth value is taken on the straight line L3dC; the point P3dC is mapped to the world coordinate system to obtain P3d, then P3d is the estimate of the 2D point in 3D space, so as to determine the target space
- the point with the smallest distance to the marker point P2d is found among these 2Ds points, and the depth value of the 3D point in the camera coordinate system corresponding to the 2Ds point is used as the intercepted depth value, and is mapped to the camera at the marker point P2d Take the point corresponding to the depth value on a straight line L3dC in the coordinate system to obtain the estimated point P3dC in the camera coordinate system, and then convert it to the world coordinate system to obtain the target space position; in some embodiments, find in these 2Ds points The point with the smallest distance from the marker point P2d, the depth value of the 3D point in the world coordinate system corresponding to this 2Ds point is used as the intercepted depth value, and the depth value is taken on a straight line L3d mapped from the marker point P2d to the world coordinate system Corresponding point, get the estimated point P3d under the world coordinate system, get the target space position; In some other embodiments, determine the weight according to the distance between these 2Ds
- the weighted average is (weight of each point*depth value)/number of points. According to the final depth value, the marker point P2d is mapped to a line in the camera coordinate system The estimated point P3dC is intercepted on the straight line L3dC, and then converted to the world coordinate system to obtain the target space position.
- the weight is determined according to the distance between these 2Ds points and the marker point P2d
- the final depth value is determined according to the weighted average of the depth values of the 3D points in the world coordinate system corresponding to these 2Ds points
- the final depth value is determined according to the final depth value , intercept the estimated point P3d on a straight line L3d mapped from the marked point P2d to the world coordinate system, and obtain the target space position.
- the above-mentioned mapping between the coordinate systems utilizes pose information of the camera device.
- the 3D point cloud information includes a plurality of feature points, and each feature point includes corresponding depth information; wherein, according to the 3D point cloud information and the 3D straight line, the The target spatial position includes: determining at least one target feature point from the 3D point cloud information according to the distance between each feature point in the 3D point cloud information and the 3D straight line; based on the depth of the at least one target feature point The information determines the depth information of the target point on the 3D straight line, so as to determine the corresponding target spatial position.
- the multiple feature points refer to 3D points in the 3D point cloud information.
- the target image position that is, the 2D marker point (target point) P2d is mapped to a straight line L3d in the world coordinate system through the corresponding real-time pose information when the marker information is input in the real-time scene image of the current scene, Calculate the distance between each feature point and the straight line L3d from the 3D point cloud in the world coordinate system obtained by the SLAM algorithm, and select at least one corresponding target feature point according to the distance between each feature point and the straight line.
- the first user equipment determines the depth information of the target point according to the depth information of at least one target feature point, such as using the depth information of a certain target feature point in the at least one target feature point as the depth information of the target point, or by weighted average
- the calculation method determines the depth information of the target point, etc., and determines the corresponding target spatial position according to the depth information of the target point, such as mapping the mark point P2d to a straight line L3d in the world coordinate system according to the depth information of the target point Take the point corresponding to the depth value to get the estimated point P3d in the world coordinate system, so as to get the target space position.
- the at least one target feature point includes a feature point with the smallest distance to the 3D straight line.
- the distance between each target feature point of the at least one target feature point and the 3D straight line is less than or equal to a distance threshold; wherein, the determination based on the depth information of the at least one target feature point
- the depth information of the target point on the 3D straight line includes: determining the weight information of each target feature point according to the distance information between the at least one target feature point and the 3D straight line; The depth information and the weight information determine the depth information of the target point on the 3D straight line.
- the 2D marker point P2d is mapped to a straight line L3d in the world coordinate system through the corresponding real-time pose information when the marker information is input in the real-time scene image of the current scene.
- the world coordinate obtained from the SLAM algorithm Find the point P3d' with the closest vertical distance to the straight line L3d in the 3D point cloud under the system.
- the depth value of the 3D point P3d' take the point corresponding to the depth value on the straight line L3d as the estimated point P3d, and obtain
- the weighted average of the distance between the 3D point cloud and L3d in the world coordinate system obtained according to the SLAM algorithm is used as the depth value, and further, the 3D point cloud participating in the weighted average calculation, which is related to the straight line L3d
- the distance of is less than or equal to the distance threshold, according to this depth value, take the point corresponding to this depth value on the straight line L3d as the estimated point P3d, and obtain the estimation in the world coordinate system.
- the weighted average of the distance between the 3D point and L3d is specifically (z value of each 3D point*weight coefficient)/number of points, and the weight is based on the vertical distance from the 3D point to the straight line L3d. The smaller the distance, the smaller the weight , and the sum of the weights is 1.
- the method further includes step S105 (not shown), in step S105, updating target scene information based on the target tag information, wherein the target scene information is stored in a scene database, and the scene
- the database includes one or more target scene information, and each target scene information includes target marker information and a corresponding 3D point cloud.
- the scene database may be stored locally on the first user equipment, or may be stored on a network device having a communication connection with the first user equipment.
- the first user equipment generates target scene information about the current scene based on target tag information, 3D point cloud information, etc., and updates the local scene database based on the target scene information, or sends the target scene information to the corresponding Network devices to update the scene database, etc.
- the first user equipment may update the target scene information after determining all the target tag information of the current scene, or the first user equipment may update the target scene information at preset time intervals, or the first user equipment may update the target scene information based on User's manual operations, such as clicking the save button, updating target scene information, etc.
- the target scene information includes 3D point cloud information corresponding to the scene, and at least one target tag information associated with the 3D point cloud information, and the corresponding target tag information includes corresponding tag information and target spatial position.
- each piece of target scene information further includes corresponding device parameter information.
- the device parameter information includes the internal reference of the device camera or the identification information of the device, such as the internal reference of the camera device of the user device that generates the 3D point cloud information or the identification information of the device.
- the first user equipment performs three-dimensional tracking on the scene to determine the 3D point cloud information of the scene and determine the target marker.
- the target scene information also includes internal parameters of the camera device of the first user equipment. After the user equipment other than the first user equipment uses the target scene, it needs to use the internal parameters of the camera device of the first user equipment when initializing the point cloud, so as to increase the accuracy of the display position of the target marker on the display devices of other user equipment.
- the target scene information may directly include the internal parameters of the equipment camera, or the internal parameters of the equipment camera may be determined through the identification information (such as ID, name, etc.) of the equipment.
- each target scene information also includes corresponding scene identification information, such as scene preview and description, such as using a picture of the scene to identify the scene, and presenting the picture of the scene on the device, etc. .
- the method further includes step S106 (not shown), in step S106, the The target scene information is sent to the corresponding network device, wherein the target scene information is sent to the second user device via the network device based on the scene call request of the second user device, and the target scene information is used in the The tag information is superimposed on the current scene image collected by the second user equipment.
- the second user holds a second user equipment, and sends a scene invocation request about the current scene to the network device through the second user equipment, and the corresponding scene invocation request includes the scene identification information of the current scene, and the network device matches the scene according to the scene identification information.
- the network device sends a plurality of target scene information to the second user device in advance, and the second user device can determine the user selection or automatic matching from the multiple target scene information based on the user selection operation or the current scene image captured by the second user device.
- the target scene information, and then the mark information added by the corresponding user of the first user equipment in the target scene information is presented through the display device of the second user equipment.
- the second user can also edit the tag information presented by the display device of the second user equipment, such as adding, deleting, modifying, and replacing tag information Or move the position of the tag information, etc., which are only examples here, and are not limited.
- the method further includes step S107 (not shown), in step S107, receiving a scene calling request about the target scene information sent by the corresponding second user equipment; responding to the scene calling request , sending the target scene information to the second user equipment, where the target scene information is used to superimpose the marker information on the current scene image collected by the second user equipment.
- the first user equipment directly establishes a communication connection with the second user equipment, without data storage and transmission through a network device.
- the second user equipment sends to the first user equipment a scene invocation request about the current scene, and the corresponding scene invocation request includes the scene identification information of the current scene, and the first user equipment determines the target scene information of the current scene according to the scene identification information, and sends The target scene information is sent to the second user equipment.
- the first user equipment shares a plurality of target scene information with the second user equipment in advance, and the second user equipment may determine from the multiple target scene information based on the user selection operation or the current scene image captured by the second user equipment or automatically
- the matching target scene information, the tag information in the target scene information is presented by the display device of the second user equipment.
- the method further includes step S108 (not shown).
- step S108 the camera of the first user equipment continues to capture subsequent scene images related to the current scene, based on the target
- the scene information superimposes and presents the mark information on a subsequent scene image corresponding to the current scene.
- the corresponding target scene information can also be used by the first user equipment to continue shooting the current scene and perform subsequent updates. For example, after the first user equipment updates the target scene information based on the target tag information, it continues to add other Mark information or edit existing mark information to continue to update target scene information, or the first user equipment reloads target scene information in a subsequent period of time, and adds marks or edits existing marks based on the target scene information.
- the first user equipment continues to shoot subsequent scene images about the current scene through the camera device, and superimposes and presents existing marker information in the subsequent scene images, such as real-time information based on the existing marker target space position, real-time pose information of the camera device, etc.
- existing marker information such as real-time information based on the existing marker target space position, real-time pose information of the camera device, etc.
- the corresponding real-time image position in the subsequent scene image is calculated for presentation.
- Fig. 2 shows a method for presenting target mark information according to another aspect of the present application, which is applied to a second user equipment, wherein the method includes step S201, step S202 and step S203.
- step S201 the target scene information matching the current scene is acquired, wherein the target scene information includes corresponding target tag information, and the target tag information includes corresponding tag information and target spatial position;
- step S202 by The imaging device shoots a current scene image about the current scene;
- step S203 the corresponding target image position is determined according to the current pose information of the imaging device and the spatial position of the target, and in the current scene image The target image position is superimposed to present the tag information.
- the second user holds the second user equipment, and the second user equipment may download one or more target scene information locally through the second user equipment by sending a scene call request to the network device or the first user equipment.
- the camera device of the second user equipment scans the current scene to obtain the current scene image, and performs point cloud initialization through three-dimensional tracking (such as SLAM), and the current world coordinates The system is aligned with the world coordinate system of the 3D point cloud information in the target scene information, so that the target space position of the marker information is aligned with the current scene.
- SLAM three-dimensional tracking
- the second user obtained through three-dimensional tracking
- the real-time pose information of the camera device of the device calculates in real time the corresponding target image position of the marker information in the current scene image, and the marker information is superimposed and displayed at the corresponding position on the display screen of the second user equipment.
- the marker information included in the target scene information Information reproduced.
- the second user equipment obtains the target scene information that matches the current scene. In some cases, it can be obtained through manual selection by the second user, and in other cases, it can be obtained through automatic matching.
- Scene image determine the matching target scene information according to the scene image, such as determining the matching target scene information through the 2D recognition initialization of three-dimensional tracking (such as a 2D recognition map in the scene); and for example, through the point cloud initialization of three-dimensional tracking, Determine the matching target scene information.
- the above-mentioned second user equipment acquires target scene information that matches the current scene, and may also determine it based on relevant information provided by the second user equipment, the relevant information includes but not limited to: location information of the second user equipment , such as GPS location information, wifi radio frequency fingerprint information, etc., the identity information of the user of the second user device, the permission information of the user of the second user device, such as the user's access permission, viewing and editing permission, etc., and the tasks being performed by the user of the second user device information etc.
- location information of the second user equipment such as GPS location information, wifi radio frequency fingerprint information, etc.
- the identity information of the user of the second user device such as the identity information of the user of the second user device
- the permission information of the user of the second user device such as the user's access permission, viewing and editing permission, etc.
- the above methods of acquiring target scene information matching the current scene can be used in combination, such as downloading the corresponding target scene information through the relevant information provided by the second user equipment or screening the downloaded target scene information Matching target scene information, and then determine the matching target scene through manual selection or automatic matching.
- downloading the corresponding target scene information through the relevant information provided by the second user equipment or screening the downloaded target scene information Matching target scene information and then determine the matching target scene through manual selection or automatic matching.
- the point cloud initialization process of SLAM includes: 1) judging the similarity between the scene image of the current scene and the point cloud in one or more target scene information, and selecting the point cloud with the highest similarity (such as using BOW( bag of words) model calculates the number of matching points between the feature points of the current image frame and the feature points of multiple key frames in the map point cloud in one or more groups of target scene information, and selects the group with the largest number of matching feature points) to load, Execute SLAM relocation and complete initialization. Among them, the relocation process is as follows:
- the SLAM system extracts ORB features from the acquired current image frame, and uses the BOW (bag of words) model to calculate the matching points between the feature points of the current image frame and multiple key frames in the map point cloud; the matching number meets a certain threshold and is considered are candidate keyframes.
- the current frame camera pose is estimated by ransac (random sampling consensus algorithm) and PNP, and then the estimated interior point (an interior point is a point suitable for the estimation model in the ransac algorithm) is updated as a map point, and then Use the graph optimization theory to optimize the camera pose of the current frame. If there are fewer interior points after optimization, repeat the above process to perform more matches on the map points of the selected candidate keyframes, and finally optimize the pose and interior points again.
- a certain threshold is met, the relocation is successful, thereby establishing a SLAM coordinate system consistent with the coordinate system in the map point cloud (obtaining the camera pose), and completing the initialization.
- the generation process of the target scene information is the same as or similar to the generation process of the target scene information in FIG. 1 , and will not be repeated here.
- the method further includes step S204 (not shown).
- step S204 an editing operation of the corresponding user on the target scene information is obtained, and the target scene information is updated based on the editing operation.
- editing operations include but are not limited to editing target tag information in target scene information, such as modifying, adding, replacing, and deleting tag information, or moving the location of tag information.
- the second user equipment presents the tag information added in the target scene information on the current scene image through the display device. After the tag information is reproduced, the second user can perform editing operations in the current scene. Specific examples are as follows: 1) The second user adds marking information to the current scene image on the display device, and can adjust the content, position, size and angle of the marking information, etc.;
- the second user can select the tag information displayed on the display device, such as the tag information that has been added in the target scene information or the tag information newly added by the second user, and a deletion option appears for deletion.
- the second user can check the tag information displayed on the display device, such as clicking on the PDF tag to open the PDF for viewing; another example is to check the real-time sensing information of the temperature sensor;
- the second user can perform corresponding operations on the tag information added in the current scene, such as calling the shortcut function of the application. After the second user clicks the tag of the application calling information, the application can be started and the corresponding shortcut function can be called; another example is filling in Form, after the second user clicks the mark of the form information, the form can be filled out.
- the editing operation includes updating the 3D point cloud information in the target scene information, for example, by moving the camera device of the second user equipment, and updating the 3D point cloud information corresponding to the target scene information through a three-dimensional tracking algorithm.
- the editing operation of the target scene by the second user is synchronized to the cloud server, and the corresponding scene database is updated.
- the editing operation of the target scene by the second user overwrites the previous target scene information.
- the editing operation of the target scene is additionally saved as a new target scene information or the like.
- the method further includes step S205 (not shown).
- step S205 the updated target scene information is sent to the network device to update the corresponding scene database.
- the tag information includes real-time sensing information, and the real-time sensing information is used to indicate real-time sensing data of a corresponding sensing device that exists between the sensing device and the second user equipment. Communication connection; wherein, the method also includes step S207 (not shown), in step S207, obtain the real-time sensing data corresponding to the sensing device; wherein, the target image position in the current scene image Superimposing and presenting the tag information includes: superimposing and presenting the real-time sensor data at a target image position of the current scene image.
- real-time sensing data includes real-time collection of any object or process that needs to be monitored, connected, and interacted with through various devices and technologies such as information sensors, radio frequency identification technology, global positioning systems, infrared sensors, and laser scanners. Sound, light, heat, electricity, mechanics, chemistry, biology, location and other necessary information.
- a communication connection is established between the corresponding sensing device and the second user equipment through wired, wireless or network equipment, so that based on the communication connection, the second user equipment can obtain real-time sensing information of the sensing device, etc., and the sensing device sets Since the target object is used to collect real-time sensing data of the target object (such as other equipment or objects, etc.), the real-time sensing information can be updated based on the currently collected real-time image, or the real-time sensing information can be updated based on a predetermined time interval renew.
- the second user device may present one or more mark information of one or more real-time sensing information added to a certain target object in the scene image, and by updating the real-time sensing information, the second user may View the real-time sensor information of the target object to understand the device status of the target object.
- the tag information includes application calling information, and the application calling information is used to call the target application in the current device; wherein, the method further includes step S206 (not shown), in step S206, Acquiring a trigger operation corresponding to the user's application invocation information in the current scene image, and invoking the target application in the second user equipment based on the trigger operation.
- step S206 Acquiring a trigger operation corresponding to the user's application invocation information in the current scene image, and invoking the target application in the second user equipment based on the trigger operation.
- multiple applications are currently installed on the second user equipment, and the application calling information is used to call one of the applications currently installed on the second user equipment, for example, after starting the corresponding application and executing related shortcut instructions, such as starting a phone application, and launching Call to Zhang XX, etc.
- the application invocation information is used to prompt the user to install the corresponding application and perform invocation after the installation is completed.
- the third user equipment is in the current scene, and the third user equipment is connected to the second user equipment through wired, wireless or network equipment. Based on the corresponding communication connection, the second user equipment may send an instruction corresponding to the application invocation information to the third user equipment, so that the third user equipment invokes a relevant application to perform a corresponding operation and the like.
- the second user device is augmented reality glasses
- the third user device is an operating device on the workbench in the current scene
- the third user device is installed with an operation application for operating the workpiece
- the second user device can view the current The identification information of the application invocation information that has been added in the real-time scene.
- the application invocation information corresponding to the operation application of the third user device is currently presented in the real-time scene image.
- the trigger operation of the second user on the application invocation information is obtained , the third user equipment invokes the corresponding operating equipment to process the workpiece and so on.
- the second user may also add identification information about the application calling information used by the operating application of the third user equipment to the real-time scene information.
- the trigger operation includes but not limited to click, touch, gesture instruction, voice instruction, button, head movement instruction and so on.
- FIG. 3 shows a method for determining and presenting target marker information according to an aspect of the present application, wherein the method includes:
- the first user equipment captures an initial scene image about the current scene by a camera device, and performs three-dimensional tracking initialization on the current scene according to the initial scene image;
- the first user equipment captures real-time scene images of the current scene through the camera device, obtains real-time pose information of the camera device and 3D point cloud information corresponding to the current scene through three-dimensional tracking, and obtains The tag information input in the real-time scene image of the current scene and the target image position corresponding to the tag information;
- the first user equipment determines a corresponding target spatial position according to the real-time pose information and the target image position, and generates corresponding target marker information according to the target spatial position and the marker information, wherein the target marker The information is used to superimpose and present the marker information at the target space position of the 3D point cloud information;
- the second user equipment acquires the target scene information that matches the current scene, where the target scene information includes corresponding target tag information, and the target tag information includes corresponding tag information and a target spatial position;
- the second user equipment captures a current scene image related to the current scene by the camera device
- the second user equipment determines a corresponding target image position according to the current pose information of the camera device and the target spatial position, and superimposes and presents the marker information on the target image position of the current scene image.
- the process of determining and presenting the target marker information is the same as or similar to the process of determining the target marker information in FIG. 1 and presenting the target marker information in FIG. 2 , and will not be repeated here.
- the above mainly introduces various embodiments of a method for determining and presenting target marker information of the present application.
- the present application also provides specific devices capable of implementing the above embodiments, which will be introduced below with reference to FIGS. 4 and 5 .
- FIG. 4 shows a first user equipment 100 for determining target tag information according to an aspect of the present application, wherein the equipment includes a one-module 101 , a two-module 102 and a three-module 103 .
- - module 101 used to take an initial scene image about the current scene through the camera device, and perform three-dimensional tracking initialization on the current scene according to the initial scene image
- module 102 used to take pictures about the current scene through the camera device
- the real-time scene image of the current scene, the real-time pose information of the camera device and the 3D point cloud information corresponding to the current scene are obtained through three-dimensional tracking, and the mark information and the mark information input by the user in the real-time scene image of the current scene are obtained.
- the target image position corresponding to the tag information a module 103, configured to determine the corresponding target spatial position according to the real-time pose information and the target image position, and generate a corresponding target position according to the target spatial position and the tag information
- the target mark information wherein the target mark information is used to superimpose and present the mark information on the target space position of the 3D point cloud information.
- the tag information includes, but is not limited to: identification information; file information; form information; application call information; real-time sensor information.
- the application calling information is used to call the first target application installed in the current device.
- the application invocation information is used to invoke a second target application installed in a corresponding third user equipment, where there is a communication connection between the third user equipment and the first user equipment.
- the tag information includes real-time sensing information, the real-time sensing information is used to indicate the real-time sensing data of the corresponding sensing device, and there is a communication connection.
- step S101, step S102 and step S103 shown in FIG. 4 are the same or similar to the embodiments of step S101, step S102 and step S103 shown in FIG. , so no further details are included here by reference.
- the device further includes a module (not shown), configured to receive the real-time sensing information returned by the network device and acquired by the sensing device; wherein, according to the target Generating corresponding target marker information based on the spatial position and the marker information includes: generating corresponding target marker information according to the target spatial position and the real-time sensing information, wherein the target marker information is used at the 3D point The target spatial position of the cloud information is superimposed to present the real-time sensor information.
- a module not shown, configured to receive the real-time sensing information returned by the network device and acquired by the sensing device; wherein, according to the target Generating corresponding target marker information based on the spatial position and the marker information includes: generating corresponding target marker information according to the target spatial position and the real-time sensing information, wherein the target marker information is used at the 3D point The target spatial position of the cloud information is superimposed to present the real-time sensor information.
- the determining the corresponding target spatial position according to the real-time pose information and the target image position includes: mapping the target image position to space according to the real-time pose information to determine the corresponding 3D straight line: determine the target spatial position corresponding to the target point according to the 3D point cloud information and the 3D straight line.
- the 3D point cloud information includes a plurality of feature points, and each feature point includes corresponding depth information; wherein, according to the 3D point cloud information and the 3D straight line, the The target spatial position includes: determining at least one target feature point from the 3D point cloud information according to the distance between each feature point in the 3D point cloud information and the 3D straight line; based on the depth of the at least one target feature point The information determines the depth information of the target point on the 3D straight line, so as to determine the corresponding target spatial position.
- the at least one target feature point includes a feature point with the smallest distance to the 3D straight line.
- the distance between each target feature point of the at least one target feature point and the 3D straight line is less than or equal to a distance threshold; wherein, the determination based on the depth information of the at least one target feature point
- the depth information of the target point on the 3D straight line includes: determining the weight information of each target feature point according to the distance information between the at least one target feature point and the 3D straight line; The depth information and the weight information determine the depth information of the target point on the 3D straight line.
- the device further includes a module (not shown), configured to update target scene information based on the target tag information, wherein the target scene information is stored in a scene database, and the scene database includes One or more target scene information, each target scene information includes target marker information and corresponding 3D point cloud. In some implementations, each target scene information also includes corresponding device parameter information.
- the device further includes a module (not shown), configured to send the target scene information to a corresponding network device, wherein the target scene information is based on a scene call request of the second user equipment The target scene information is sent to the second user equipment via the network device, and the target scene information is used to superimpose the marker information on the current scene image collected by the second user equipment.
- the device further includes a module (not shown), configured to receive a scene call request about the target scene information sent by the corresponding second user equipment; in response to the scene call request, the The target scene information is sent to the second user equipment, where the target scene information is used to superimpose the marker information on the current scene image collected by the second user equipment.
- the device further includes a module (not shown), configured to continue shooting subsequent scene images about the current scene through the camera device of the first user equipment, based on the target scene information Superimposing and presenting the mark information on a subsequent scene image corresponding to the current scene.
- step S104 to step S108 the specific implementation manners corresponding to the first four modules to one eighth modules are the same as or similar to the above-mentioned embodiments of step S104 to step S108, so they are not repeated here, and are included here by reference.
- FIG. 5 shows a second user equipment 200 for presenting target marking information according to another aspect of the present application, wherein the equipment includes a two-one module 201 , a two-two module 202 and a two-three module 203 .
- the two-one module 201 is configured to acquire target scene information matching the current scene, wherein the target scene information includes corresponding target tag information, and the target tag information includes corresponding tag information and target spatial position;
- two-two module 202 used to shoot the current scene image about the current scene through the camera device;
- the second and third modules 203 used to determine the corresponding target image position according to the current pose information of the camera device and the target spatial position, in The target image position of the current scene image is superimposed to present the tag information.
- step S201, step S202, and step S203 shown in FIG. 5 are the same or similar to the embodiments of step S201, step S202, and step S203 shown in FIG. 2 , so no further details are included here by reference.
- the device further includes a 24 module (not shown), configured to acquire an editing operation of a corresponding user on the target scene information, and update the target scene information based on the editing operation.
- the device further includes a 25 module (not shown), configured to send the updated target scene information to the network device to update the corresponding scene database.
- the tag information includes real-time sensing information, and the real-time sensing information is used to indicate real-time sensing data of a corresponding sensing device that exists between the sensing device and the second user equipment.
- the device further includes a twenty-seven module (not shown), used to obtain real-time sensing data corresponding to the sensing device; wherein, the target image position of the current scene image is superimposed and presented
- the marking information includes: superimposing and presenting the real-time sensor data at the target image position of the current scene image.
- the tag information includes application calling information, and the application calling information is used to call the target application in the current device; wherein, the device further includes a module (not shown) for obtaining the corresponding The user performs a trigger operation on the application invocation information in the current scene image, and invokes the target application in the second user equipment based on the trigger operation.
- the present application also provides a computer-readable storage medium, the computer-readable storage medium stores computer codes, and when the computer codes are executed, as described in any one of the preceding items The described method is carried out.
- the present application also provides a computer program product, when the computer program product is executed by a computer device, the method described in any one of the preceding items is executed.
- the present application also provides a kind of computer equipment, and described computer equipment comprises:
- processors one or more processors
- memory for storing one or more computer programs
- the one or more processors are made to implement the method as described in any one of the preceding items.
- FIG. 6 illustrates an exemplary system that may be used to implement various embodiments described in this application
- system 300 can be used as any one of the above-mentioned devices in each of the above-mentioned embodiments.
- system 300 may include one or more computer-readable media (e.g., system memory or NVM/storage device 320 ) having instructions and be coupled to and configured to execute The instructions are one or more processors (eg, processor(s) 305 ) that implement a module to perform the actions described in this application.
- processors e.g, processor(s) 305
- system control module 310 may include any suitable interface controller to provide at least one of processor(s) 305 and/or any suitable device or component in communication with system control module 310 Any suitable interface.
- the system control module 310 may include a memory controller module 330 to provide an interface to the system memory 315.
- the memory controller module 330 may be a hardware module, a software module and/or a firmware module.
- System memory 315 may be used, for example, to load and store data and/or instructions for system 300 .
- system memory 315 may include any suitable volatile memory, such as suitable DRAM.
- system memory 315 may include Double Data Rate Type Quad Synchronous Dynamic Random Access Memory (DDR4 SDRAM).
- DDR4 SDRAM Double Data Rate Type Quad Synchronous Dynamic Random Access Memory
- system control module 310 may include one or more input/output (I/O) controllers to provide interfaces to NVM/storage devices 320 and communication interface(s) 325 .
- I/O input/output
- NVM/storage 320 may be used to store data and/or instructions.
- NVM/storage 320 may include any suitable non-volatile memory (e.g., flash memory) and/or may include any suitable non-volatile storage device(s) (e.g., one or more hard drives (HDD), one or more compact disc (CD) drives, and/or one or more digital versatile disc (DVD) drives).
- suitable non-volatile memory e.g., flash memory
- suitable non-volatile storage device(s) e.g., one or more hard drives (HDD), one or more compact disc (CD) drives, and/or one or more digital versatile disc (DVD) drives.
- HDD hard drives
- CD compact disc
- DVD digital versatile disc
- NVM/storage device 320 may include a storage resource that is physically part of the device on which system 300 is installed, or it may be accessible by the device without necessarily being part of the device. For example, NVM/storage 320 may be accessed over a network via communication interface(s) 325 .
- Communication interface(s) 325 may provide an interface for system 300 to communicate over one or more networks and/or with any other suitable device.
- System 300 may communicate wirelessly with one or more components of a wireless network according to any of one or more wireless network standards and/or protocols.
- processor(s) 305 may be packaged with logic of one or more controllers of system control module 310 (eg, memory controller module 330 ).
- processor(s) 305 may be packaged with the logic of one or more controllers of the system control module 310 to form a system-in-package (SiP).
- SiP system-in-package
- at least one of the processor(s) 305 may be integrated on the same die as the logic of the one or more controllers of the system control module 310 .
- at least one of the processor(s) 305 may be integrated on the same die with the logic of the one or more controllers of the system control module 310 to form a system on chip (SoC).
- SoC system on chip
- system 300 may be, but is not limited to, a server, workstation, desktop computing device, or mobile computing device (eg, laptop computing device, handheld computing device, tablet computer, netbook, etc.). In various embodiments, system 300 may have more or fewer components and/or a different architecture. For example, in some embodiments, system 300 includes one or more cameras, a keyboard, a liquid crystal display (LCD) screen (including a touchscreen display), non-volatile memory ports, multiple antennas, graphics chips, application-specific integrated circuits ( ASIC) and speakers.
- LCD liquid crystal display
- ASIC application-specific integrated circuits
- the present application can be implemented in software and/or a combination of software and hardware, for example, it can be implemented by using an application specific integrated circuit (ASIC), a general-purpose computer or any other similar hardware devices.
- ASIC application specific integrated circuit
- the software program of the present application can be executed by a processor to realize the steps or functions described above.
- the software program (including associated data structures) of the present application can be stored in a computer-readable recording medium such as RAM memory, magnetic or optical drive or floppy disk and the like.
- some steps or functions of the present application may be implemented by hardware, for example, as a circuit that cooperates with a processor to execute each step or function.
- a part of the present application can be applied as a computer program product, such as a computer program instruction.
- a computer program product such as a computer program instruction.
- the method and/or technical solution according to the present application can be invoked or provided through the operation of the computer.
- computer program instructions exist in computer-readable media in forms including but not limited to source files, executable files, installation package files, etc. Limited to: the computer directly executes the instruction, or the computer compiles the instruction and then executes the corresponding compiled program, or the computer reads and executes the instruction, or the computer reads and installs the instruction and then executes the corresponding post-installation program program.
- a computer readable medium may be any available computer readable storage medium or communication medium that can be accessed by a computer.
- Communication media includes the media whereby communication signals embodying, for example, computer readable instructions, data structures, program modules or other data are transmitted from one system to another.
- Communication media can include guided transmission media such as cables and wires (e.g., fiber optics, coaxial, etc.) and wireless (unguided transmission) media capable of propagating waves of energy, such as acoustic, electromagnetic, RF, microwave, and infrared .
- Computer readable instructions, data structures, program modules or other data may be embodied, for example, as a modulated data signal in a wireless medium such as a carrier wave or similar mechanism such as embodied as part of spread spectrum technology.
- modulated data signal means a signal that has one or more of its characteristics changed or set in such a manner as to encode information in the signal. Modulation can be analog, digital or mixed modulation techniques.
- computer-readable storage media may include volatile and nonvolatile, volatile, volatile, or Removable and non-removable media.
- computer-readable storage media include, but are not limited to, volatile memories such as random access memories (RAM, DRAM, SRAM); and nonvolatile memories such as flash memory, various read-only memories (ROM, PROM, EPROM) , EEPROM), magnetic and ferromagnetic/ferroelectric memory (MRAM, FeRAM); and magnetic and optical storage devices (hard disks, tapes, CDs, DVDs); or other media known now or developed in the future capable of storing data for computer systems Computer readable information/data used.
- volatile memories such as random access memories (RAM, DRAM, SRAM
- nonvolatile memories such as flash memory, various read-only memories (ROM, PROM, EPROM) , EEPROM), magnetic and ferromagnetic/ferroelectric memory (MRAM, FeRAM); and magnetic and optical storage devices (hard disks, tapes, CDs, DVDs); or other media known now or developed
- an embodiment according to the present application includes an apparatus comprising a memory for storing computer program instructions and a processor for executing the program instructions, wherein when the computer program instructions are executed by the processor, triggering
- the operation of the device is based on the foregoing methods and/or technical solutions according to multiple embodiments of the present application.
Landscapes
- Engineering & Computer Science (AREA)
- Theoretical Computer Science (AREA)
- Physics & Mathematics (AREA)
- General Physics & Mathematics (AREA)
- General Engineering & Computer Science (AREA)
- Computer Vision & Pattern Recognition (AREA)
- Data Mining & Analysis (AREA)
- Multimedia (AREA)
- Bioinformatics & Cheminformatics (AREA)
- Bioinformatics & Computational Biology (AREA)
- Artificial Intelligence (AREA)
- Evolutionary Biology (AREA)
- Evolutionary Computation (AREA)
- Life Sciences & Earth Sciences (AREA)
- Human Computer Interaction (AREA)
- Computer Graphics (AREA)
- Computer Hardware Design (AREA)
- Software Systems (AREA)
- User Interface Of Digital Computer (AREA)
Abstract
Description
Claims (26)
- 一种确定目标标记信息的方法,应用于第一用户设备,其中,该方法包括:通过摄像装置拍摄关于当前场景的初始场景图像,根据所述初始场景图像对所述当前场景进行三维跟踪初始化;通过所述摄像装置拍摄关于所述当前场景的实时场景图像,通过三维跟踪获取所述摄像装置的实时位姿信息及所述当前场景对应的3D点云信息,获取用户在所述当前场景的实时场景图像中输入的标记信息及所述标记信息对应的目标图像位置;根据所述实时位姿信息及所述目标图像位置确定对应的目标空间位置,根据所述目标空间位置及所述标记信息生成对应的目标标记信息,其中,所述目标标记信息用于在所述3D点云信息的目标空间位置叠加呈现所述标记信息。
- 根据权利要求1所述的方法,其中,所述标记信息包括以下至少任一项:标识信息;文件信息;表单信息;应用调用信息;实时传感信息。
- 根据权利要求2所述的方法,其中,所述应用调用信息用于调用当前设备中已安装的第一目标应用。
- 根据权利要求2所述的方法,其中,所述应用调用信息用于调用对应第三用户设备中已安装的第二目标应用,其中,所述第三用户设备与所述第一用户设备之间存在通信连接。
- 根据权利要求2所述的方法,其中,所述标记信息包括实时传感信息,所述实时传感信息用于指示对应传感装置的实时传感数据,所述传感装置与所述第一用户设备之间存在通信连接。
- 根据权利要求5所述的方法,其中,所述方法还包括:向对应网络设备发送关于所述传感装置的数据获取请求;接收所述网络设备返回的、所述传感装置获取的实时传感信息;其中,所述根据所述目标空间位置及所述标记信息生成对应的目标标记信息,包括:根据所述目标空间位置及所述实时传感信息生成对应的目标标记信息,其中,所述目标标记信息用于在所述3D点云信息的目标空间位置叠加呈现所述实时传感信息。
- 根据权利要求1所述的方法,其中,所述根据所述实时位姿信息及所述目标图像位置确定对应的目标空间位置,包括:根据所述实时位姿信息将所述目标图像位置映射至空间中确定对应的3D直线;根据所述3D点云信息及所述3D直线确定目标点位对应的目标空间位置。
- 根据权利要求7所述的方法,其中,所述3D点云信息包括多个特征点,每个特征点包含对应的深度信息;其中,所述根据所述3D点云信息及所述3D直线确定目标点位对应的目标空间位置,包括:根据所述3D点云信息中各个特征点与所述3D直线的距离,从所述3D点云信息中确定至少一个目标特征点;基于所述至少一个目标特征点的深度信息确定所述3D直线上的目标点位的深度信息,从而确定对应的目标空间位置。
- 根据权利要求8所述的方法,其中,所述至少一个目标特征点包括与所述3D直线的距离最小的特征点。
- 根据权利要求8所述的方法,其中,所述至少一个目标特征点中每个目标特征点与所述3D直线的距离小于或等于距离阈值;其中,所述基于所述至少一个目标特征点的深度信息确定所述3D直线上的目标点位的深度信息,包括:根据所述至少一个目标特征点与所述3D直线的距离信息确定每个目标特征点的权重信息;基于所述每个目标特征点的深度信息、权重信息确定所述3D直线上的目标点位的深度信息。
- 根据权利要求1所述的方法,其中,所述方法还包括:基于所述目标标记信息更新目标场景信息,其中,所述目标场景信息存储于场景数据库,所述场景数据库包括一个或多个目标场景信息,每个目标场景信息包括目标标记信息及对应的3D点云信息。
- 根据权利要求11所述的方法,其中,所述每个目标场景信息还包括对应的设备参数信息。
- 根据权利要求11或12所述的方法,其中,所述方法还包括:将所述目标场景信息发送至对应网络设备,其中,所述目标场景信息基于第二用户设备的场景调用请求经由所述网络设备发送至所述第二用户设备,所述目标场景信息用于在所述第二用户设备采集的当前场景图像中叠加所述标记信息。
- 根据权利要求11或12所述的方法,其中,所述方法还包括:接收对应第二用户设备发送的关于所述目标场景信息的场景调用请求;响应于所述场景调用请求,将所述目标场景信息发送至所述第二用户设备,其中,所述目标场景信息用于在所述第二用户设备采集的当前场景图像中叠加所述标记信息。
- 根据权利要求11或12所述的方法,其中,所述方法还包括:通过所述第一用户设备的摄像装置继续拍摄关于所述当前场景的后续场景图像,基于所述目标场景信息将所述标记信息叠加呈现于所述当前场景对应的后续场景图像。
- 一种呈现目标标记信息的方法,应用于第二用户设备,其中,该方法包括:获取与当前场景相匹配的目标场景信息,其中,所述目标场景信息包括对应的目标标记信息,所述目标标记信息包括对应标记信息及目标空间位置;通过所述摄像装置拍摄关于所述当前场景的当前场景图像;根据所述摄像装置的当前位姿信息及所述目标空间位置确定对应的目标图像位置,在所述当前场景图像的目标图像位置叠加呈现所述标记信息。
- 根据权利要求16所述的方法,其中,所述方法还包括:获取对应用户关于所述目标场景信息的编辑操作,基于所述编辑操作更新所述目标场景信息。
- 根据权利要求17所述的方法,其中,所述方法还包括:将更新后的目标场景信息发送至网络设备以更新对应场景数据库。
- 根据权利要求16所述的方法,其中,所述标记信息包括应用调用信息,所述应用调用信息用于调用当前设备中的目标应用;其中,所述方法还包括:获取对应用户关于所述当前场景图像中应用调用信息的触发操作,基于所述触发操作调用所述第二用户设备中的目标应用。
- 根据权利要求16所述的方法,其中,所述标记信息包括实时传感信息,所述实时传感信息用于指示对应传感装置的实时传感数据,所述传感装置与所述第二用户设备之间存在通信连接;其中,所述方法还包括:获取所述传感装置对应的实时传感数据;其中,所述在所述当前场景图像的目标图像位置叠加呈现所述标记信息,包括:在所述当前场景图像的目标图像位置叠加呈现所述实时传感数据。
- 一种确定并呈现目标标记信息的方法,其中,所述方法包括:第一用户设备通过摄像装置拍摄关于当前场景的初始场景图像,根据所述初始场景图像对所述当前场景进行三维跟踪初始化;所述第一用户设备通过所述摄像装置拍摄关于所述当前场景的实时场景图像,通过三维跟踪获取所述摄像装置的实时位姿信息及所述当前场景对应的3D点云信息,获取用户在所述当前场景的实时场景图像中输入的标记信息及所述标记信息对应的目标图像位置;所述第一用户设备根据所述实时位姿信息及所述目标图像位置确定对应的目标空间位置,根据所述目标空间位置及所述标记信息生成对应的目标标记信息,其中,所述目标标记信息用于在所述3D点云信息的目标空间位置叠加呈现所述标记信息;所述第二用户设备获取与所述当前场景相匹配的所述目标场景信息,其中,所述目标场景信息包括对应的目标标记信息,所述目标标记信息包括对应标记信息及目标空间位置;所述第二用户设备通过所述摄像装置拍摄关于所述当前场景的当前场景图像;所述第二用户设备根据所述摄像装置的当前位姿信息及所述目标空间位置确定对应的目标图像位置,在所述当前场景图像的目标图像位置叠加呈现所述标记信息。
- 一种确定目标标记信息的第一用户设备,其中,该设备包括:一一模块,用于通过摄像装置拍摄关于当前场景的初始场景图像,根据所述初始场景图像对所述当前场景进行三维跟踪初始化;一二模块,用于通过所述摄像装置拍摄关于所述当前场景的实时场景图像,通过三维跟踪获取所述摄像装置的实时位姿信息及所述当前场景对应的3D点云信息,获取用户在所述当前场景的实时场景图像中输入的标记信息及所述标记信息对应的目标图像位置;一三模块,用于根据所述实时位姿信息及所述目标图像位置确定对应的目标空间位置,根据所述目标空间位置及所述标记信息生成对应的目标标记信息,其中,所述目标标记信息用于在所述3D点云信息的目标空间位置叠加呈现所述标记信息。
- 一种呈现目标标记信息的第二用户设备,其中,该设备包括:二一模块,用于获取与当前场景相匹配的目标场景信息,其中,所述目标场景信息包括对应的目标标记信息,所述目标标记信息包括对应标记信息及目标空间位置;二二模块,用于通过所述摄像装置拍摄关于所述当前场景的当前场景图像;二三模块,用于根据所述摄像装置的当前位姿信息及所述目标空间位置确定对应的目标图像位置,在所述当前场景图像的目标图像位置叠加呈现所述标记信息。
- 一种计算机设备,其中,该设备包括:处理器;以及被安排成存储计算机可执行指令的存储器,所述可执行指令在被执行时使所述处理器执行如权利要求1至20中任一项所述方法的步骤。
- 一种计算机可读存储介质,其上存储有计算机程序/指令,其特征在于,该计算机程序/指令在被执行时使得系统进行执行如权利要求1至20中任一项所述方法的步骤。
- 一种计算机程序产品,包括计算机程序/指令,其特征在于,该计算机程序/指令被处理器执行时实现权利要求1至20中任一项所述方法的步骤。
Priority Applications (1)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| EP22866303.5A EP4394554A4 (en) | 2021-09-09 | 2022-08-05 | METHOD FOR DETERMINING AND DISPLAYING TARGET MARKING INFORMATION AND DEVICE |
Applications Claiming Priority (2)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| CN202111056350.6 | 2021-09-09 | ||
| CN202111056350.6A CN113741698B (zh) | 2021-09-09 | 2021-09-09 | 一种确定和呈现目标标记信息的方法与设备 |
Publications (1)
| Publication Number | Publication Date |
|---|---|
| WO2023035829A1 true WO2023035829A1 (zh) | 2023-03-16 |
Family
ID=78737547
Family Applications (1)
| Application Number | Title | Priority Date | Filing Date |
|---|---|---|---|
| PCT/CN2022/110472 Ceased WO2023035829A1 (zh) | 2021-09-09 | 2022-08-05 | 一种确定和呈现目标标记信息的方法与设备 |
Country Status (3)
| Country | Link |
|---|---|
| EP (1) | EP4394554A4 (zh) |
| CN (1) | CN113741698B (zh) |
| WO (1) | WO2023035829A1 (zh) |
Cited By (3)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| CN116363331A (zh) * | 2023-04-03 | 2023-06-30 | 北京百度网讯科技有限公司 | 图像生成方法、装置、设备以及存储介质 |
| CN118587391A (zh) * | 2024-05-23 | 2024-09-03 | 中交一公局绿建(厦门)科技有限公司 | 一种基于slam与ar的建筑信息模型管理系统 |
| CN118819280A (zh) * | 2024-03-04 | 2024-10-22 | 中移动金融科技有限公司 | 元宇宙空间分享方法、装置、设备、存储介质及产品 |
Families Citing this family (15)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| CN113741698B (zh) * | 2021-09-09 | 2023-12-15 | 亮风台(上海)信息科技有限公司 | 一种确定和呈现目标标记信息的方法与设备 |
| CN114332417B (zh) * | 2021-12-13 | 2023-07-14 | 亮风台(上海)信息科技有限公司 | 一种多人场景交互的方法、设备、存储介质及程序产品 |
| CN114089836B (zh) * | 2022-01-20 | 2023-02-28 | 中兴通讯股份有限公司 | 标注方法、终端、服务器和存储介质 |
| US12307771B2 (en) * | 2022-02-18 | 2025-05-20 | Omnivision Technologies, Inc. | Image processing method and apparatus implementing the same |
| WO2023168836A1 (zh) * | 2022-03-11 | 2023-09-14 | 亮风台(上海)信息科技有限公司 | 一种投影交互方法、设备、介质及程序产品 |
| CN115100649B (zh) * | 2022-05-06 | 2025-01-07 | 广东虚拟现实科技有限公司 | 标记图案的生成方法、装置、电子设备及存储介质 |
| CN115439635B (zh) * | 2022-06-30 | 2024-04-26 | 亮风台(上海)信息科技有限公司 | 一种呈现目标对象的标记信息的方法与设备 |
| CN115460539B (zh) * | 2022-06-30 | 2023-12-15 | 亮风台(上海)信息科技有限公司 | 一种获取电子围栏的方法、设备、介质及程序产品 |
| CN117768627B (zh) * | 2022-09-16 | 2026-01-09 | 华为技术有限公司 | 一种增强现实方法和计算装置 |
| CN115268658A (zh) * | 2022-09-30 | 2022-11-01 | 苏芯物联技术(南京)有限公司 | 一种基于增强现实的多方远程空间圈画标记方法 |
| CN118509419B (zh) * | 2023-02-14 | 2025-11-11 | 广州视源电子科技股份有限公司 | 标记信息的同步方法、装置、计算机设备及远程协作系统 |
| CN116431880A (zh) * | 2023-04-20 | 2023-07-14 | 亮风台(上海)信息科技有限公司 | 一种用于呈现记录的方法、设备及介质 |
| CN116664806A (zh) * | 2023-06-07 | 2023-08-29 | 亮风台(上海)信息科技有限公司 | 一种用于呈现增强现实数据的方法、设备与介质 |
| CN117369633B (zh) * | 2023-10-07 | 2024-08-23 | 九转棱镜(北京)科技有限公司 | 一种基于ar的信息交互方法及系统 |
| CN117745988B (zh) * | 2023-12-20 | 2024-09-13 | 亮风台(上海)信息科技有限公司 | 一种用于呈现ar标签信息的方法与设备 |
Citations (5)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| CN108830894A (zh) * | 2018-06-19 | 2018-11-16 | 亮风台(上海)信息科技有限公司 | 基于增强现实的远程指导方法、装置、终端和存储介质 |
| CN109669541A (zh) * | 2018-09-04 | 2019-04-23 | 亮风台(上海)信息科技有限公司 | 一种用于配置增强现实内容的方法与设备 |
| CN111709973A (zh) * | 2020-06-16 | 2020-09-25 | 北京百度网讯科技有限公司 | 目标跟踪方法、装置、设备及存储介质 |
| CN112907671A (zh) * | 2021-03-31 | 2021-06-04 | 深圳市慧鲤科技有限公司 | 点云数据生成方法、装置、电子设备及存储介质 |
| CN113741698A (zh) * | 2021-09-09 | 2021-12-03 | 亮风台(上海)信息科技有限公司 | 一种确定和呈现目标标记信息的方法与设备 |
Family Cites Families (10)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| CN103529959B (zh) * | 2013-01-21 | 2016-06-29 | Tcl集团股份有限公司 | 基于关键点射线碰撞检测的框选方法、系统及电子设备 |
| US9911235B2 (en) * | 2014-11-14 | 2018-03-06 | Qualcomm Incorporated | Spatial interaction in augmented reality |
| US20160358383A1 (en) * | 2015-06-05 | 2016-12-08 | Steffen Gauglitz | Systems and methods for augmented reality-based remote collaboration |
| US20190033989A1 (en) * | 2017-07-31 | 2019-01-31 | Google Inc. | Virtual reality environment boundaries using depth sensors |
| CN111199583B (zh) * | 2018-11-16 | 2023-05-16 | 广东虚拟现实科技有限公司 | 一种虚拟内容显示方法、装置、终端设备及存储介质 |
| CN110197148B (zh) * | 2019-05-23 | 2020-12-01 | 北京三快在线科技有限公司 | 目标物体的标注方法、装置、电子设备和存储介质 |
| CN111415388B (zh) * | 2020-03-17 | 2023-10-24 | Oppo广东移动通信有限公司 | 一种视觉定位方法及终端 |
| CN111311684B (zh) * | 2020-04-01 | 2021-02-05 | 亮风台(上海)信息科技有限公司 | 一种进行slam初始化的方法与设备 |
| CN111950521A (zh) * | 2020-08-27 | 2020-11-17 | 深圳市慧鲤科技有限公司 | 一种增强现实交互的方法、装置、电子设备及存储介质 |
| CN113048980B (zh) * | 2021-03-11 | 2023-03-14 | 浙江商汤科技开发有限公司 | 位姿优化方法、装置、电子设备及存储介质 |
-
2021
- 2021-09-09 CN CN202111056350.6A patent/CN113741698B/zh active Active
-
2022
- 2022-08-05 WO PCT/CN2022/110472 patent/WO2023035829A1/zh not_active Ceased
- 2022-08-05 EP EP22866303.5A patent/EP4394554A4/en not_active Withdrawn
Patent Citations (5)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| CN108830894A (zh) * | 2018-06-19 | 2018-11-16 | 亮风台(上海)信息科技有限公司 | 基于增强现实的远程指导方法、装置、终端和存储介质 |
| CN109669541A (zh) * | 2018-09-04 | 2019-04-23 | 亮风台(上海)信息科技有限公司 | 一种用于配置增强现实内容的方法与设备 |
| CN111709973A (zh) * | 2020-06-16 | 2020-09-25 | 北京百度网讯科技有限公司 | 目标跟踪方法、装置、设备及存储介质 |
| CN112907671A (zh) * | 2021-03-31 | 2021-06-04 | 深圳市慧鲤科技有限公司 | 点云数据生成方法、装置、电子设备及存储介质 |
| CN113741698A (zh) * | 2021-09-09 | 2021-12-03 | 亮风台(上海)信息科技有限公司 | 一种确定和呈现目标标记信息的方法与设备 |
Non-Patent Citations (1)
| Title |
|---|
| See also references of EP4394554A4 |
Cited By (4)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| CN116363331A (zh) * | 2023-04-03 | 2023-06-30 | 北京百度网讯科技有限公司 | 图像生成方法、装置、设备以及存储介质 |
| CN116363331B (zh) * | 2023-04-03 | 2024-02-23 | 北京百度网讯科技有限公司 | 图像生成方法、装置、设备以及存储介质 |
| CN118819280A (zh) * | 2024-03-04 | 2024-10-22 | 中移动金融科技有限公司 | 元宇宙空间分享方法、装置、设备、存储介质及产品 |
| CN118587391A (zh) * | 2024-05-23 | 2024-09-03 | 中交一公局绿建(厦门)科技有限公司 | 一种基于slam与ar的建筑信息模型管理系统 |
Also Published As
| Publication number | Publication date |
|---|---|
| CN113741698B (zh) | 2023-12-15 |
| EP4394554A4 (en) | 2025-01-08 |
| EP4394554A1 (en) | 2024-07-03 |
| CN113741698A (zh) | 2021-12-03 |
Similar Documents
| Publication | Publication Date | Title |
|---|---|---|
| CN113741698B (zh) | 一种确定和呈现目标标记信息的方法与设备 | |
| CN111311684B (zh) | 一种进行slam初始化的方法与设备 | |
| CN111161347B (zh) | 一种进行slam初始化的方法与设备 | |
| CN109887003B (zh) | 一种用于进行三维跟踪初始化的方法与设备 | |
| JP2023504775A5 (zh) | ||
| US20110310227A1 (en) | Mobile device based content mapping for augmented reality environment | |
| WO2023109153A1 (zh) | 一种多人场景交互的方法、设备、存储介质及程序产品 | |
| CN107909612A (zh) | 一种基于3d点云的视觉即时定位与建图的方法与系统 | |
| CN107784671A (zh) | 一种用于视觉即时定位与建图的方法与系统 | |
| CN109584377B (zh) | 一种用于呈现增强现实内容的方法与设备 | |
| KR20140090078A (ko) | 이미지 처리 방법 및 그 방법을 처리하는 전자 장치 | |
| US12236537B2 (en) | Spatially aware environment relocalization | |
| WO2017181699A1 (zh) | 一种三维展现监控视频的方法及装置 | |
| KR20230049969A (ko) | 글로벌 측위 장치 및 방법 | |
| CN108681389B (zh) | 一种通过阅读设备进行阅读的方法与设备 | |
| CN116740314A (zh) | 一种用于生成增强现实数据的方法、设备及介质 | |
| CN115278084A (zh) | 图像处理方法、装置、电子设备及存储介质 | |
| CN116645493A (zh) | 一种用于呈现增强现实数据的方法、设备及介质 | |
| CN114170366B (zh) | 基于点线特征融合的三维重建方法及电子设备 | |
| CN112102145B (zh) | 图像处理方法及装置 | |
| US20170213383A1 (en) | Displaying Geographic Data on an Image Taken at an Oblique Angle | |
| CN117745988B (zh) | 一种用于呈现ar标签信息的方法与设备 | |
| WO2024250493A1 (zh) | 一种用于呈现增强现实数据的方法、设备与介质 | |
| CN116684540A (zh) | 一种用于呈现增强现实数据的方法、设备及介质 | |
| CN109636922B (zh) | 一种用于呈现增强现实内容的方法与设备 |
Legal Events
| Date | Code | Title | Description |
|---|---|---|---|
| 121 | Ep: the epo has been informed by wipo that ep was designated in this application |
Ref document number: 22866303 Country of ref document: EP Kind code of ref document: A1 |
|
| WWE | Wipo information: entry into national phase |
Ref document number: 2022866303 Country of ref document: EP |
|
| NENP | Non-entry into the national phase |
Ref country code: DE |
|
| ENP | Entry into the national phase |
Ref document number: 2022866303 Country of ref document: EP Effective date: 20240329 |
|
| WWW | Wipo information: withdrawn in national office |
Ref document number: 2022866303 Country of ref document: EP |