A head-wearable extended reality (XR) device includes a display arrangement. The display arrangement has a display to display virtual content, and also has one or more optical elements to direct the virtual content along an optical path to an eye of a user of the XR device. The virtual content is presented in a virtual content field of view. The display arrangement further includes an adjustment mechanism to alter the optical path so as to adjust the virtual content field of view between at least two display modes.
H04N 13/344 - Displays for viewing with the aid of special glasses or head-mounted displays [HMD] with head-mounted left-right displays
H04N 13/361 - Reproducing mixed stereoscopic imagesReproducing mixed monoscopic and stereoscopic images, e.g. a stereoscopic image overlay window on a monoscopic image background
H04N 13/383 - Image reproducers using viewer tracking for tracking with gaze detection, i.e. detecting the lines of sight of the viewer's eyes
2.
CONTENT COLLECTION INDICATORS WITHIN A GROUP MESSAGING SYSTEM
In one aspect, a method, includes identifying members of a group in a messaging application, detecting an active sharing of content collection from a member of the group, and generating a visual indicator corresponding to the group in a user interface of the messaging application. The method may also include further includes receiving user input from the member of the group to share the content collection with one or more members of the group, and in response to receiving the user input, enabling an ephemeral display of the content collection only from devices associated with the one or more members of the group, where detecting the active sharing of content collection from the member of the group is based on receiving the user input from the member of the group.
An AR system includes multiple input-modalities. A hand-tracking pipeline supports Direct Manipulation of Virtual Object (DMVO) and gesture input methodologies. In addition, a voice processing pipeline provides for speech inputs. Direct memory buffer access to preliminary hand-tracking data, such as skeletal models, allows for low latency communication of the data for use by DMVO-based user interfaces. A system framework component routes higher level hand-tracking data, such as gesture identification and symbols generated based on hand positions, via a Snips protocol to gesture-based user interfaces.
Systems and methods are provided for notifying users about videos in a playback sequence. The systems and methods determine that a video that meets a criterion is currently available on a video server associated with a messaging client. In response to determining that the video that meets the criterion is currently available, the messaging client on the client device prefetches a sequence of videos from a recommendation engine that match a profile of a user associated with the messaging client. The recommendation engine is being used to provide sequence of videos to a video playback graphical user interface (GUI) that automatically plays back the videos in the sequence. The systems and methods determine that the video that meets the criterion is in a first position in the sequence of videos and, in response, present a notification that indicates the availability of the video on the video playback GUI.
H04N 21/458 - Scheduling content for creating a personalised stream, e.g. by combining a locally stored advertisement with an incoming streamUpdating operations, e.g. for OS modules
H04N 21/431 - Generation of visual interfacesContent or additional data rendering
H04N 21/45 - Management operations performed by the client for facilitating the reception of or the interaction with the content or administrating data related to the end-user or to the client device itself, e.g. learning user preferences for recommending movies or resolving scheduling conflicts
H04N 21/466 - Learning process for intelligent management, e.g. learning user preferences for recommending movies
H04N 21/472 - End-user interface for requesting content, additional data or servicesEnd-user interface for interacting with content, e.g. for content reservation or setting reminders, for requesting event notification or for manipulating displayed content
A search query received from a mobile computing device includes image query data and ambient audio data captured at a location of the mobile computing device. The ambient audio data is filtered using a low-pass filter to isolate background audio data, which is compared to a set of stored audio fingerprints respectively associated with a plurality of sub-locations of a geographic location. Location-identifying text is generated based on a matching audio fingerprint. Image analysis of the image query data generates a set of image-derived keywords. An augmented search query incorporating the location-identifying text and the image-derived keywords is compiled and correlated with respective sub-location indexes to identify, as a search result, a specific sub-location of the plurality of sub-locations. The search result identifying the specific sub-location is returned to the mobile computing device.
G06F 16/48 - Retrieval characterised by using metadata, e.g. metadata not derived from the content or metadata generated manually
G06F 16/487 - Retrieval characterised by using metadata, e.g. metadata not derived from the content or metadata generated manually using geographical or spatial information, e.g. location
G06F 16/58 - Retrieval characterised by using metadata, e.g. metadata not derived from the content or metadata generated manually
G06F 16/583 - Retrieval characterised by using metadata, e.g. metadata not derived from the content or metadata generated manually using metadata automatically derived from the content
G06F 16/587 - Retrieval characterised by using metadata, e.g. metadata not derived from the content or metadata generated manually using geographical or spatial information, e.g. location
G11B 27/11 - IndexingAddressingTiming or synchronisingMeasuring tape travel by using information not detectable on the record carrier
The systems and methods display, by at least one processor on a first portion of a real-world environment visible on a display of a user device, one or more virtual objects at a first virtual coordinate in three-dimensional (3D) space. The systems and methods receive a request to move the one or more virtual objects to a second virtual coordinate in 3D space. The systems and methods in response to receiving the request, maintain display of the one or more virtual objects at the second virtual coordinate as a second portion of the real-world environment is visible on the display of the user device.
Methods and systems are disclosed for performing operations comprising: receiving a video that includes a depiction of a person wearing a fashion item; generating a segmentation of the fashion item worn by the person depicted in the video; applying one or more augmented reality elements to the fashion item worn by the person based on the segmentation of the fashion item worn by the person; detecting a gesture performed by the person in the video; and modifying the one or more augmented reality elements that have been applied to the fashion item worn by the person based on the gesture performed by the person.
An eXtended Reality (XR) device is provided that uses images of a dorsal surface of a hand of a user to authenticate the user. The XR device captures, using a set of cameras, images of a hand of a user and processes these images to generate enhanced vein patterns. The XR device generates authentication features based on the enhanced vein patterns and performs classification using the authentication features. Based on the classification, the XR device controls access to specific content and applications. The XR device includes an infrared camera and emitter to optimize vein pattern visibility, processes images using computer vision algorithms, and performs lightweight classification on-device while maintaining user privacy.
In some implementations, a system may establish a video call between a first device associated with a first user and a second device associated with a second user of a communications platform. The system may present a video interface for the video call, the video interface comprising a first video stream generated by the first device of the first user and a second video stream generated by the second device associated with the second user. The system may present a first set of image augmentations selected by the communications platform in the video interface, the first set of video augmentations being selectable by the first user for augmentation of the first video stream generated by the first user device. The system may identify a second set of image augmentations used by a further set of users of the communications platform. The system may present the second set of image augmentations in the video interface, the second set of image augmentations being selectable by the first user for augmentation of the first video stream generated by the first user device.
One aspect disclosed is a method including determining a location from a positioning system receiver, determining, using a hardware processor and the location, that the location is approaching a path of direction of visual direction information, displaying the visual direction information on a display of a wearable device in response to the determining, determining, using the positioning system receiver, whether the turn of the visual direction information has been made, determining, by the hardware processor, a first period of time for display of the content data based on whether the turn of the visual direction information has been made, powering on the display and displaying, using the display, content data for the first period of time, turning off the display and the hardware processor following display of the content data.
Examples relate to display systems and techniques for reducing visual artifacts in field sequential color displays. A system includes a field sequential color display device that presents visual content according to a first display configuration, and an eye tracking subsystem that detects rapid eye movements. The system predicts a duration of a detected rapid eye movement and temporarily modifies display parameters during the predicted duration according to a second display configuration that reduces color breakup artifacts. The second configuration can include reducing content opacity through optical filter control, adjusting display brightness or contrast, or modifying color data at region boundaries to use monochromatic colors. After the predicted duration, the system returns to presenting content according to the first display configuration. The prediction of movement duration utilizes relationships between peak velocity and acceleration profiles of eye movements to overcome eye tracking and display latency constraints.
G09G 5/02 - Control arrangements or circuits for visual indicators common to cathode-ray tube indicators and other visual indicators characterised by the way in which colour is displayed
An extended Reality (XR) system is provided that monitors neurological signals to determine an engagement of a user with a real -world environment. The XR system continuously monitors neurological signals of a user through a processor operating in a lowpower mode. The XR system generates an engagement signal by analyzing endogenous brain patterns in the neurological signals. In response to the engagement signal, the XR system activates environmental sensors to capture real-world environment data. The XR system generates contextual data from the captured environment data and determines XR content to provide to the user based on the contextual data. The XR system selectively activates XR capabilities to display the determined XR content.
An extended Reality (XR) device is provided that uses images of a dorsal surface of a hand of a user to authenticate the user. The XR device captures, using a set of cameras, images of a hand of a user and processes these images to generate enhanced vein patterns. The XR device generates authentication features based on the enhanced vein patterns and performs classification using the authentication features. Based on the classification, the XR device controls access to specific content and applications. The XR device includes an infrared camera and emitter to optimize vein pattern visibility, processes images using computer vision algorithms, and performs lightweight classification on-device while maintaining user privacy.
Methods and systems are disclosed for generating video by applying a template to various content items. The methods and systems select, by an interaction application, a video generation template comprising instructions for combining a set of content items into a video using one or more augmented reality (AR) elements. The methods and systems identify a subset of content items from a plurality of previously captured content items and modify one or more content items of the identified subset of content items based on the AR elements of the video generation template. The methods and systems generate a video comprising a collection of content items including the identified subset of content items and the modified one or more content items based on the instructions of the video generation template.
Systems and methods in the present disclosure relate to surface-based user input for extended reality (XR) devices. An XR device tracks a hand of a user across multiple captured image frames to obtain positions of the hand in a real-world environment. The XR device detects an input plane associated with a physical surface in the real-world environment and projects the positions onto the input plane. The XR device continuously monitors an input state with respect to the input plane to identify when the user is providing user input via the input plane and to differentiate between ongoing user input and finalized user input. Based on the monitoring of the input state, the XR device records one or more of the projected positions as input data. The input data is processed to interpret the user input. The XR device executes an action based on the interpreted user input.
Aspects of the present disclosure involve a system comprising a computer-readable storage medium storing a program and method for video synthesis. The program and method provide for accessing a primary generative adversarial network (GAN) comprising a pre-trained image generator, a motion generator comprising a plurality of neural networks, and a video discriminator; generating an updated GAN based on the primary GAN, by performing operations comprising identifying input data of the updated GAN, the input data comprising an initial latent code and a motion domain dataset, training the motion generator based on the input data, and adjusting weights of the plurality of neural networks of the primary GAN based on an output of the video discriminator; and generating a synthesized video based on the primary GAN and the input data.
Methods, devices, media, and other embodiments are described for a state-space system for pseudorandom animation. In one embodiment animation elements within a computer model are identified, and for each animation element motion patterns and speed harmonics are identified. A set of motion data values comprising a state-space description of the motion patterns and the speed harmonics are generated, and a probability assigned to each value of the set of motion data values for the state-space description. The probability can then be used to select and update a particular motion used in an animation generated from the computer model.
Examples relate to computer-implemented systems and methods for generating four-dimensional (4D) video content. A 4D video generation model receives a freeze-time video showing a scene varying in viewpoint and a fixed-view video showing the scene varying in time. The model processes these inputs through parallel pathways—a freeze-time pathway and a fixed-view pathway—each comprising pretrained diffusion transformer blocks. The pathways are synchronized via interleaved synchronization layers to generate a grid of video frames varying consistently in both time and viewpoint. The generated frame grid can be used to reconstruct a three-dimensional representation of the scene for real-time rendering of views from arbitrary viewpoints.
H04N 13/117 - Transformation of image signals corresponding to virtual viewpoints, e.g. spatial image interpolation the virtual viewpoint locations being selected by the viewers or determined by viewer tracking
G06T 7/70 - Determining position or orientation of objects or cameras
H04N 13/167 - Synchronising or controlling image signals
19.
ADJUSTING PERFORMANCE OF PARALLEL AUGMENTED REALITY EXPERIENCES
Systems, methods, and computer readable media for power and temperature attribution on mobile devices. Example methods include launching a native augmented reality (AR) application together with an external AR application on the user system. The example methods include determining that usage of the external AR application transgresses a usage budget for the external AR application and adjusting one or more operations of the external AR application without modifying operation of the native AR application in response to determining that the usage of the external AR application transgresses the usage budget for the external AR application.
A dynamic application theme system detects a trigger event for modifying a visual appearance of an application and, in response, causes dynamic theme assets to be downloaded to a subset of a plurality of computing devices having the application installed. The system further causes modification of a first visual appearance of multiple user interface elements of the application, on each computing device of the subset of the plurality of computing devices, by applying the downloaded dynamic theme assets to the multiple user interface elements of the application. The system maintains the modified visual appearance for a specified duration and then reverts the visual appearance of the multiple user interface elements to the first visual appearance before the modification, on each computing device of the subset of the plurality of computing devices.
Systems and methods are provided for retrieving first query result data associated with a first user account and rendering the first query result data into a first result item, generating a shareable search result stream comprising the first result item associated with the first user account, retrieving second query result data associated with a second user account and rendering the second query result data into a second result item, adding the second result item to the shareable search result stream associated with the first user account, and providing the sharable search result stream comprising the first result item and the second result item to a first computing device associated with the first user account and a second computing device associated with the second user account.
A model generation system captures, using cameras, hand-tracking data of a gesture made by a user demonstrating the gesture. The model generation system generates a three-dimensional model of the gesture using the hand-tracking data and provides a display of the three-dimensional model to the user. The model generation system receives, from the user, model refining data refining the three-dimensional model. The model generation system generates a refined three-dimensional model of the gesture using the model refining data and the three-dimensional model. The refined three-dimensional model is used for detecting the gesture.
A method for improving the startup time of a six-degrees of freedom tracking system is described. An augmented reality system receives a device initialization request and activates a first set of sensors in response to the device initialization request. The augmented reality system receives first tracking data from the first set of sensors. The augmented reality system receives an augmented reality experience request and in response to the augmented reality request, causes display of a set of augmented reality content items based on the first tracking data and simultaneously activates a second set of sensors. The augmented reality system receives second tracking data from the activated second set of sensors. The augmented reality system updates the display of the set of augmented reality content items based on the second tracking data.
G06F 3/01 - Input arrangements or combined input and output arrangements for interaction between user and computer
G01C 21/16 - NavigationNavigational instruments not provided for in groups by using measurement of speed or acceleration executed aboard the object being navigatedDead reckoning by integrating acceleration or speed, i.e. inertial navigation
The present application discloses examples of various apparatuses and systems that can be utilized for augmented reality. According to one example, a wearable device that can optionally comprise: a frame configured for wearing by a user; one or more optical elements mounted on the frame; an array having a plurality of light emitting diodes coupled to the one or more optical elements, wherein the one or more optical elements and the array are mounted within a field of view of the user when the frame is worn by the user; and additional onboard electronic components carried by the frame including at least a battery that is configured to provide for electrically powered operation of the array.
G06T 11/60 - Editing figures and textCombining figures or text
G09G 3/32 - Control arrangements or circuits, of interest only in connection with visual indicators other than cathode-ray tubes for presentation of an assembly of a number of characters, e.g. a page, by composing the assembly by combination of individual elements arranged in a matrix using controlled light sources using electroluminescent panels semiconductive, e.g. using light-emitting diodes [LED]
Aspects of the present disclosure involve a system and a method for performing operations comprising: receiving, by a messaging application implemented on a client device, input that selects a sound option to add sound to one or more images; in response to receiving the input, presenting a sound editing user interface element that visually indicates a played portion of the sound and separately visually indicates an un-played portion of the sound; receiving an interaction with the sound editing user interface element to modify a start point of the sound; embedding a graphical element representing the sound in the one or more images; playing, by the messaging application, the sound associated with the graphical element starting from the start point together with displaying the one or more images.
Techniques for training a neural network having a plurality of computational layers with associated weights and activations for computational layers in fixed-point formats include determining an optimal fractional length for weights and activations for the computational layers; training a learned clipping-level with fixed-point quantization using a PACT process for the computational layers; and quantizing on effective weights that fuses a weight of a convolution layer with a weight and running variance from a batch normalization layer. A fractional length for weights of the computational layers is determined from current values of weights using the determined optimal fractional length for the weights of the computational layers. A fixed-point activation between adjacent computational layers is related using PACT quantization of the clipping-level and an activation fractional length from a node in a following computational layer. The resulting fixed-point weights and activation values are stored as a compressed representation of the neural network.
Methods and systems are disclosed for measuring and quantifying inner speech production. The methods and systems present an instruction to a user to produce inner speech and collect electromyograph (EMG) data corresponding to the inner speech. The methods and systems process the EMG data by a machine learning model to generate a prediction comprising a word or phrase corresponding to the inner speech and concurrently generate an evaluation of inner speech production and detection.
A system to automatically increment read-watermarks based on a set of predefined rules and criteria and configured to perform operations that include: accessing a message thread that comprises a plurality of messages; detecting a display of a message from among the plurality of messages at a client device, the message corresponding with an identification number from among a plurality of sequentially assigned identification numbers associated with the plurality of messages; applying the identification number that corresponds with the message from among the plurality of messages to a data object within a database associated with the message thread, the data object indicating a most recent message read by a user of the client device based on the identification number; detecting a trigger event; and automatically incrementing the identification number associated with the data object within the database responsive to the trigger event.
A map-based graphical user interface (GUI) for a social media platform displays an interactive map populated with user representations positioned at respective geographic locations of respective users. Upon user selection of a user representation on the interactive map, a target user associated with the selected representation is identified and a communication mechanism is caused to be displayed that enables initiation of direct communication with the target user from within the map-based GUI. In one embodiment, a friend panel is launched at a bottom portion of the display screen over the interactive map, the friend panel including the communication mechanism alongside identifying information about the target user. The communication mechanism enables composition and transmission of text-based messages directly to the target user without navigating outside the map-based GUI.
G06F 3/04817 - Interaction techniques based on graphical user interfaces [GUI] based on specific properties of the displayed interaction object or a metaphor-based environment, e.g. interaction with desktop elements like windows or icons, or assisted by a cursor's changing behaviour or appearance using icons
G06F 16/9535 - Search customisation based on user profiles and personalisation
G06F 16/9537 - Spatial or temporal dependent retrieval, e.g. spatiotemporal queries
A device and system for visual reasoning in augmented reality environments employs adaptive multi-frame capture triggered by detection of user speech. Upon detecting speech, the device or system captures image frames at an initial frame capture rate, increasing capture frequency when a hand is detected in a captured image. Timestamped frames and transcribed speech form a prompt for a multimodal large language model, which extracts relevant details with constrained output. A separate language model then generates a final response. This two-stage approach optimizes processing efficiency and accuracy while preserving privacy by limiting continuous visual data collection. The system enables more natural and context-aware interactions in AR settings without complex gesture recognition algorithms.
Systems and methods for radial gesture navigation are provided. In example embodiments, user input data is received from a user device. The user input data indicates a continuous physical user interaction associated with a display screen of the user device. An initial point and a current point are detected from the user input data. A radius distance for a circle that includes the current point and is centered about the initial point is determined. An action is selected from among multiple actions based on the radius distance being within a particular range among successive ranges along a straight line that starts at the initial point and extends through the circle. Each range among the successive ranges corresponds to a particular action among the multiple actions. The selected action is performed in response to detecting a completion of the continuous physical user interaction.
Examples in the present disclosure relate to systems and methods for reducing noise in object tracking data. Images of an object are obtained via one or more cameras. The images are processed to obtain first pose data indicative of a pose of the object over time. The first pose data is represented in a camera space. The first pose data is transformed to second pose data represented in a world space. The second pose data is filtered using a smoothing filter to generate filtered pose data. The filtering includes, for each pose data item in a time series of the second pose data, using a rotation transformation between the world space and camera space to apply one or more camera space-specific filter parameters to the pose data item that is represented in the world space. The pose of the object is dynamically tracked based on the filtered pose data.
A first neural network is trained to generate a ground truth using a small set of example images that illustrate the goal ground truth output images, which can be full-body images of people in an AR style. The first neural network is used to generate ground truth output images from random input images. Example methods of the first neural network include determining poses in input images, changing values of pixels within areas of the input images, inputting the poses, the areas of the changed input images, and a text prompt describing the input images, into a neural network, to generate output images. The methods further include determining losses between the output images and the input images and updating weights of the neural network based on the losses. A second neural network is then trained using the generated ground truth. And, an application is generated that uses the second neural network.
Methods and systems are disclosed for performing operations for estimating a 3D scene representation from one or multiple 2D images. The operations include: receiving one or multiple two-dimensional (2D) images representing a real-world environment; and generating, by a machine learning model, a three-dimensional (3D) scene representation of the 2D image, which explicitly (separately) defines the a 3D shape and appearance of the background as well as a 3D position, 3D shape and appearance of each object of the scene depicted in the set of images, where the machine learning model has been trained in an unsupervised approach from a dataset of images and their camera poses (e.g. without any manually labelled annotations, such as depth maps, segmentation masks, object poses).
A method and a system include, for each predetermined time period in a plurality of predetermined time periods, writing a plurality of data rows comprising a set of data associated with a plurality of active entities, and updating an index table based on the plurality of data rows in the stats table, wherein the index table comprises an index row. The method further includes receiving from an electronic device via an interface a query corresponding to an entity, retrieving an index value from an index row included an index row, retrieving the current value from the stats table using the index value, generating a response to the query using the index value and the current value, and displaying the response on a display of the electronic device.
A waveguide includes a first region with a plurality of first diffractive structures and a second region with a plurality of second diffractive structures. The first and second diffractive structures have first and second values of a first physical property giving rise to a first and second values of an optical property in the respective regions. The waveguide also includes at least one interstitial region located between the first region and the second region, with a plurality of interstitial diffractive structures. Each interstitial diffractive structure has a value of at least one additional physical property giving rise to an intermediate value of the optical property in the at least one interstitial region that is between the first value and the second value of the optical property.
The method involves accessing image data from consecutive frames generated by one or more cameras of a device. A visual Simultaneous Localization and Mapping (SLAM) processing is performed on the image data to detect and track visual features. In the image data, a region of interest is identified by detecting clusters of visual features that either appear and then disappear in consecutive frames or existing tracked features that suddenly lose tracking. Following the identification of the region of interest, a hand detection processing is performed within this area.
G06V 40/10 - Human or animal bodies, e.g. vehicle occupants or pedestriansBody parts, e.g. hands
G06T 7/579 - Depth or shape recovery from multiple images from motion
G06V 10/25 - Determination of region of interest [ROI] or a volume of interest [VOI]
G06V 10/762 - Arrangements for image or video recognition or understanding using pattern recognition or machine learning using clustering, e.g. of similar faces in social networks
G06V 10/82 - Arrangements for image or video recognition or understanding using pattern recognition or machine learning using neural networks
38.
MANUFACTURABLE DIFFRACTIVE STRUCTURES WITH SMOOTHLY VARYING PROPERTIES
A waveguide includes a first region with a plurality of first diffractive structures and a second region with a plurality of second diffractive structures. The first and second diffractive structures have first and second values of a first physical property giving rise to a first and second values of an optical property in the respective regions. The waveguide also includes at least one interstitial region located between the first region and the second region, with a plurality of interstitial diffractive structures. Each interstitial diffractive structure has a value of at least one additional physical property giving rise to an intermediate value of the optical property in the at least one interstitial region that is between the first value and the second value of the optical property.
Described is a system for improving machine learning models. In some cases, the system improves such models by identifying an autoencoder for a latent diffusion machine learning model, the latent diffusion machine learning model is trained to receive text as input and output an image based on the received text. The system identifies a number of channels in a decoder of the autoencoder, the decoder being configured to receive latent features as input and output images. The system further identifies a performance characteristic of the decoder and changes the node topology of the decoder based on the performance characteristic to generate an updated decoder. The system retrains the latent diffusion machine learning model using the updated decoder by inputting latent features to the updated decoder, receiving an outputted image from the updated decoder, and updating one or more weights of the decoder based on an assessment of the outputted image.
An XR system is provided. This system captures images including images of a first hand of a user and a second hand of the user using one or more cameras. The XR system generates cropped images using the images, each cropped image including a surface of the first hand. The XR system detects a hand touch of the surface of the hand by a digit of the second hand using the cropped images. The hand touch is used as an input into an XR user interface of the XR system. The surface of the hand can be palmar surface or a hand dorsal surface.
Systems and methods described herein relate to generation of media collections in a messaging system. The media collection may be created by the user, other users, or an entity. Example embodiments further allow users to set access criteria through privacy settings assigned to one or more media content items themselves, as well as to a media collection, such that some or all of the media collection may only be viewed by users authorized by the user sharing the media content item or media collection (e.g., only to one or more users designated by the user as a “friend”).
G06F 3/04817 - Interaction techniques based on graphical user interfaces [GUI] based on specific properties of the displayed interaction object or a metaphor-based environment, e.g. interaction with desktop elements like windows or icons, or assisted by a cursor's changing behaviour or appearance using icons
G06F 3/04842 - Selection of displayed objects or displayed text elements
G06F 16/9537 - Spatial or temporal dependent retrieval, e.g. spatiotemporal queries
G06F 21/62 - Protecting access to data via a platform, e.g. using keys or access control rules
H04W 4/02 - Services making use of location information
42.
METHOD, SYSTEM, AND MACHINE-READABLE STORAGE MEDIUM FOR VR-BASED CONNECTED PORTAL SHOPPING
Aspects of the present disclosure involve a system comprising a computer-readable storage medium storing programs and methods for performing operations comprising: receiving a request from a client device of a first user to engage in a shared virtual reality shopping experience with a second user; generating, for display on respective client devices of the first and second users, the shared virtual reality shopping experience comprising a plurality of virtual reality items that represent real-world products; receiving, from the client device of the second user, data indicating a selection of a first virtual reality item of the plurality of virtual reality items made by the second user; and modifying a display attribute of the first virtual item in the display of the shared virtual reality shopping experience on the client device of the first user to indicate the selection of the first virtual reality item made by the second user.
Systems and methods are provided for receiving a request to generate a media overlay corresponding to a home of a first user, and for generating the media overlay corresponding to the home of the first user using media content received in the request. The systems and methods further provide for associating, with the media overlay, a location of the home of the first user and a selection of users to grant permission to access the media overlay corresponding to the home of the first user. The systems and methods further provide for determining whether a second user and a location corresponding to the second computing device trigger access to the media overlay and providing the media overlay to the second computing device, based on determining that the second user and location corresponding to the second computing device trigger access to the media overlay.
H04N 21/236 - Assembling of a multiplex stream, e.g. transport stream, by combining a video stream with other content or additional data, e.g. inserting a URL [Uniform Resource Locator ] into a video stream, multiplexing software data into a video streamRemultiplexing of multiplex streamsInsertion of stuffing bits into the multiplex stream, e.g. to obtain a constant bit-rateAssembling of a packetised elementary stream
G06F 3/04842 - Selection of displayed objects or displayed text elements
G06F 3/04847 - Interaction techniques to control parameter settings, e.g. interaction with sliders or dials
Examples in the present disclosure relate to scale estimation for facilitating extended reality (XR) experiences. An image of a hand of a user is captured via one or more optical sensors of an XR device. The image is processed to detect a hand pose relative to the XR device. A hand scale estimate corresponding to the detected hand pose is accessed. The hand scale estimate is one of a plurality of hand scale estimates each uniquely associated with a respective hand pose. The hand scale estimate is applied to generate positional data for one or more features of the hand of the user. The XR device tracks the hand of the user based on the positional data while the user uses the XR device.
The subject technology receives information for a product. The subject technology generates a 3D model file of the product in a first format. The subject technology converts the 3D model file to a 3D object file in a second format. The subject technology associates the 3D object file to the product in a product catalog service. The subject technology publishes an augmented reality (AR) content generator corresponding to the product.
A two-stage approach for learning and generating an expressive text-to-motion animation from partially annotated datasets (T2M-X). In an example implementation, T2M-X builds a unified motion dataset based on partially annotated datasets. In the first stage, T2M-X uses the unified motion dataset to train three vector-quantized variational autoencoders (VQ-VAE) for body, hand, and face, respectively, and generate high-quality motion outputs. In the second stage, T2M-X uses the high-quality motion outputs to train a multi-indexing generative pre-trained transformer (GPT) model that includes motion consistency loss and sequence length consistency for learning and then generating coordinated and expressive animations.
The systems and techniques described herein relate to predicting user conversions in online advertising. Input data associated with user and advertisement features may be processed through neural networks to generate embedding representations or feature cross representations. A multi-task layer calculates probabilities associated with multiple user actions like clicks, page views, sign-ups, or purchases. Click-through and view-through conversion probabilities may be calculated to generate a score. The systems and techniques described herein perform predictions on multiple types of user actions despite data sparsity and negative transfer challenges, enhancing advertisement targeting and improving conversion metrics.
Optical display engines with reduced size and/or increased efficiency for Augmented Reality (AR) and near-eye devices that incorporate LED, microLED, and/or OLED displays. Example optical systems provide different light paths for polarized (or split unpolarized) light, with recombining oppositely polarized light corresponding to different image portions/channels, in a compact system using polarized beam splitters, quarter wave plates, half wave plates, and reflective elements such as curved mirrors or reflective lenses, resulting in high optical power. Images and/or portions thereof are presented at multiple observation planes to project a more realistic synthetic image, and at different angular resolutions to enable creation of a large composite field of view with high apparent resolution from a single small and efficient display.
Example systems take as input a text prompt describing events and generate a multi-event video, audio, or data sequence based on the text prompt. Example operations include accessing a sequence of events, the sequence of events comprising time ranges for the events of the sequence of events and text descriptions for the events, accessing a global caption comprising a text description for the sequence of events, and inputting the sequence of events and the global caption simultaneously into a trained neural network to generate a video, audio, or data sequence, where frames of the video, sound of the audio, or data of the data sequence during a time range of the time ranges are in accordance with a corresponding text description of the text descriptions. Additionally, systems are disclosed that train a neural network to generate the multi-event video, audio, or a data sequence from the text prompt.
The method involves accessing image data from consecutive frames generated by one or more cameras of a device. A visual Simultaneous Localization and Mapping (SLAM) processing is performed on the image data to detect and track visual features. In the image data, a region of interest is identified by detecting clusters of visual features that either appear and then disappear in consecutive frames or existing tracked features that suddenly lose tracking. Following the identification of the region of interest, a hand detection processing is performed within this area.
G06V 10/25 - Determination of region of interest [ROI] or a volume of interest [VOI]
G06V 10/62 - Extraction of image or video features relating to a temporal dimension, e.g. time-based feature extractionPattern tracking
G06V 10/762 - Arrangements for image or video recognition or understanding using pattern recognition or machine learning using clustering, e.g. of similar faces in social networks
G06V 20/20 - ScenesScene-specific elements in augmented reality scenes
G06V 40/10 - Human or animal bodies, e.g. vehicle occupants or pedestriansBody parts, e.g. hands
41 - Education, entertainment, sporting and cultural services
42 - Scientific, technological and industrial services, research and design
45 - Legal and security services; personal services for individuals.
Goods & Services
Online retail store services featuring computer hardware, peripherals, cameras, video cameras, and digital media, namely, pre-recorded music, videos, photographs, images, and audiovisual content; Facilitating the exchange and sale of services and products of third parties via the internet and communication networks, namely, facilitating transactions between buyers and sellers through providing buyers with information about sellers, goods, and/or services via the internet and communication networks; Advertising, marketing, and promotion services; Marketing, advertising, and promotional services using artificial intelligence software, chatbot software, and augmented reality software; Dissemination of advertising for others via computer and other communication networks; Online retail store services featuring a wide variety of consumer goods of others; Promoting the goods and services of others by providing an internet website portal featuring links to the online retail web sites of others; Facilitating the exchange and sale of services and products of third parties via computer and communication networks, namely, operating on-line marketplaces for sellers and buyers of goods and services; Consumer profiling for commercial or marketing purposes; Providing commercial consumer information and advice for consumers in the selection of products to buy Online electronic publishing services, namely, publishing online works of others featuring user-created photographs, images, videos, text and graphics; Providing information and online databases via the Internet in the fields of entertainment and music; Mobile media and entertainment services in the nature of animated and non-animated content preparation, namely, creation and production of multimedia entertainment content in the form of avatars, graphic icons, symbols, images representing individuals, fanciful designs, comics, comic series, phrases, and graphical depictions of people, places and things; Entertainment services, namely, providing online, non-downloadable graphics in the nature of avatars, graphic icons, symbols, images representing individuals, fanciful designs, comics, comic series, phrases, and graphical depictions of people, places and things that end users can transmit and receive by means of the Internet or other computer or telecommunication networks, wireless communications networks, or by using computers, laptops, mobile equipment, and handheld digital electronic devices Rental of computer hardware and computer peripherals; Providing temporary use of online non-downloadable software development kits (SDKs); Development of augmented reality game software; Software design and development; Software development in the framework of software publishing; computer software design; computer software development; customizing computer software; research and development of computer software; Providing temporary use of online non-downloadable middleware for providing an interface between augmented reality devices and operating systems; providing temporary use of online non-downloadable software for providing an interface between augmented reality devices and operating systems; providing temporary use of online non-downloadable software for providing an interface between computer peripheral devices and operating systems; Providing online non-downloadable computer software using artificial intelligence for machine learning; Product research and product development in the field of artificial intelligence; Providing online non-downloadable software using artificial intelligence for natural language processing, generation, understanding, and analysis; Providing online non-downloadable software for developing, running and analyzing algorithms that are able to learn to analyze, classify, and take actions in response to exposure to data; Providing on-line non-downloadable software using artificial intelligence for image recognition and generation; Providing on-line non-downloadable software using artificial intelligence for text recognition and generation; Providing online non-downloadable software for the generation of advertisements and promotional materials; Providing on-line non-downloadable software using artificial intelligence for music generation and suggestions; Providing on-line non-downloadable software using artificial intelligence for image and video editing and retouching; Providing on-line non-downloadable software using artificial intelligence for the generation of text, images, photos, videos, audio, and multimedia content; Providing on-line non-downloadable software using artificial intelligence for connecting consumers with promotional advertisements; Providing temporary use of online non-downloadable chatbot software using artificial intelligence for connecting consumers with advertisements; Providing temporary use of online non-downloadable chatbot software using artificial intelligence for simulating human conversations; Providing temporary use of online non-downloadable chatbot software using artificial intelligence for responding to oral and written prompts; Providing temporary use of online non-downloadable chatbot software using artificial intelligence for image recognition and generation; providing temporary use of online non-downloadable chatbot software using artificial intelligence for text recognition and generation; providing temporary use of online non-downloadable chatbot software using artificial intelligence for the generation of text, images, photos, video, audio, text, and multimedia content Computer software licensing; Licensing of software in the framework of software publishing; Providing online computer databases in the field of social introduction; Internet-based social introduction services
09 - Scientific and electric apparatus and instruments
Goods & Services
Computer hardware; Computer peripherals; Wearable computer hardware; Wearable computer peripherals; Computer hardware, peripherals and recorded software for remotely accessing, capturing, transmitting and displaying pictures, video, audio and data; Cameras; Digital cameras; Digital video cameras; Video cameras; Video recorders; Remote controls for cameras and video recorders; Downloadable software for cameras, video cameras and video recorders; Downloadable software for setting up, configuring, and controlling wearable computer hardware and peripherals; Electric wires and cables for camera electricity mains; Downloadable computer software and software applications for use in uploading, downloading, capturing, editing, storing, accessing, posting, displaying, tagging, distributing, streaming, linking, sharing, transmitting or otherwise providing photos, videos, images, text, electronic media, photographic and video content, digital data, or information via the Internet, communication networks, mobile phones and mobile devices; Wearable computer hardware and peripherals in the nature of smartglasses featuring software and display screens to enable augmented reality and virtual reality experiences; Electronic sensors, cameras, and microphones for object, landscape, gesture, facial, and voice detection, tracking, capture, and recognition; downloadable computer software and downloadable application programming interface (API) for use in creating and designing augmented reality and virtual reality experiences; Computer hardware, peripherals, and downloadable software for tracking motion in, visualizing, manipulating, viewing, transmitting, and displaying images, video, audio, and data for augmented reality and virtual reality experiences; Downloadable software development kits (SDK); Downloadable augmented reality software for use in mobile devices for integrating electronic data with real-world environments for the purpose of creating video content, text, graphics, animations, and links; Recorded and downloadable augmented reality computer software platform for developers to use to create immersive augmented reality experiences for augmented reality glasses; Downloadable computer software for creating augmented reality experiences; Downloadable software for creating augmented reality software; Downloadable computer software, namely, augmented reality software for integrating electronic data with real world environments for the purpose of experiencing, viewing, capturing, recording and editing augmented images, videos, and audio; Downloadable computer programs and downloadable computer software using artificial intelligence for the purpose of for machine learning; Downloadable computer programs and downloadable computer software using artificial intelligence for natural language processing, generation, understanding and analysis; Downloadable computer programs and downloadable computer software for image recognition and generation; downloadable computer programs and downloadable computer software using artificial intelligence for music generation and suggestions; downloadable computer programs and downloadable computer software for artificial intelligence, namely, computer software for developing, running and analyzing algorithms that are able to learn to analyze, classify, and take actions in response to exposure to data; downloadable computer software using artificial intelligence for image and video editing and retouching; downloadable computer software using artificial intelligence for the generation of text, images, photos, videos, audio, and multimedia content; downloadable computer software using artificial intelligence for connecting consumers with targeted promotional advertisements; downloadable computer software using artificial intelligence for the generation of advertisements and promotional materials; downloadable computer software using artificial intelligence for creating and generating text; downloadable computer software using artificial intelligence for translating words or text from one language to another; Downloadable chatbot software using artificial intelligence for generating augmented reality software for integrating electronic data with real world environments for the purposes of experiencing, viewing, capturing, recording, and editing augmented images, videos, audio, and sensory content; downloadable chatbot software using artificial intelligence for generating augmented reality experiences and augmented reality content; downloadable chatbot software using artificial intelligence for generating entertainment and educational content; downloadable chatbot software for connecting consumers with promotional messaging; downloadable chatbot software for simulating human conversations; downloadable chatbot software for suggesting image, video, audio, text, and multimedia content; downloadable chatbot software for responding to oral and written prompts; Downloadable augmented reality and virtual reality software for use in mobile devices for integrating electronic data with real and virtual world environment in the fields of entertainment, photography, communications, and online social sharing; Downloadable multimedia files containing digital photos, video, audio, and other digital data all in the fields of entertainment, photography and online social sharing
09 - Scientific and electric apparatus and instruments
35 - Advertising and business services
41 - Education, entertainment, sporting and cultural services
42 - Scientific, technological and industrial services, research and design
45 - Legal and security services; personal services for individuals.
Goods & Services
Computer hardware; Computer peripherals; Wearable computer hardware; Wearable computer peripherals; Computer hardware, peripherals and recorded software for remotely accessing, capturing, transmitting and displaying pictures, video, audio and data; Cameras; Digital cameras; Digital video cameras; Video cameras; Video recorders; Remote controls for cameras and video recorders; Downloadable software for cameras, video cameras and video recorders; Downloadable software for setting up, configuring, and controlling wearable computer hardware and peripherals; Electric wires and cables for camera electricity mains; Downloadable computer software and software applications for use in uploading, downloading, capturing, editing, storing, accessing, posting, displaying, tagging, distributing, streaming, linking, sharing, transmitting or otherwise providing photos, videos, images, text, electronic media, photographic and video content, digital data, or information via the Internet, communication networks, mobile phones and mobile devices; Wearable computer hardware and peripherals in the nature of smartglasses featuring software and display screens to enable augmented reality and virtual reality experiences; Electronic sensors, cameras, and microphones for object, landscape, gesture, facial, and voice detection, tracking, capture, and recognition; downloadable computer software and downloadable application programming interface (API) for use in creating and designing augmented reality and virtual reality experiences; Computer hardware, peripherals, and downloadable software for tracking motion in, visualizing, manipulating, viewing, transmitting, and displaying images, video, audio, and data for augmented reality and virtual reality experiences; Downloadable software development kits (SDK); Downloadable augmented reality software for use in mobile devices for integrating electronic data with real-world environments for the purpose of creating video content, text, graphics, animations, and links; Recorded and downloadable augmented reality computer software platform for developers to use to create immersive augmented reality experiences for augmented reality glasses; Downloadable computer software for creating augmented reality experiences; Downloadable software for creating augmented reality software; Downloadable computer software, namely, augmented reality software for integrating electronic data with real world environments for the purpose of experiencing, viewing, capturing, recording and editing augmented images, videos, and audio; Downloadable computer programs and downloadable computer software using artificial intelligence for the purpose of for machine learning; Downloadable computer programs and downloadable computer software using artificial intelligence for natural language processing, generation, understanding and analysis; Downloadable computer programs and downloadable computer software for image recognition and generation; downloadable computer programs and downloadable computer software using artificial intelligence for music generation and suggestions; downloadable computer programs and downloadable computer software for artificial intelligence, namely, computer software for developing, running and analyzing algorithms that are able to learn to analyze, classify, and take actions in response to exposure to data; downloadable computer software using artificial intelligence for image and video editing and retouching; downloadable computer software using artificial intelligence for the generation of text, images, photos, videos, audio, and multimedia content; downloadable computer software using artificial intelligence for connecting consumers with targeted promotional advertisements; downloadable computer software using artificial intelligence for the generation of advertisements and promotional materials; downloadable computer software using artificial intelligence for creating and generating text; downloadable computer software using artificial intelligence for translating words or text from one language to another; Downloadable chatbot software using artificial intelligence for generating augmented reality software for integrating electronic data with real world environments for the purposes of experiencing, viewing, capturing, recording, and editing augmented images, videos, audio, and sensory content; downloadable chatbot software using artificial intelligence for generating augmented reality experiences and augmented reality content; downloadable chatbot software using artificial intelligence for generating entertainment and educational content; downloadable chatbot software for connecting consumers with promotional messaging; downloadable chatbot software for simulating human conversations; downloadable chatbot software for suggesting image, video, audio, text, and multimedia content; downloadable chatbot software for responding to oral and written prompts; Downloadable augmented reality and virtual reality software for use in mobile devices for integrating electronic data with real and virtual world environment in the fields of entertainment, photography, communications, and online social sharing; Downloadable multimedia files containing digital photos, video, audio, and other digital data all in the fields of entertainment, photography and online social sharing Online retail store services featuring computer hardware, peripherals, cameras, video cameras, and digital media, namely, pre-recorded music, videos, photographs, images, and audiovisual content; Facilitating the exchange and sale of services and products of third parties via the internet and communication networks, namely, facilitating transactions between buyers and sellers through providing buyers with information about sellers, goods, and/or services via the internet and communication networks; Advertising, marketing, and promotion services; Marketing, advertising, and promotional services using artificial intelligence software, chatbot software, and augmented reality software; Dissemination of advertising for others via computer and other communication networks; Online retail store services featuring a wide variety of consumer goods of others; Promoting the goods and services of others by providing an internet website portal featuring links to the online retail web sites of others; Facilitating the exchange and sale of services and products of third parties via computer and communication networks, namely, operating on-line marketplaces for sellers and buyers of goods and services; Consumer profiling for commercial or marketing purposes; Providing commercial consumer information and advice for consumers in the selection of products to buy Online electronic publishing services, namely, publishing online works of others featuring user-created photographs, images, videos, text and graphics; Providing information and online databases via the Internet in the fields of entertainment and music; Mobile media and entertainment services in the nature of animated and non-animated content preparation, namely, creation and production of multimedia entertainment content in the form of avatars, graphic icons, symbols, images representing individuals, fanciful designs, comics, comic series, phrases, and graphical depictions of people, places and things; Entertainment services, namely, providing online, non-downloadable graphics in the nature of avatars, graphic icons, symbols, images representing individuals, fanciful designs, comics, comic series, phrases, and graphical depictions of people, places and things that end users can transmit and receive by means of the Internet or other computer or telecommunication networks, wireless communications networks, or by using computers, laptops, mobile equipment, and handheld digital electronic devices Rental of computer hardware and computer peripherals; Providing temporary use of online non-downloadable software development kits (SDKs); Development of augmented reality game software; Software design and development; Software development in the framework of software publishing; computer software design; computer software development; customizing computer software; research and development of computer software; Providing temporary use of online non-downloadable middleware for providing an interface between augmented reality devices and operating systems; providing temporary use of online non-downloadable software for providing an interface between augmented reality devices and operating systems; providing temporary use of online non-downloadable software for providing an interface between computer peripheral devices and operating systems; Providing online non-downloadable computer software using artificial intelligence for machine learning; Product research and product development in the field of artificial intelligence; Providing online non-downloadable software using artificial intelligence for natural language processing, generation, understanding, and analysis; Providing online non-downloadable software for developing, running and analyzing algorithms that are able to learn to analyze, classify, and take actions in response to exposure to data; Providing on-line non-downloadable software using artificial intelligence for image recognition and generation; Providing on-line non-downloadable software using artificial intelligence for text recognition and generation; Providing online non-downloadable software for the generation of advertisements and promotional materials; Providing on-line non-downloadable software using artificial intelligence for music generation and suggestions; Providing on-line non-downloadable software using artificial intelligence for image and video editing and retouching; Providing on-line non-downloadable software using artificial intelligence for the generation of text, images, photos, videos, audio, and multimedia content; Providing on-line non-downloadable software using artificial intelligence for connecting consumers with promotional advertisements; Providing temporary use of online non-downloadable chatbot software using artificial intelligence for connecting consumers with advertisements; Providing temporary use of online non-downloadable chatbot software using artificial intelligence for simulating human conversations; Providing temporary use of online non-downloadable chatbot software using artificial intelligence for responding to oral and written prompts; Providing temporary use of online non-downloadable chatbot software using artificial intelligence for image recognition and generation; providing temporary use of online non-downloadable chatbot software using artificial intelligence for text recognition and generation; providing temporary use of online non-downloadable chatbot software using artificial intelligence for the generation of text, images, photos, video, audio, text, and multimedia content Computer software licensing; Licensing of software in the framework of software publishing; Providing online computer databases in the field of social introduction; Internet-based social introduction services
54.
Bystander speech command rejection for wearable devices
A system comprising a head-wearable apparatus, which includes a plurality of microphones, each microphone of the plurality of microphones being spatially separated from each other microphone of the plurality of microphones, thereby defining a set of spatial separations. The system also includes one or more processors. The system also includes a non-transitory computer readable storage medium including instructions that, when executed by the one or more processors, cause the one or more processors to perform operations including generating, by each microphone of the plurality of microphones, an audio signal corresponding to sound detected by the microphone, thereby generating a plurality of audio signals, and processing the plurality of audio signals, using a trained machine learning model, to generate wearer speech detection data representative of a likelihood that a wearer of the head-wearable apparatus is speaking.
A system is disclosed, including a processor and a memory. The memory stores instructions that, when executed by the processor, configure the system to perform operations. Autoexposure (AE) primary camera information is obtained, identifying a first camera as an AE primary camera to be used for computing AE settings of the first camera and a second camera. First camera region of interest (ROI) information and second camera ROI information are obtained, representative of a first number of ROIs within a field of view (FOV) of the first camera and a second number of ROIs within a FOV of the second camera. In response to determining that the first number is zero and the second number is greater than zero, the AE primary camera information is updated to identify the second camera as the AE primary camera, and the second camera ROI information is processed to generate the AE settings.
A map-based graphical user interface (GUI) for a social media application displays an interactive map on a user device showing two visually distinct types of icons simultaneously overlaid on the map. A first type comprises user icons representing the respective geographic locations of friend users, each user icon displayed at a display location determined based on a most recently updated location of the user device of the respective friend. A second type comprises content icons representing geo-anchored collections of location-tagged content uploaded by respective individual users, each content icon displayed at a geographically fixed location determined by the location-tag data of the associated content and being selectable to facilitate user access to that content via the interactive map. The first and second icon types have visually distinct shapes, differentiating dynamically positioned friend location indicators from geographically fixed content indicators on the map. The content icons are ephemeral, each having a respective ephemeral lifetime, and each content icon is automatically excluded from the interactive map upon expiry of its associated ephemeral lifetime, causing the map to be updated to reflect current content availability without user intervention.
G06F 3/04817 - Interaction techniques based on graphical user interfaces [GUI] based on specific properties of the displayed interaction object or a metaphor-based environment, e.g. interaction with desktop elements like windows or icons, or assisted by a cursor's changing behaviour or appearance using icons
G06F 3/0482 - Interaction with lists of selectable items, e.g. menus
G06F 3/04842 - Selection of displayed objects or displayed text elements
G06F 3/0487 - Interaction techniques based on graphical user interfaces [GUI] using specific features provided by the input device, e.g. functions controlled by the rotation of a mouse with dual sensing arrangements, or of the nature of the input device, e.g. tap gestures based on pressure sensed by a digitiser
G06F 3/0488 - Interaction techniques based on graphical user interfaces [GUI] using specific features provided by the input device, e.g. functions controlled by the rotation of a mouse with dual sensing arrangements, or of the nature of the input device, e.g. tap gestures based on pressure sensed by a digitiser using a touch-screen or digitiser, e.g. input of commands through traced gestures
G06F 16/487 - Retrieval characterised by using metadata, e.g. metadata not derived from the content or metadata generated manually using geographical or spatial information, e.g. location
H04L 41/22 - Arrangements for maintenance, administration or management of data switching networks, e.g. of packet switching networks comprising specially adapted graphical user interfaces [GUI]
H04L 41/28 - Restricting access to network management systems or functions, e.g. using authorisation function to access network configuration
H04L 51/52 - User-to-user messaging in packet-switching networks, transmitted according to store-and-forward or real-time protocols, e.g. e-mail for supporting social networking services
H04L 67/12 - Protocols specially adapted for proprietary or special-purpose networking environments, e.g. medical networks, sensor networks, networks in vehicles or remote metering networks
H04L 67/52 - Network services specially adapted for the location of the user terminal
H04W 4/02 - Services making use of location information
H04W 4/029 - Location-based management or tracking services
H04W 4/18 - Information format or content conversion, e.g. adaptation by the network of the transmitted or received information for the purpose of wireless delivery to users or terminals
H04W 4/21 - Services signallingAuxiliary data signalling, i.e. transmitting data via a non-traffic channel for social networking applications
H04W 12/02 - Protecting privacy or anonymity, e.g. protecting personally identifiable information [PII]
57.
MAP INTERFACE WITH STORY ACCESS VIA MAP-DISPLAYED FRIEND ICONS
A map-based graphical user interface (GUI) for a social media application displays an interactive map on a user device, with a plurality of friend icons rendered at respective geographic locations representing the current or last known locations of friends in the user's social network. User selection of a friend icon via a touchscreen input coincident with the icon's on-screen location launches a full-screen display of the selected friend's story—a collection of media items uploaded by that friend—the full-screen display temporarily replacing the interactive map. The story is compiled without regard to geographic considerations, such that all unexpired media items uploaded by the friend are included irrespective of location. The full-screen display comprises automated sequential playback in chronological order, each item displayed for a respective duration before automatic advancement. Upon completion or dismissal, the interactive map is restored at the same geographic area and zoom level previously displayed. In some embodiments, selection of a friend icon instead causes display of a friend panel from which the story is accessible via a secondary selection.
G06F 3/04817 - Interaction techniques based on graphical user interfaces [GUI] based on specific properties of the displayed interaction object or a metaphor-based environment, e.g. interaction with desktop elements like windows or icons, or assisted by a cursor's changing behaviour or appearance using icons
G06F 16/9535 - Search customisation based on user profiles and personalisation
G06F 16/9537 - Spatial or temporal dependent retrieval, e.g. spatiotemporal queries
A system to display a route of a user over a period of time is configured to perform operations that include: causing display of a map image that depicts a location; accessing user profile data associated with a user profile, the user profile data comprising a user identifier and location data associated with the user profile; identifying a sequence of locations associated with the user profile based on the user profile data; and causing display of a presentation of a trail indicating the sequence of locations associated with the user profile, the trail terminating at a display of the user identifier.
A map-based graphical user interface (GUI) for a social media application displays an interactive map on which friend icons represent respective geographic locations of friend users. The map-based GUI provides location-based access mechanisms, each enabling access to social media content based on associated location information, and additionally provides a friend-level access mechanism enabling location-agnostic access to friend user content. Accessibility of friend content via the friend-level access mechanism is agnostic to the identity of or variation in the geographic area displayed on the interactive map. Responsive to user input selecting a target friend user via the friend-level access mechanism, a location-agnostic collection of social media items is displayed — the composition of that collection being likewise agnostic to the geographic area displayed. The friend-level access mechanism is operable via friend icons on the map, a friend panel, or a search interface provided as part of the map-based GUI.
G06F 3/04817 - Interaction techniques based on graphical user interfaces [GUI] based on specific properties of the displayed interaction object or a metaphor-based environment, e.g. interaction with desktop elements like windows or icons, or assisted by a cursor's changing behaviour or appearance using icons
G06F 16/9535 - Search customisation based on user profiles and personalisation
G06F 16/9537 - Spatial or temporal dependent retrieval, e.g. spatiotemporal queries
A server system manages a plurality of location-based social media galleries, each gallery having an associated anchor location and comprising a set of social media items whose geo-tag data corresponds to that anchor location. A map-based graphical user interface (GUI) for the social media application is caused to display on a user device, the map-based GUI including a map representing a geographical area. For each gallery to be surfaced, a gallery icon is displayed at the anchor location on the map, the gallery icon being an interactive, user-selectable GUI element rendered at the geographic coordinate of the anchor location and representing the corresponding location-based social media gallery. User selection of a gallery icon triggers automated sequential replay on the user device of the set of social media items of the corresponding gallery, each item being displayed for a respective display duration before automatic advancement to the next, the automated sequential replay being performed in a full-screen mode temporarily replacing the map-based GUI. In some embodiments, the gallery icon bears a thumbnail image derived from one of the social media items in the gallery set and a text label identifying the anchor location.
G06F 3/04817 - Interaction techniques based on graphical user interfaces [GUI] based on specific properties of the displayed interaction object or a metaphor-based environment, e.g. interaction with desktop elements like windows or icons, or assisted by a cursor's changing behaviour or appearance using icons
G06F 3/0482 - Interaction with lists of selectable items, e.g. menus
G06F 3/04842 - Selection of displayed objects or displayed text elements
G06F 3/0487 - Interaction techniques based on graphical user interfaces [GUI] using specific features provided by the input device, e.g. functions controlled by the rotation of a mouse with dual sensing arrangements, or of the nature of the input device, e.g. tap gestures based on pressure sensed by a digitiser
G06F 3/0488 - Interaction techniques based on graphical user interfaces [GUI] using specific features provided by the input device, e.g. functions controlled by the rotation of a mouse with dual sensing arrangements, or of the nature of the input device, e.g. tap gestures based on pressure sensed by a digitiser using a touch-screen or digitiser, e.g. input of commands through traced gestures
G06F 16/487 - Retrieval characterised by using metadata, e.g. metadata not derived from the content or metadata generated manually using geographical or spatial information, e.g. location
H04L 41/22 - Arrangements for maintenance, administration or management of data switching networks, e.g. of packet switching networks comprising specially adapted graphical user interfaces [GUI]
H04L 41/28 - Restricting access to network management systems or functions, e.g. using authorisation function to access network configuration
H04L 51/52 - User-to-user messaging in packet-switching networks, transmitted according to store-and-forward or real-time protocols, e.g. e-mail for supporting social networking services
H04L 67/12 - Protocols specially adapted for proprietary or special-purpose networking environments, e.g. medical networks, sensor networks, networks in vehicles or remote metering networks
H04L 67/52 - Network services specially adapted for the location of the user terminal
H04W 4/02 - Services making use of location information
H04W 4/029 - Location-based management or tracking services
H04W 4/18 - Information format or content conversion, e.g. adaptation by the network of the transmitted or received information for the purpose of wireless delivery to users or terminals
H04W 4/21 - Services signallingAuxiliary data signalling, i.e. transmitting data via a non-traffic channel for social networking applications
H04W 12/02 - Protecting privacy or anonymity, e.g. protecting personally identifiable information [PII]
61.
LOCATION VISIBILITY MECHANISMS FOR MAP-BASED SOCIAL MEDIA GRAPHICAL USER INTERFACE USER INTERFACE
A map-based graphical user interface (GUI) for a social media platform displays an interactive map on a requesting user's device showing geographic locations of other users sharing their locations with the requesting user. The GUI provides a visibility sharing mechanism comprising a user-operable control element enabling selective switching between a visible mode, in which the requesting user's device location is shared with and visible to other users via their respective map-based GUIs, and an invisible mode, in which the requesting user's location is excluded from display on all other users' map-based GUIs. Upon activation of the invisible mode, the server system ceases transmission of the requesting user's location data to other user devices. The requesting user's map-based GUI continues to display other users' locations and the requesting user's own location while the invisible mode is active. Granular visibility control is supported at general, group, and per-user levels, with per-user settings overriding group-level settings.
G06F 3/04817 - Interaction techniques based on graphical user interfaces [GUI] based on specific properties of the displayed interaction object or a metaphor-based environment, e.g. interaction with desktop elements like windows or icons, or assisted by a cursor's changing behaviour or appearance using icons
G06F 3/0482 - Interaction with lists of selectable items, e.g. menus
G06F 3/04842 - Selection of displayed objects or displayed text elements
G06F 3/0487 - Interaction techniques based on graphical user interfaces [GUI] using specific features provided by the input device, e.g. functions controlled by the rotation of a mouse with dual sensing arrangements, or of the nature of the input device, e.g. tap gestures based on pressure sensed by a digitiser
G06F 3/0488 - Interaction techniques based on graphical user interfaces [GUI] using specific features provided by the input device, e.g. functions controlled by the rotation of a mouse with dual sensing arrangements, or of the nature of the input device, e.g. tap gestures based on pressure sensed by a digitiser using a touch-screen or digitiser, e.g. input of commands through traced gestures
G06F 16/487 - Retrieval characterised by using metadata, e.g. metadata not derived from the content or metadata generated manually using geographical or spatial information, e.g. location
H04L 41/22 - Arrangements for maintenance, administration or management of data switching networks, e.g. of packet switching networks comprising specially adapted graphical user interfaces [GUI]
H04L 41/28 - Restricting access to network management systems or functions, e.g. using authorisation function to access network configuration
H04L 51/52 - User-to-user messaging in packet-switching networks, transmitted according to store-and-forward or real-time protocols, e.g. e-mail for supporting social networking services
H04L 67/12 - Protocols specially adapted for proprietary or special-purpose networking environments, e.g. medical networks, sensor networks, networks in vehicles or remote metering networks
H04L 67/52 - Network services specially adapted for the location of the user terminal
H04W 4/02 - Services making use of location information
H04W 4/029 - Location-based management or tracking services
H04W 4/18 - Information format or content conversion, e.g. adaptation by the network of the transmitted or received information for the purpose of wireless delivery to users or terminals
H04W 4/21 - Services signallingAuxiliary data signalling, i.e. transmitting data via a non-traffic channel for social networking applications
H04W 12/02 - Protecting privacy or anonymity, e.g. protecting personally identifiable information [PII]
Methods and systems are disclosed for detecting whether a wearable device is being worn by a user. The system transmits a radio signal from a first communication device of a wearable device to a second communication device of the wearable device and measures a signal strength associated with the radio signal received by the second communication device. The system compares the signal strength to a threshold value and generates an indication of a wear status associated with the wearable device based on comparing the signal strength to the threshold value.
A head-worn device system includes one or more cameras, one or more display devices and one or more processors. The system also includes a memory storing instructions that, when executed by the one or more processors, configure the system to receive image data from one or more cameras of an augmented reality device, and detect a hand of a user within the image data. Based on the system determining that the hand is holding the object, the system sets an object-in-hand detection flag and can enable or disable at least one user interface element corresponding the hand that is holding the object.
A head-worn device system includes one or more cameras, one or more display devices and one or more processors. The system also includes a memory storing instructions that, when executed by the one or more processors, configure the system to receive image data from one or more cameras of an augmented reality device, and detect a hand of a user within the image data. Based on the system determining that the hand is holding the object, the system sets an object-in-hand detection flag and can enable or disable at least one user interface element corresponding the hand that is holding the object.
Systems, methods, and computer readable media for graphical assistance with tasks using an augmented reality (AR) wearable devices are disclosed. Embodiments capture an image of a first user view of a real-world scene and access indications of surfaces and locations of the surfaces detected in the image. The AR wearable device displays indications of the surfaces on a display of the AR wearable device where the locations of the indications are based on the locations of the surfaces and a second user view of the real-world scene. The locations of the surfaces are indicated with 3D world coordinates. The user views are determined based on a location of the user. The AR wearable device enables a user to add graphics to the surfaces and select tasks to perform. Tools such as a bubble level or a measuring tool are available for the user to utilize to perform the task.
An eXtended Reality (XR) system provides methodologies for clipping virtual content displayed to a user. The XR system generates an XR user interface with virtual content using an XR user interface model. The XR system generates clipped virtual content from the virtual content by clipping virtual content that is located outside of the user's stereoscopic field of view and provides the XR user interface containing the clipped virtual content to the user.
System and method including receiving visual content for display on a display device, the visual content comprising a plurality of layers comprising a background layer and visual media layer, the plurality of layers being arranged in a presentation stack of layers, evaluating the visual content to determine at least one display parameter, determining a desired brightness level for displaying the visual content on a display of the display device based on the at least one display parameter, determining adjustment of display parameters of the presentation stack of layers to achieve the desired brightness level for displaying the visual content, based on adjusting the display parameters of the presentation stack of layers, generating adjusted visual content, and causing presentation of the adjusted visual content on the display device.
A server system of a social media platform automatically compiles multiple location-based galleries from social media items submitted by users to the platform. For each gallery, social media items whose geotag data satisfies predefined location matching criteria relative to a respective anchor location are identified and assembled into the gallery's content set. The compilation is performed as a continuous, ongoing operation such that newly submitted items are automatically added upon submission and expired items automatically removed. A map-based graphical user interface displays a plurality of gallery icons on an interactive map, each gallery icon positioned at the anchor location of a respective location-based gallery. Responsive to user selection of a gallery icon, automated sequential reproduction of the social media items in the corresponding gallery is caused on the user device, each item being displayed for a respective display duration before automatically advancing to the next. The compilation may incorporate automated content evaluation procedures and a server-side curation interface operable by a human operator to review and modify gallery content.
G06F 3/04817 - Interaction techniques based on graphical user interfaces [GUI] based on specific properties of the displayed interaction object or a metaphor-based environment, e.g. interaction with desktop elements like windows or icons, or assisted by a cursor's changing behaviour or appearance using icons
G06F 3/0482 - Interaction with lists of selectable items, e.g. menus
G06F 3/04842 - Selection of displayed objects or displayed text elements
G06F 3/0487 - Interaction techniques based on graphical user interfaces [GUI] using specific features provided by the input device, e.g. functions controlled by the rotation of a mouse with dual sensing arrangements, or of the nature of the input device, e.g. tap gestures based on pressure sensed by a digitiser
G06F 3/0488 - Interaction techniques based on graphical user interfaces [GUI] using specific features provided by the input device, e.g. functions controlled by the rotation of a mouse with dual sensing arrangements, or of the nature of the input device, e.g. tap gestures based on pressure sensed by a digitiser using a touch-screen or digitiser, e.g. input of commands through traced gestures
G06F 16/487 - Retrieval characterised by using metadata, e.g. metadata not derived from the content or metadata generated manually using geographical or spatial information, e.g. location
H04L 41/22 - Arrangements for maintenance, administration or management of data switching networks, e.g. of packet switching networks comprising specially adapted graphical user interfaces [GUI]
H04L 41/28 - Restricting access to network management systems or functions, e.g. using authorisation function to access network configuration
H04L 51/52 - User-to-user messaging in packet-switching networks, transmitted according to store-and-forward or real-time protocols, e.g. e-mail for supporting social networking services
H04L 67/12 - Protocols specially adapted for proprietary or special-purpose networking environments, e.g. medical networks, sensor networks, networks in vehicles or remote metering networks
H04L 67/52 - Network services specially adapted for the location of the user terminal
H04W 4/02 - Services making use of location information
H04W 4/029 - Location-based management or tracking services
H04W 4/18 - Information format or content conversion, e.g. adaptation by the network of the transmitted or received information for the purpose of wireless delivery to users or terminals
H04W 4/21 - Services signallingAuxiliary data signalling, i.e. transmitting data via a non-traffic channel for social networking applications
H04W 12/02 - Protecting privacy or anonymity, e.g. protecting personally identifiable information [PII]
69.
GESTURE-BASED INTERACTIVE MAP INTERFACE FOR MEDIA CONTENT
A mobile user device presents an interactive map interface for a social media application on a touchscreen display. A plurality of gallery icons are displayed overlaid on the interactive map at respective geographic locations corresponding to anchor locations of respective location-based social media galleries. The interactive map supports gesture-based navigation via haptic input on the touchscreen display, including a pinch-out gesture that increases the magnification level of the map and a pinch-in gesture that decreases the magnification level, each treated as a distinct gesture type producing a distinct navigational response, and a dragging gesture performed at any location on the touchscreen display that pans the map to change the displayed geographic area.
G06F 3/04817 - Interaction techniques based on graphical user interfaces [GUI] based on specific properties of the displayed interaction object or a metaphor-based environment, e.g. interaction with desktop elements like windows or icons, or assisted by a cursor's changing behaviour or appearance using icons
G06F 3/0482 - Interaction with lists of selectable items, e.g. menus
G06F 3/04842 - Selection of displayed objects or displayed text elements
G06F 3/0487 - Interaction techniques based on graphical user interfaces [GUI] using specific features provided by the input device, e.g. functions controlled by the rotation of a mouse with dual sensing arrangements, or of the nature of the input device, e.g. tap gestures based on pressure sensed by a digitiser
G06F 3/0488 - Interaction techniques based on graphical user interfaces [GUI] using specific features provided by the input device, e.g. functions controlled by the rotation of a mouse with dual sensing arrangements, or of the nature of the input device, e.g. tap gestures based on pressure sensed by a digitiser using a touch-screen or digitiser, e.g. input of commands through traced gestures
G06F 16/487 - Retrieval characterised by using metadata, e.g. metadata not derived from the content or metadata generated manually using geographical or spatial information, e.g. location
H04L 41/22 - Arrangements for maintenance, administration or management of data switching networks, e.g. of packet switching networks comprising specially adapted graphical user interfaces [GUI]
H04L 41/28 - Restricting access to network management systems or functions, e.g. using authorisation function to access network configuration
H04L 51/52 - User-to-user messaging in packet-switching networks, transmitted according to store-and-forward or real-time protocols, e.g. e-mail for supporting social networking services
H04L 67/12 - Protocols specially adapted for proprietary or special-purpose networking environments, e.g. medical networks, sensor networks, networks in vehicles or remote metering networks
H04L 67/52 - Network services specially adapted for the location of the user terminal
H04W 4/02 - Services making use of location information
H04W 4/029 - Location-based management or tracking services
H04W 4/18 - Information format or content conversion, e.g. adaptation by the network of the transmitted or received information for the purpose of wireless delivery to users or terminals
H04W 4/21 - Services signallingAuxiliary data signalling, i.e. transmitting data via a non-traffic channel for social networking applications
H04W 12/02 - Protecting privacy or anonymity, e.g. protecting personally identifiable information [PII]
70.
DUAL-CONFIGURATION MAP-BASED GRAPHICAL USER INTERFACE
A graphical user interface (GUI) for a social media platform is displayable on a user device associated with a particular user in either of two selectively switchable configurations. In a first configuration, the GUI provides an interactive live map that enables location-based exploration of and access to geo-tagged ephemeral items uploaded to the social media platform by other users, each ephemeral item having a predefined limited ephemeral lifetime and becoming inaccessible via the first configuration upon expiry of its respective lifetime. Responsive to selective user input, the GUI is caused to display in a second configuration providing an interactive historical map that enables location-based exploration of and access to geo-tagged social media content previously uploaded by the particular user, the content comprising expired ephemeral items previously uploaded by that user. In the second configuration, the GUI provides access exclusively to the particular user's own geo-tagged historical content, such that content uploaded by other users is not accessible via the second configuration. Each configuration displays location indicators on the interactive map at geographic locations indicated by the geo-tag data of the respective accessible content, enabling the user to explore personal content history geographically. In some embodiments, the GUI includes a toggle operable to switch back and forth between the first and second configurations.
G06F 3/04817 - Interaction techniques based on graphical user interfaces [GUI] based on specific properties of the displayed interaction object or a metaphor-based environment, e.g. interaction with desktop elements like windows or icons, or assisted by a cursor's changing behaviour or appearance using icons
G06F 3/0482 - Interaction with lists of selectable items, e.g. menus
G06F 3/04842 - Selection of displayed objects or displayed text elements
G06F 3/0487 - Interaction techniques based on graphical user interfaces [GUI] using specific features provided by the input device, e.g. functions controlled by the rotation of a mouse with dual sensing arrangements, or of the nature of the input device, e.g. tap gestures based on pressure sensed by a digitiser
G06F 3/0488 - Interaction techniques based on graphical user interfaces [GUI] using specific features provided by the input device, e.g. functions controlled by the rotation of a mouse with dual sensing arrangements, or of the nature of the input device, e.g. tap gestures based on pressure sensed by a digitiser using a touch-screen or digitiser, e.g. input of commands through traced gestures
G06F 16/487 - Retrieval characterised by using metadata, e.g. metadata not derived from the content or metadata generated manually using geographical or spatial information, e.g. location
H04L 41/22 - Arrangements for maintenance, administration or management of data switching networks, e.g. of packet switching networks comprising specially adapted graphical user interfaces [GUI]
H04L 41/28 - Restricting access to network management systems or functions, e.g. using authorisation function to access network configuration
H04L 51/52 - User-to-user messaging in packet-switching networks, transmitted according to store-and-forward or real-time protocols, e.g. e-mail for supporting social networking services
H04L 67/12 - Protocols specially adapted for proprietary or special-purpose networking environments, e.g. medical networks, sensor networks, networks in vehicles or remote metering networks
H04L 67/52 - Network services specially adapted for the location of the user terminal
H04W 4/02 - Services making use of location information
H04W 4/029 - Location-based management or tracking services
H04W 4/18 - Information format or content conversion, e.g. adaptation by the network of the transmitted or received information for the purpose of wireless delivery to users or terminals
H04W 4/21 - Services signallingAuxiliary data signalling, i.e. transmitting data via a non-traffic channel for social networking applications
H04W 12/02 - Protecting privacy or anonymity, e.g. protecting personally identifiable information [PII]
Systems, methods, and computer instructions are provided. The method includes retrieving a first set of a media content transmitted by a plurality of interaction clients based on a chronological order, wherein the first set of media content has been saved as part of communications of ephemeral messages between at least two users of the plurality of interaction clients. The method further includes creating a visual representation of the first set of media content, and causing to display, on at least one of the plurality of interaction clients, the visual representation the first set of media content.
G06F 3/0484 - Interaction techniques based on graphical user interfaces [GUI] for the control of specific functions or operations, e.g. selecting or manipulating an object, an image or a displayed text element, setting a parameter value or selecting a range
G06F 3/0488 - Interaction techniques based on graphical user interfaces [GUI] using specific features provided by the input device, e.g. functions controlled by the rotation of a mouse with dual sensing arrangements, or of the nature of the input device, e.g. tap gestures based on pressure sensed by a digitiser using a touch-screen or digitiser, e.g. input of commands through traced gestures
H04L 51/046 - Interoperability with other network applications or services
H04L 51/216 - Handling conversation history, e.g. grouping of messages in sessions or threads
72.
DEVICE-BASED IMAGE MODIFICATION OF DEPICTED OBJECTS
A system of machine learning schemes can be configured to efficiently perform image processing tasks on a user device, such as a mobile phone. The system can selectively detect and transform individual regions within each frame of a live streaming video. The system can selectively partition and toggle image effects within the live streaming video.
G06F 3/04845 - Interaction techniques based on graphical user interfaces [GUI] for the control of specific functions or operations, e.g. selecting or manipulating an object, an image or a displayed text element, setting a parameter value or selecting a range for image manipulation, e.g. dragging, rotation, expansion or change of colour
G06F 3/04817 - Interaction techniques based on graphical user interfaces [GUI] based on specific properties of the displayed interaction object or a metaphor-based environment, e.g. interaction with desktop elements like windows or icons, or assisted by a cursor's changing behaviour or appearance using icons
G06F 3/0482 - Interaction with lists of selectable items, e.g. menus
Systems, methods, and computer readable media for an authentication orchestration system. Example methods include receiving, from an authentication client, an authentication request, the authentication request comprising an indication of an account and an indication of a goal authentication level. The method further includes accessing a current authentication level and adjusting, based on a risk level, the goal authentication level to an adjusted goal authentication level. The method further includes selecting a challenge method of a plurality of challenge methods based on a difference between the adjusted goal authentication level and the current authentication level. The method further includes performing the selected challenge method with a user associated with the account, and causing to be sent, to the authentication client, an indication of whether the adjusted authentication level was achieved.
A waveguide for use in a virtual reality or augmented reality device, is disclosed. The waveguide comprising an input region configured to couple light into the waveguide so that it propagates under total internal reflection (TIR) within the waveguide, and an output region comprising optical structures configured to receive image bearing light from the input region. The output region comprises a first and second zones comprising optical structures, and a transition zone between the first and second zones for blending between the first and the second zone thereby reducing the visible appearance of the boundary between them.
A head-worn device system includes one or more cameras, one or more display devices and one or more processors. The system also includes a memory storing instructions that, when executed by the one or more processors, configure the system to generate a virtual object, generate a virtual object collider for the virtual object, determine a conic collider for the virtual object, provide the virtual object to a user, detect a landmark on the user's hand in the real-world, generate a landmark collider for the landmark, and determine a selection of the first virtual object by the user based on detecting a collision between the landmark collider with the conic collider and with the virtual object collider.
G06F 3/04815 - Interaction with a metaphor-based environment or interaction object displayed as three-dimensional, e.g. changing the user viewpoint with respect to the environment or object
G06T 19/00 - Manipulating 3D models or images for computer graphics
G06V 40/20 - Movements or behaviour, e.g. gesture recognition
Aspects of the present disclosure involve a system comprising a computer-readable storage medium storing a program and a method for generating a summary based on trip information. The program and method include operations for: determining that one or more criteria associated with a user correspond to a trip taken by the user during a given time interval; retrieving a plurality of visual media items generated by a client device of the user during the given time interval; determining location information for the plurality of visual media items; automatically generating a trip graphic to represent the trip based on the plurality of visual media items generated by the user during the given time interval and the determined location information; and causing the trip graphic to be displayed on the client device.
G06F 16/783 - Retrieval characterised by using metadata, e.g. metadata not derived from the content or metadata generated manually using metadata automatically derived from the content
G06F 16/787 - Retrieval characterised by using metadata, e.g. metadata not derived from the content or metadata generated manually using geographical or spatial information, e.g. location
H04L 51/043 - Real-time or near real-time messaging, e.g. instant messaging [IM] using or handling presence information
H04L 51/046 - Interoperability with other network applications or services
H04L 51/52 - User-to-user messaging in packet-switching networks, transmitted according to store-and-forward or real-time protocols, e.g. e-mail for supporting social networking services
78.
EFFICIENT AVATAR CREATION WITH MESH PENETRATION AVOIDANCE
Described is a system for efficient avatar creation with mesh penetration avoidance by receiving a first body characteristic for a first virtual avatar, accessing a first body mesh of the first virtual avatar that corresponds to the first body characteristic, receiving a first accessory characteristic for the first virtual avatar, accessing a first accessory mesh that corresponds to the first accessory characteristic, and identifying a first portion of the first body mesh that is determined to penetrate the first accessory mesh, the penetration being determined prior to receiving the first body characteristic. The system modifies the first portion of the first body mesh that is determined to penetrate the first accessory mesh to generate an updated body mesh, applies the first accessory mesh to the updated first body mesh to generate the first virtual avatar, and displays the first virtual avatar.
Method of generating modified media content items for sharing to external applications starts with a processor receiving a media content item from a client device. Processor causes a sharing interface to be displayed on the client device. Sharing interface includes selectable items associated with external applications. Processor receives from the client device a selection of a first selectable item of the selectable items that is associated with a first external application of the external applications. Processor determines an attribute associated with the media content item. Processor generates a modified media content item based on the first external application and the attribute associated with the media content item and causes the modified media content item to be displayed in the first external application activated on the client device. Other embodiments are disclosed herein.
G06F 3/0482 - Interaction with lists of selectable items, e.g. menus
H04L 51/52 - User-to-user messaging in packet-switching networks, transmitted according to store-and-forward or real-time protocols, e.g. e-mail for supporting social networking services
80.
EXTENDED FIELD-OF-VIEW CAPTURE OF AUGMENTED REALITY EXPERIENCES
Augmented reality experiences of a user wearing an electronic eyewear device are captured by at least one camera on a frame of the electronic eyewear device, the at least one camera having a field of view that is larger than a field of view of a display of the electronic eyewear device. An augmented reality feature or object is applied to the captured scene. A photo or video of the augmented reality scene is captured and a first portion of the captured photo or video is displayed in the display. The display is adjusted to display a second portion of the captured photo or video with the augmented reality features as the user moves the user’s head to view the second portion of the captured photo or video. The captured photo or video may be transferred to another device for viewing the larger field of view augmented reality image.
A system is disclosed, including a processor and a memory. The memory stores instructions that, when executed by the processor, configure the system to perform operations. Raw region of interest (ROI) information is obtained, identifying a raw ROI having a raw ROI size, within a first video frame captured by a camera. Motion information representative of motion of the raw ROI relative to a field of view (FOV) of the camera is obtained. The raw ROI information and the motion information are processed to generate a dynamic ROI having a dynamic ROI size larger than the raw ROI size. A second video frame is captured by the camera. A portion of the second video frame defined by the dynamic ROI is processed to generate autoexposure (AE) settings for the camera.
A mixed-reality media content system may be configured to perform operations that include: causing display of image data at a client device, the image data comprising a depiction of an object that includes a graphical code at a position upon the object; detecting the graphical code at the position upon the depiction of the object based on the image data; accessing media content within a media repository based on the graphical code scanned by the client device; and causing display of a presentation of the media content at the position of the graphical code upon the depiction of the object at the client device.
H04N 21/8545 - Content authoring for generating interactive applications
G06F 3/04817 - Interaction techniques based on graphical user interfaces [GUI] based on specific properties of the displayed interaction object or a metaphor-based environment, e.g. interaction with desktop elements like windows or icons, or assisted by a cursor's changing behaviour or appearance using icons
G06K 19/06 - Record carriers for use with machines and with at least a part designed to carry digital markings characterised by the kind of the digital marking, e.g. shape, nature, code
83.
Gallery user interface with last posted message indication
A server maintains a gallery of ephemeral messages. Each ephemeral message is posted to the gallery by a user for viewing by recipients via recipient devices. In response to a gallery view request from any of the recipient devices, the ephemeral messages in the gallery are displayed on the requesting device in automated sequence, each message being displayed for a respective display duration before display of the next message in the gallery. A user interface via which the gallery is viewable includes indicia based on a posting time of that one of the plurality of ephemeral messages which was posted to the gallery last.
G06F 15/16 - Combinations of two or more digital computers each having at least an arithmetic unit, a program unit and a register, e.g. for a simultaneous processing of several programs
G06F 3/0482 - Interaction with lists of selectable items, e.g. menus
G06F 16/16 - File or folder operations, e.g. details of user interfaces specifically adapted to file systems
G06F 16/215 - Improving data qualityData cleansing, e.g. de-duplication, removing invalid entries or correcting typographical errors
Examples relate to systems and methods for generating videos with precise camera control using a camera-conditioned video diffusion transformer (DiT) model. The model performs a denoising process through a series of pretrained video DiT blocks, where camera trajectory information is processed through a camera conditioning branch to generate camera activations. During an initial portion of denoising passes, the process is conditioned on camera activations for an initial subset of video DiT blocks, while later denoising passes proceed without camera conditioning. This approach leverages the insight that camera motion is established early in the denoising process, enabling precise camera control while maintaining high visual quality.
H04N 19/176 - Methods or arrangements for coding, decoding, compressing or decompressing digital video signals using adaptive coding characterised by the coding unit, i.e. the structural portion or semantic portion of the video signal being the object or the subject of the adaptive coding the unit being an image region, e.g. an object the region being a block, e.g. a macroblock
G06V 20/40 - ScenesScene-specific elements in video content
H04N 19/18 - Methods or arrangements for coding, decoding, compressing or decompressing digital video signals using adaptive coding characterised by the coding unit, i.e. the structural portion or semantic portion of the video signal being the object or the subject of the adaptive coding the unit being a set of transform coefficients
85.
SCALABLE MACHINE LEARNING PLATFORM WITH INTEGRATED FEATURE GENERATION AND REAL-TIME MODEL SERVING
A system and method for machine learning platform management includes a unified architecture for developing and deploying machine learning models at scale. The platform integrates feature generation, model training, and inference services through a centralized interface. The system processes source data through a feature platform to generate training datasets and real-time features. A multi-stage training pipeline enables automated model experimentation through configurable workflows combining core frameworks and user modeling code. The platform implements specialized inference services optimized for high-throughput ranking and recommendation use cases, with distributed feature stores and local caching for efficient feature serving. A comprehensive monitoring system tracks model performance, feature distributions, and prediction quality through automated anomaly detection. The platform enables rapid experimentation while maintaining production reliability through automated deployment orchestration, optimized inference engines, and continuous feedback loops for model improvement.
Uploading of a video file is performed by transcoding, encrypting, and uploading portions of the video file in parallel, to reduce total processing and upload time. The processing of the video file may include applying associated augmented reality effects to a raw video recording, to generate an enhanced video recording for transmission and viewing at a recipient device. The uploaded portions of the video file may be assembled into a fragmented file format such as fMP4, in which portions of the video file are stored as fragments.
H04N 19/436 - Methods or arrangements for coding, decoding, compressing or decompressing digital video signals characterised by implementation details or hardware specially adapted for video compression or decompression, e.g. dedicated software implementation using parallelised computational arrangements
H04N 19/40 - Methods or arrangements for coding, decoding, compressing or decompressing digital video signals using video transcoding, i.e. partial or full decoding of a coded input stream followed by re-encoding of the decoded output stream
H04N 19/85 - Methods or arrangements for coding, decoding, compressing or decompressing digital video signals using pre-processing or post-processing specially adapted for video compression
An event detection system is configured to access a repository that contains a collection of media content. The media content may for example include images, videos, audio clips, and the like, wherein the media content comprises features that include: tags (e.g., hashtags or other similar mechanisms to label and sort content); captions that comprises one or more words or phrases; continuous numerical values; geolocation data (e.g., geo-hash, check-in data, coordinates); as well as temporal data (e.g., timestamps).
G06F 16/48 - Retrieval characterised by using metadata, e.g. metadata not derived from the content or metadata generated manually
G06F 16/487 - Retrieval characterised by using metadata, e.g. metadata not derived from the content or metadata generated manually using geographical or spatial information, e.g. location
Techniques are provided for measuring interpupillary distance (IPD) using a mobile device application that integrates with augmented reality (AR) eyewear. The mobile application guides users through a measurement process using the device's camera and sensors to capture facial data. Computer vision algorithms analyze the data to determine IPD measurements while considering facial geometry and depth perception. The system provides real-time positioning feedback and allows manual entry of known IPD values. Measurements are stored in user profiles rather than on AR devices, enabling cross-device synchronization. The integration of IPD measurement into the initial device setup process, combined with flexible measurement options and profile-based storage, provides an accessible solution that maintains accuracy while eliminating the need for specialized measurement hardware.
H04N 23/12 - Cameras or camera modules comprising electronic image sensorsControl thereof for generating image signals from different wavelengths with one sensor only
Aspects of the present disclosure involve a system comprising a computer-readable storage medium storing a program and method for displaying timestamps for views of content collections. The program and method provide for providing access to a content collection of a first user of an application, the first user being associated with a first device; storing a respective timestamp for each view of the content collection by second users of the application, the second users being associated with second devices; receiving, from the first device, a request to display an indication of each view of the content collection by the second users; determining whether the first user has a paid subscription with respect to the application; and providing, on the first device and upon determining that the first user has a paid subscription, display of the views with timestamp information based on the respective timestamp for each view.
G06F 3/0484 - Interaction techniques based on graphical user interfaces [GUI] for the control of specific functions or operations, e.g. selecting or manipulating an object, an image or a displayed text element, setting a parameter value or selecting a range
G06F 21/10 - Protecting distributed programs or content, e.g. vending or licensing of copyrighted material
Eyewear for capturing and processing electroencephalogram (EEG) signals. The eyewear includes one or more biopotential measurement systems attached to rearward extensions of the temple arms of the eyewear. The biopotential measurement systems include an electrode arrangement that contacts the scalp of a user of the eyewear for capturing EEG signals of the user. The eyewear further includes a reference electrode mounted on a frame of the eyewear and in contact with the nose of the user. The eyewear processes the EEG signals to provide user inputs into software applications.
Aspects of the present disclosure involve a system comprising a computer-readable storage medium storing at least one program, method, and user interface to facilitate augmented reality based communication between multiple users over a network. A first user of a first device is enabled to view a real-world environment that is visible to a second user via a second device by causing display, at the first device, of a live camera feed generated at the second device. The live camera feed comprises images of the real-world environment that is visible to the second user. Input data indicative of a selection by the first user of a virtual content item to apply to the real-world environment that is visible to the second user is received. The first device and second device present media objects overlaid on the real-world environment based on the input data.
A system for creating and presenting enhanced voice notes in augmented reality (AR) environments is disclosed. The system enables users to generate personalized voice notes with visual and audio enhancements on mobile devices, and deliver them to recipients wearing AR devices. Voice notes can be customized with AI-generated voice styles, animated visual representations, and spatial audio effects. Recipients experience immersive playback through AR glasses, with voice notes appearing at specified locations or anchored to body parts. The system leverages computer vision, spatial audio processing, and real-time tracking to create context-aware and spatially relevant communications. This approach transforms traditional voice messaging into an engaging, three-dimensional experience that seamlessly integrates with the user's physical environment.
G06F 3/0484 - Interaction techniques based on graphical user interfaces [GUI] for the control of specific functions or operations, e.g. selecting or manipulating an object, an image or a displayed text element, setting a parameter value or selecting a range
G06K 9/00 - Methods or arrangements for reading or recognising printed or written characters or for recognising patterns, e.g. fingerprints
G06T 13/40 - 3D [Three Dimensional] animation of characters, e.g. humans, animals or virtual beings
G06T 19/00 - Manipulating 3D models or images for computer graphics
G10L 21/003 - Changing voice quality, e.g. pitch or formants
H04M 1/72433 - User interfaces specially adapted for cordless or mobile telephones with means for local support of applications that increase the functionality with interactive means for internal management of messages for voice messaging, e.g. dictaphones
93.
SPATIAL SPARSITY EXPLOITATION IN NEURAL NETWORK PROCESSING
Examples described herein relate to neural network processing. Each of a plurality of input feature maps may be processed to obtain respective output feature maps. For each input feature map, a differential feature map is obtained based on differences between corresponding values of a first set of values of the input feature map in respective spatially adjacent segments thereof. A transformation operation is performed on at least a subset of a second set of values of the differential feature map to generate a transformed differential feature map that includes a third set of values. An output feature map is generated and includes a fourth set of values. At least a subset of the fourth set of values is obtained by accumulating respective values of the third set of values with corresponding values of the fourth set of values that were previously generated in the output feature map.
The subject technology receives, by one or more hardware processors implementing a local wireless network, a request from a client device to mirror media content displayed on a screen of the client device on a wearable device. In response to the request, the subject technology causes a display of the media content in a mirroring lens of the wearable device. While the media content is being displayed in the mirroring lens of the wearable device, the subject technology tracks hand gestures of a user wearing the wearable device and viewing the media content displayed in the mirroring lens of the wearable device. The subject technology processes navigational or manipulation data based on the tracked hand gestures and sends a navigation or manipulation instruction to the client device or a mirroring lens processor of the wearable device based on the tracked hand gestures.
A methodology is described that provides access to an augmented reality (AR) component maintained by a messaging server system directly from a web view application. When a user activates, from a web view application executing in the messaging client, a user selectable element that references an AR component, a web view AR system obtains the identification of the AR component, performs validation of the identification and of any additional launch data, and launches a camera view user interface (UI) with the AR component loaded in the camera view UI. Content captured from the camera view UI can be shared to other computing devices.
G06T 15/00 - 3D [Three Dimensional] image rendering
G06F 3/0488 - Interaction techniques based on graphical user interfaces [GUI] using specific features provided by the input device, e.g. functions controlled by the rotation of a mouse with dual sensing arrangements, or of the nature of the input device, e.g. tap gestures based on pressure sensed by a digitiser using a touch-screen or digitiser, e.g. input of commands through traced gestures
Examples relate to systems and methods for integrated content in a conversation interface of a messaging system. The system presents a conversation interface comprising a plurality of conversation cells, each of the plurality of conversation cells corresponding to a different conversation between an individual user and one or more users in a set of users, the set of users being associated with the individual user. The system detects a condition for presenting sponsored content to the individual user and, in response, generates a sponsored conversation cell associated with the sponsored content. The system presents the sponsored conversation cell among the plurality of conversation cells in the conversation interface, the sponsored conversation cell comprising a visual indicator that identifies the sponsored conversation cell as being associated with the sponsored content.
G06F 3/04817 - Interaction techniques based on graphical user interfaces [GUI] based on specific properties of the displayed interaction object or a metaphor-based environment, e.g. interaction with desktop elements like windows or icons, or assisted by a cursor's changing behaviour or appearance using icons
G06F 3/0482 - Interaction with lists of selectable items, e.g. menus
H04L 51/04 - Real-time or near real-time messaging, e.g. instant messaging [IM]
97.
INTERACTIVITY WITH CONTENT INTEGRATED AMONG CONVERSATION CELLS
Examples relate to systems and methods for integrated content in a conversation interface of a messaging system. The system presents, in a conversation interface, a sponsored conversation cell among a plurality of conversation cells, the sponsored conversation cell comprising a visual indicator that identifies the sponsored conversation cell as being associated with sponsored content. The system receives input that selects the sponsored conversation cell and in response to receiving the input, generates a query comprising a request for specific information associated with the sponsored content. The system presents a response to the query comprising the specific information associated with the sponsored content.
Systems and methods described herein relate to a video see-through (VST) head- mounted display (HMD) and to methods for operating such an HMD. In some examples, the HMD has a foveal capture system and a peripheral capture system. The foveal capture system captures high-resolution images in a direction corresponding to a viewing direction of a user and with a narrow field of view. The peripheral capture system captures additional images with a wider field of view and lower angular resolution. The HMD includes one or more processors that dynamically adjust the foveal capture system based on the viewing direction to enable the HMD to capture high-resolution images for providing a foveal view. The HMD may render processed images by combining foveal and peripheral captures. In some examples, this enables recording the real world with resolution similar to that of the human eye while maintaining feasible bandwidth requirements.
Examples relate to systems and methods for integrated content in a conversation interface of a messaging system. The system presents, in a conversation interface, a sponsored conversation cell among a plurality of conversation cells, the sponsored conversation cell comprising a visual indicator that identifies the sponsored conversation cell as being associated with sponsored content. The system receives input that selects the sponsored conversation cell and in response to receiving the input, generates a query comprising a request for specific information associated with the sponsored content. The system presents a response to the query comprising the specific information associated with the sponsored content.
A system includes one or more hardware processors and at least one memory storing instructions that cause the one or more hardware processors to perform operations including retrieving a first set of a media content captured by an interaction client included in a client device, and retrieving a second set of media content captured by the interaction client included in the client device. The operations also include assigning the first set of media content a first ranking value, and assigning the second set of media content a second ranking value, creating a first visual representation of the first set of media content and a second visual representation of the second set of the second set of media content based on the first ranking value and on the second ranking value, and causing to display, on a display of the client device, the first visual representation and the second visual representation.
G06T 11/60 - Editing figures and textCombining figures or text
G06F 3/04845 - Interaction techniques based on graphical user interfaces [GUI] for the control of specific functions or operations, e.g. selecting or manipulating an object, an image or a displayed text element, setting a parameter value or selecting a range for image manipulation, e.g. dragging, rotation, expansion or change of colour