System and method of generating bounding polygons
Abstract
An example system includes a first and second digital device. The first digital device may be configured to provide an interface displaying an image including a depiction of an object, place a bounding shape around the object, and crop contents of the bounding shape to create a portion. The second digital device may be configured to receive the portion, retrieve high-level features and low-level features, apply first Atrous Spatial Pyramid Pooling (ASPP) to the high-level features to aggregate the high-level features as aggregate features, concatenate results to create the aggregate features, up-sample, apply a convolution to the low-level features, concatenate the aggregate features with the low-level features after convolution to form combined features, segment the combined features to generate a polygonal shape outline along outer boundaries of the first object, and provide the first polygonal shape outline to the first digital device for display.
Claims
exact text as granted — not AI-modified1 . A system comprising: at least one processor; and first memory, the first memory containing instructions to control any number of the at least one processor to: provide a first user interface displaying at least one first image taken by a first image capture device, the first image capture device including a first field of view, the at least one first image including a depiction of a first object, the first object being of a particular type, the first image including a first bounding shape placed by a first user around the depiction of the first object using a shape tool provided by the first interface; extract a first portion of the first image, the first portion including only contents of the bounding shape that are contained within the first bounding shape including the depiction of the first object; retrieve first high-level features and first low-level features from the portion, the first high-level features including low spatial information content and high semantic content of the first portion, the first low-level features including high spatial information content and low semantic content of the first portion; apply first Atrous Spatial Pyramid Pooling (ASPP) to the first high-level features of the first portion to aggregate the first high-level features as first aggregate features, the applying the first ASPP including applying any number of convolutional layers in parallel and at different rates from each other to the first high-level features and concatenating results to create the first aggregate features; up-sample the first aggregate features; apply a convolution to the first low-level features; concatenate the first aggregate features after upsampling with the first low-level features after convolution to form first combined features; segment the first combined features to generate a first polygonal shape outline along first outer boundaries of the first object in the first portion, segmenting comprising batch normalization and application of a rectified linear activation function on the first combined features; and display the first image in the user interface and the first polygonal shape outline of the first object in the user interface, the first image not including the first bounding shape placed by the first user.
Join the waitlist — get patent alerts
Track US2023334813A1 — get alerts on status changes and closely related new filings.
We store only your email — no account needed. See our privacy policy.