Patent · US Active

Image manipulation by text instruction

US11900517B2 · kind B2 · utility

0Cited by
0References
26Claims
0Family size

Assignee

Inventors

Key dates

Filing dateDec 20, 2022
Grant dateFeb 13, 2024
Priority date
Expiry dateDec 20, 2042

Classification

  • Technology area (CPC G)Physics
  • CPC primaryG06T9/002
  • WIPO fieldComputer technology
  • WIPO sectorElectrical engineering

Abstract

A method for generating an output image from an input image and an input text instruction that specifies a location and a modification of an edit applied to the input image using a neural network is described. The neural network includes an image encoder, an image decoder, and an instruction attention network. The method includes receiving the input image and the input text instruction; extracting, from the input image, an input image feature that represents features of the input image using the image encoder; generating a spatial feature and a modification feature from the input text instruction using the instruction attention network; generating an edited image feature from the input image feature, the spatial feature and the modification feature; and generating the output image from the edited image feature using the image decoder.

Source: USPTO / EPO open patent data. Objective bibliographic and citation counts.