AI Times has published Yongduck Kim's patent analysis of OpenAI's multimodal interaction interface.
The patent analyzed is a technique for creating a natural language response based on a visual context, recognizing the user behavior itself as an input to the LLM, such as clicking or dragging a specific area in the image. The US patent US 12,039,431 B1 registered by OpenAI describes a multimodal interaction structure that handles text, images, and user UI input integratively.
Yongduck Kim patent attorney explained the patent’s technical content easily and clearly, and introduced the potential for use in a variety of industries such as healthcare, education, and commerce.
You can find the full article at the link below.
Read the column
This article reflects the information available when it was published. Contact us to discuss your circumstances.
