{"schema_version":"1.0","canonical_url":"https://patentable.app/patents/US-11948078","patent":{"patent_number":"US-11948078","title":"Joint representation learning from images and text","assignee":null,"inventors":[],"filing_date":"2020-08-21T00:00:00.000Z","publication_date":"2024-04-02T00:00:00.000Z","cpc_codes":["G06N","G06F","G06F","G06N","G06V","G06V","G06V","G06V","G06V"],"num_claims":20,"abstract":"The disclosure provides a framework or system for learning visual representation using a large set of image/text pairs. The disclosure provides, for example, a method of visual representation learning, a joint representation learning system, and an artificial intelligence (AI) system that employs one or more of the trained models from the method or system. The AI system can be used, for example, in autonomous or semi-autonomous vehicles. In one example, the method of visual representation learning includes: (1) receiving a set of image embeddings from an image representation model and a set of text embeddings from a text representation model, and (2) training, employing mutual information, a critic function by learning relationships between the set of image embeddings and the set of text embeddings."},"analysis":{"summary":null,"layman_explanation":null,"technical_analysis":null,"business_analysis":null,"faqs":null,"topics":[],"tech_cluster":null},"seo":{"title":"Joint representation learning from images and text","description":"The disclosure provides a framework or system for learning visual representation using a large set of image/text pairs. The disclosure provides, for example, a method of visual representation learning","keywords":[]},"attribution":{"source":"Patentable","source_url":"https://patentable.app","canonical_url":"https://patentable.app/patents/US-11948078","license":"CC-BY-4.0-like","license_terms":"AI-generated analysis on this page (summary, layman_explanation, technical_analysis, business_analysis, faqs) may be reused with attribution and a visible link back to the canonical URL above. Patent abstracts, claims, and bibliographic data are USPTO public domain.","required_link":"https://patentable.app/patents/US-11948078","citation_suggestion":"Patentable. \"Joint representation learning from images and text\" (US-11948078). https://patentable.app/patents/US-11948078","copyright_holder":"Nomic Interactive Technology LLC"},"links":{"html":"https://patentable.app/patents/US-11948078","json":"https://patentable.app/api/llm-context/US-11948078","site":"https://patentable.app","llms_txt":"https://patentable.app/llms.txt"},"generated_at":"2026-05-31T01:32:52.076Z"}