Text-image retrieval (T2I) refers to the task of recovering all images relevant to a keyword query. Popular datasets for text-image retrieval, such as Flickr30k, VG, or MS-COCO, utilize annotated image captions, e.g., “a man playing with a kid”, as a surrogate for queries. With such surrogate queries, current multi-modal machine learning models, such as CLIP or BLIP, perform remarkably well.
Articles
Related Articles
March 31, 2026
Tempranillo: Non-Speculative Early Register Release
Abstract: Limited by the breakdown of technology scaling, CPU architects are looking for creative solutions to...
Read More >
1 MIN READING
May 9, 2024
A Balanced Distributed Cascode Power Amplifier With an Integrated Chebyshev Load Balancer for Full-Duplex Wireless Operation
This work proposes a fully integrated transmitter front end based on a balanced distributed cascode power...
Read More >
1 MIN READING
March 30, 2017
Carrier aggregation receiver employing direct recentred offset receivers
Carrier aggregation supports an increased total bandwidth, data rate and utilization of available fragmented spectrum, where...
Read More >
1 MIN READING