2 papers
cs.CV2025
Composite Sketch+Text Queries for Retrieving Objects with Elusive Names and Complex Interactions
Prajwal Gatti, Kshitij Parikh, Dhriti Prasanna Paul +2
Non-native speakers with limited vocabulary often struggle to name specific objects despite being able to visualize them, e.g., people outside Australia searching for numbats. Furt…
cs.CV2025
PatentLMM: Large Multimodal Model for Generating Descriptions for Patent Figures
Shreya Shukla, Nakul Sharma, Manish Gupta +1
Writing comprehensive and accurate descriptions of technical drawings in patent documents is crucial to effective knowledge sharing and enabling the replication and protection of i…