Name: Using Visual Feature Space as a Pivot Across Languages
Start: 2021-10-26T11:15:00-0500
End: 2021-10-26T11:30:00-0500

Back To Schedule

Using Visual Feature Space as a Pivot Across Languages

Feedback form is now closed.

Technical Presentations Group 2: AI for Good + Business Impact/Industry

People can create image descriptions using thousands of languages, but these languages share only one visual space. The aim of this work is to leverage visual feature space to pass information across languages. We show that models trained to generate textual captions in more than one language conditioned on an input image can leverage their jointly trained feature space during inference to pivot across languages. We particularly demonstrate improved quality on a caption generated from an input image, by leveraging a caption in a second language. More importantly, we demonstrate that even without conditioning on any visual input, the model demonstrates to have learned implicitly to perform to some extent machine translation from one language to another in their shared visual feature space even though the multilingual captions used for training are created independently.

Authors: Ziyan Yang (Rice University), Leticia Pinto-Alva (University of Southern California), Franck Dernoncourt (Adobe Research), and Vicente Ordóñez (Rice University)

Speakers

Ziyan Yang

Rice University

Tuesday October 26, 2021 11:15am - 11:30am CDT

Technical Presentation

2021 Ken Kennedy AI and Data Science Conference

Ziyan Yang

Attendees (2)

2021 Ken Kennedy AI and Data Science Conference

Sign up or log in to save this to your schedule, view media, leave feedback and see who's attending!

Ziyan Yang

Attendees (2)