
125 - VQA for Real Users, with Danna Gurari
How can we build Visual Question Answering system…
42 Minuten
Podcast
Podcaster
Beschreibung
vor 1 Jahr
How can we build Visual Question Answering systems for real users?
For this episode, we chatted with Danna Gurari, about her work in
building datasets and models towards VQA for people who are blind.
We talked about the differences between the existing datasets, and
Vizwiz, a dataset built by Gurari et al., and the resulting
algorithmic changes. We also discussed the unsolved challenges in
this field, and the new tasks they result in. Danna Gurari is an
Assistant Professor as well as Founding Director of the Image and
Video Computing group in the School of Information at University of
Texas at Austin (UT-Austin). Vizwiz project page:
https://vizwiz.org/ The hosts for this episode are Ana Marasović
and Pradeep Dasigi.
For this episode, we chatted with Danna Gurari, about her work in
building datasets and models towards VQA for people who are blind.
We talked about the differences between the existing datasets, and
Vizwiz, a dataset built by Gurari et al., and the resulting
algorithmic changes. We also discussed the unsolved challenges in
this field, and the new tasks they result in. Danna Gurari is an
Assistant Professor as well as Founding Director of the Image and
Video Computing group in the School of Information at University of
Texas at Austin (UT-Austin). Vizwiz project page:
https://vizwiz.org/ The hosts for this episode are Ana Marasović
and Pradeep Dasigi.
Weitere Episoden

47 Minuten
vor 1 Jahr


48 Minuten
vor 1 Jahr

46 Minuten
vor 1 Jahr

48 Minuten
vor 1 Jahr
Kommentare (0)