Aishwarya Agrawal, PhD student at Virginia Tech shares how her team is using NVIDIA GPUs and deep learning to automatically answer a wide range of questions about arbitrary images.
According to Agrawal and her collaborators, the system may one day be used by the visually impaired to help navigate real-world environments, such as informing the user when it is safe to cross the street.
To learn more, try the online demo or read their research paper “VQA: Visual Question Answering”.
Share your GPU-accelerated science with us at http://nvda.ly/Vpjxr and with the world on #ShareYourScience.
Watch more scientists and researchers share how accelerated computing is benefiting their work at http://nvda.ly/X7WpH
Share Your Science: Training a Machine to Answer Questions About Images
May 13, 2016
Discuss (0)
AI-Generated Summary
- Aishwarya Agrawal and her collaborators are using NVIDIA GPUs and deep learning to automatically answer a wide range of questions about arbitrary images.
- The system may one day help the visually impaired navigate real-world environments, such as determining when it is safe to cross the street.
Next Steps
- Try the online demo to experience the visual question answering system.
- Read the VQA: Visual Question Answering research paper for detailed methodology and results.
- Share GPU-accelerated science at http://nvda.ly/Vpjxr and with the world on #ShareYourScience.
- Watch more researchers share how accelerated computing benefits their work at http://nvda.ly/X7WpH.
Powered by NVIDIA Nemotron. AI-generated content may summarize information incompletely. Verify important information. Learn more