Computer Vision / Video Analytics

Share Your Science: Training a Machine to Answer Questions About Images

AI-Generated Summary

  • Aishwarya Agrawal and her collaborators are using NVIDIA GPUs and deep learning to automatically answer a wide range of questions about arbitrary images.
  • The system may one day help the visually impaired navigate real-world environments, such as determining when it is safe to cross the street.

Next Steps

Powered by NVIDIA Nemotron. AI-generated content may summarize information incompletely. Verify important information. Learn more

Aishwarya Agrawal, PhD student at Virginia Tech shares how her team is using NVIDIA GPUs and deep learning to automatically answer a wide range of questions about arbitrary images.
According to Agrawal and her collaborators, the system may one day be used by the visually impaired to help navigate real-world environments, such as informing the user when it is safe to cross the street.

To learn more, try the online demo or read their research paper “VQA: Visual Question Answering”.
Share your GPU-accelerated science with us at http://nvda.ly/Vpjxr and with the world on #ShareYourScience.
Watch more scientists and researchers share how accelerated computing is benefiting their work at http://nvda.ly/X7WpH
 

Discuss (0)

Tags