Computer Vision / Video Analytics

Share Your Science: Microsoft Developing Applications for the Visually Impaired

AI-Generated Summary

  • Microsoft Research uses deep learning to create applications for people who are blind or visually impaired.
  • Researchers trained their language model to describe images or scenes in natural language using Tesla M40 and TITAN X GPUs with the cuDNN-accelerated Caffe framework.
  • The work powers Microsoft's Seeing AI research project, which uses computer vision and natural language processing to describe surroundings, read text, answer questions and identify emotions on people's faces.

Next Steps

  • Watch the Seeing AI video to see the project in action.
  • Share your GPU-accelerated science at http://nvda.ly/Vpjxr and with the world on #ShareYourScience.
  • Watch more researchers share how accelerated computing benefits their work at http://nvda.ly/X7WpH.
Powered by NVIDIA Nemotron. AI-generated content may summarize information incompletely. Verify important information. Learn more

Ken Tran, Senior Research Engineer at Microsoft Research shares how they are using deep learning to create applications for people who are blind or visually impaired.
Using Tesla M40 and TITAN X GPUs with the cuDNN-accelerated Caffe deep learning framework, they have trained their language model to describe images or scenes in natural language.
The work is being used for Microsoft’s research project called Seeing AI that uses computer vision and natural language processing to describe a person’s surroundings, read text, answer questions and even identify emotions on people’s faces.

Share your GPU-accelerated science with us at http://nvda.ly/Vpjxr and with the world on #ShareYourScience.
Watch more scientists and researchers share how accelerated computing is benefiting their work at http://nvda.ly/X7WpH

Discuss (0)

Tags