$ cat /etc/cookies.conf
We use cookies to understand how people use this site.
Analytics cookies help us improve your experience.
They are off by default. Nothing tracks you until you say so.
$ select cookie_preferences
Members-Only
Recent Talks & Demos are for members only
You must be an AI Tinkerers active member to view these talks and demos.
Explore a deep learning model that transforms images into spoken captions using attention on Flickr data, providing real‑time object and scene descriptions for visually‑impaired users.
A deep learning model that can explain the content of an image in the form of speech through caption generation with the attention mechanism(backbone of LLM) on the Flickr data. This is used for visually impaired individuals using artificial intelligence and computer vision. These projects typically focus on creating systems that can describe the surrounding environment, read text, and recognize objects, providing real-time assistance.
Loading recent emails...