$ cat /etc/cookies.conf
We use cookies to understand how people use this site.
Analytics cookies help us improve your experience.
They are off by default. Nothing tracks you until you say so.
$ select cookie_preferences
Members-Only
Recent Talks & Demos are for members only
You must be an AI Tinkerers active member to view these talks and demos.
A hands‑on comparison of multimodal AI models, demonstrating how they encode and reason over text, image, and structured data for analytical tasks.
Many different deep learning models, especially LLMs, can be focused and fine-tuned on specific data modals. One very interesting feature of today’s LLMs is multimodal input; this has its uses in encoding image, text, structured, and other sorts of data into the same embedding space. This talk will briefly go over the technical use cases of multimodal models by running an in-code comparison across different AI models and their ability to analyze and reason with data.
Loading recent emails...