Screenshot Describer
A vision model doesn't take a file path. It takes bytes, turned into text, sitting inside the same message list you already use for text-only prompts. This lesson builds the one new piece every later lesson depends on: getting an image into the request correct
8 lessons, each with runnable code in the browser.