What it may solve
Digests the images in a prompt with a vision model before admission, so any model — a text-only one included — can read them without the session ever switching models, and adds a describe_image tool for image paths.
Imported third-party catalog description; not a Registry verification conclusion.