AI image captioning that adapts to your workflow.
What used to take hours, VisionAI Foto handles in seconds: your photos are automatically captioned using artificial intelligence — in the language, writing style, and content you configure yourself. Matched to your brand, your product, and your audiences.
- Processing exclusively in Germany
- No training on your images
- Fully GDPR-compliant

Photo: Marc Conzelmann
One photo in, one searchable asset out
The same archive, once without and once with VisionAI Foto — headline, description, and keywords are created in a single pass.
Manual captioning
Hours
for headlines, descriptions, and keywords by hand
- Open each image individually and caption it by hand
- Think up and type keywords image by image
- Every additional language means double the maintenance
- Older stock stays uncaptioned — and therefore unfindable
- Style and scope vary depending on who's captioning
With VisionAI Foto
Seconds
per image — automatically on every upload, if you like
- Headline, description, and keywords in a single pass
- In any language you need — including Easy Language
- Logos and lettering are recognized
- Consistent style by your rules, matched to your brand
- Runs automatically on every upload or targeted at selected images
- Processing in Germany, GDPR-compliant, no AI training
You set the rules
In every single workflow, you define precisely how our AI should interpret and caption your images — from context through style and language to the question of when it even runs.
Context
Specify where and when a photo was taken. That way our AI knows, for example, that it's your Christmas party in Munich — information that could never be inferred from the image alone.
Headline
Whether short for internal use or detailed for external communication: your own prompt determines the style and length of the headline, in any language you need, and optionally in Easy Language too.
Image description
So every image becomes findable by content too, you configure the description independently of the headline — also in any language, and optionally in Easy Language for maximum accessibility.
Keywords
Define which elements get captured as keywords — logos and lettering are recognized too, of course. If BMW, say, was a mobility partner of an event, the keyword "BMW" instantly finds every matching image.
Prefixes and suffixes
Add fixed text blocks before or after every headline and description, such as a campaign name or a copyright notice.
Overwrite rule
Decide whether VisionAI Foto replaces existing headlines and descriptions or leaves them untouched — useful, for example, if you don't want to overwrite older stock that's already been captioned by hand.
Fixed keywords
Define keywords that get added automatically on every run, such as your company or campaign name — regardless of what's actually visible in the image.
Automatically on every upload
Every new image is captioned instantly, with zero manual effort — ideal for editorial or marketing teams with a high upload volume.
Limited to spaces or folders
Automate specific areas only, such as just the folder for press or social media content, while other areas stay untouched.
Manually on selected images
Apply VisionAI Foto selectively, for example for after-the-fact corrections or individual images from an older archive.
How VisionAI Foto reads a picture
Real photos from different archives — the picture on the left, on the right the three IPTC fields VisionAI Foto wrote for it — in German, the language this workflow was set to. Click through, or just keep scrolling.

Photo: Marc Conzelmann1 / 5
- Headline
- Verleihung des Freiheitspreises der Medien 2025 an Joachim Gauck
- Description
- Das Bild zeigt eine Veranstaltung des Ludwig Erhard Gipfels, organisiert von der Weimer Media Group. Im Vordergrund steht ein Rednerpult mit dem Logo des Gipfels, an dem Joachim Gauck eine Rede hält. Im Hintergrund ist eine Leinwand zu sehen, die den Preisträger Joachim Gauck für den Freiheitspreis der Medien 2025 ankündigt. Das Publikum sitzt aufmerksam vor der Bühne, während Kameras die Veranstaltung dokumentieren.
- Keywords
- Weimer Media GroupFreiheitspreis der MedienJoachim Gauck2025VeranstaltungRednerpultPublikumKamerasMedienpreisTegernsee SummitPreisverleihungMarketingEventfotografieBühneLudwig-Erhard-GipfelPorträtKameratechnikLogo
Always the most capable AI behind the scenes
Standing still isn't an option — not even for the AI behind it. At VAULO we designed our DAM so we're never locked into a single provider: if a more capable model establishes itself on the market, we swap out the technology on short notice.
Right now, VisionAI Foto runs on OpenAI's models. For you, nothing about that ever changes — not the contract or terms, and not the data protection standards: processing happens exclusively in Germany, your images never feed into any AI training, and everything stays fully GDPR-compliant.
Your data stays your data
Your images belong to you — and stay that way even after processing by our AI. All processing happens exclusively on servers in Germany — your image data never leaves German jurisdiction at any point.
Your photos and the text generated from them are never used for training purposes, neither by us nor by the AI model in use. The entire process is designed to be fully GDPR-compliant, from processing through storage.
Answers to the most common questions
Still have questions? We're happy to help you personally.
Test VisionAI Foto on your own stock
Show us a typical selection of your images — we'll configure a workflow and you'll see the result on your own material.
- Processing exclusively in Germany
- No training on your images
- Fully GDPR-compliant



