Turn any image into a production-ready prompt. The vision AI reads the picture and the language model writes the prompt — entirely on your device.
🔒 Everything runs on your device — your video, images and ideas never leave it. The model downloads once, then works offline.
Your ideas and prompts never leave your device. No uploads, no logs, no one watching. True for every single generation.
The model runs on YOUR hardware, so there's no server bill for anyone. No API keys, no quotas, no paywalls — generate 100 prompts a day.
After the one-time download, the model lives in your browser cache. Turn off Wi-Fi and it still generates — airplane mode, no problem.
No account, no email, no analytics on your prompts. The model has no idea who you are — and neither does anyone else.
Hand your ideas to a cloud tool and they live on a server you can't see. A local model keeps every choice — and every word — with you.
| VideoPrompt — on your device | Typical cloud tools | |
|---|---|---|
| Your content | ✓ Never leaves your device | ✗ Uploaded to their servers, then out of your control |
| Privacy | ✓ Zero risk — no server ever sees it | ✗ May be stored, analyzed, or used for AI training (check their terms) |
| Works offline | ✓ Yes — after the one-time download | ✗ No — needs a connection |
| File size limits | ✓ None | ✗ Common: 50–100MB caps |
| Account & email | ✓ Not needed | ✗ Usually required |
No. The image is analyzed entirely in your browser. Nothing is uploaded — there is no server that could receive it. Load the page, switch on airplane mode, and it still works.
Any format your browser can display (JPG, PNG, WebP, GIF and more) — with no size limit, since the image never leaves your device.
Yes: a compact vision model reads the image — Standard (~560MB) or High Precision (~1.5GB) — and the text model you already use writes the final prompt.
The AI reconstructs a prompt from what the image shows: subject, composition, lighting, style. It cannot know the original prompt used to create the image — but it produces an honest, production-ready interpretation you can refine.
Yes. Prompts generated on your device are yours — the underlying model (Qwen2.5) is open-weight under the Apache 2.0 license. Use them for client work, marketing videos, or anything else.