1,000 
首次购买 $50 或以上可获赠
Generate detailed text descriptions from any image. Useful for accessibility alt text, content analysis, OCR-style readouts, and visual Q&A.
Accessibility (alt text and aria-labels for screen readers), content moderation and tagging, OCR-style readouts of text in images, visual Q&A ("what brand is this watch?"), product cataloguing, and turning reference images into prompts for other generators.
Choose from short, standard, or detailed output lengths. Optionally ask a specific question about the image alongside the description. Powered by Qwen VL Max — Alibaba's flagship multimodal model — with more vision-capable models being added soon.
本网站使用 cookie 实现基本功能、其他功能以及统计用途。详情请参阅Cookie 政策。
首次购买 $50 或以上可获赠
适用于全部 100 多个生成器首次购买 $10 或以上可获赠
适用于全部 100 多个生成器