Qwen2.5-VL-7B-Instruct
Vision-language model from Qwen that understands images, documents and video and answers in text.
Apache-2.0Permissive
SaaS: YesClosed product: Yes
- Deployment
- transformers
GPT-4o vision in the AI category. These are the open alternatives we recommend looking at.
Vision-language model from Qwen that understands images, documents and video and answers in text.
SaaS: YesClosed product: Yes
Small vision-language model from Hugging Face that describes and answers questions about images. Based on HuggingFaceTB/SmolLM2-1.7B-Instruct.
SaaS: YesClosed product: Yes
This is guidance, not legal advice (inte juridisk rådgivning). Check with a lawyer before deciding.
We use no tracking cookies. We only store what is necessary in your browser: your theme choice, your sign-in and that you have seen this notice. Read more