Q

Qwen2.5 VL 3B Instruct FP8 Dynamic

Developed by RedHatAI
The FP8 quantized version of Qwen2.5-VL-3B-Instruct, supporting visual-text input and text output, and optimizing inference efficiency.
Downloads 112
Release Time : 2/6/2025

Model Overview

This model is a quantized version based on Qwen2.5-VL-3B-Instruct. It is optimized through FP8 weight quantization and activation quantization, and supports efficient inference using vLLM. It is suitable for multimodal understanding and generation tasks.

Model Features

FP8 Quantization
Both weight quantization and activation quantization are FP8, significantly improving inference efficiency.
Multimodal Support
Supports visual-text input and text output, suitable for complex multimodal tasks.
Efficient Inference
After optimization, it supports efficient deployment using vLLM, improving inference speed.

Model Capabilities

Visual Question Answering
Image Description Generation
Multimodal Inference
Document Understanding
Chart Analysis

Use Cases

Education
Educational Content Understanding
Parse the image and text content in educational materials to assist learning.
Achieved an accuracy of 45.78% on the MMMU validation set.
Business
Document Analysis
Automatically parse the image and text information in business documents.
Achieved an ANLS score of 92.40% on the DocVQA validation set.
Research
Scientific Chart Understanding
Parse the charts and data in scientific papers.
Achieved a lenient correct rate of 80.72% on the ChartQA test set.
Featured Recommended AI Models
AIbase
Empowering the Future, Your AI Solution Knowledge Base
Š 2025AIbase