Gme Qwen2 VL 2B Instruct
Qwen logo

Gme Qwen2 VL 2B Instruct

gme-qwen2-vl-2b-instructllms.txt
Qwen
The GME-Qwen2VL series is a unified multimodal Embedding model trained based on the Qwen2-VL multimodal large language model (MLLMs). The GME model supports three types of inputs: text, images, and image-text pairs. All these input types can generate universal vector representations and exhibit excellent retrieval performance.

Pricing

  • Input Tokens: $0.138 /M tokens
  • Output Tokens: $0.138 /M tokens

Input Modalities

  • Text
  • Vision
  • Video

Providers

Baidu gme-qwen2-vl-2b-instruct
Pricing$0.138$0.138
Context0
Max output0
Latency-
Throughput-
Uptime
0.00% uptime 2 days ago
0.00% uptime yesterday
0.00% uptime today

Performance for gme-qwen2-vl-2b-instruct

Uptime is the percentage of requests that succeeded over the past 72 hours. AIHubMix continuously monitors every provider and automatically retries with the next-best provider when one returns an error or responds too slowly; Latency is total round-trip time (lower is better); Throughput is how fast the model writes (tokens per second, higher is better).

Uptime
Loading...
Latency
Loading...
Throughput
Loading...

Frequently asked questions

What is Gme Qwen2 VL 2B Instruct?

The GME-Qwen2VL series is a unified multimodal Embedding model trained based on the Qwen2-VL multimodal large language model (MLLMs). The GME model supports three types of inputs: text, images, and image-text pairs. All these input types can generate universal vector representations and exhibit excellent retrieval performance.