ERNIE 4.5 Turbo VL
Baidu · text, image → text
The new version of the Wenxin Yiyan large model significantly improves capabilities in image understanding, creation, translation, and coding. It supports a context length of up to 32K tokens for the first time, with a notable reduction in the latency of the first token.
Input$0.4 /M
Output$1.2 /M
