ByteDance
Every ByteDance model UltraGPT can reach, kept in sync with the live catalog — pricing, context window and capabilities for each one.
1 model from ByteDance Up to 128K tokens of context
1
Models tracked
1
Providers
0
Reasoning models
1
See images (vision)
128K
Largest context window
0
Free to run
All ByteDance models.
Search and filter this provider's lineup — pricing per million tokens, context window, modalities, reasoning effort and independent benchmark scores where available.
1 of 1 models
UI-TARS-1.5 is a multimodal vision-language agent optimized for GUI-based environments, including desktop interfaces, web browsers, mobile systems, and games. Built by ByteDance, it builds upon the UI-TARS framework with reinforcement...
VisionStreamingJSON mode
Context window128K tok
Max output2.0K tok
ReleasedJul 22, 2025
Knowledge cutoffJan 31, 2025
Image + TextText
Input$0.1 / 1M tok
Output$0.2 / 1M tok
Cache read$0.1 / 1M tok
