Загрузка видео...
Не удалось загрузить видео
Same model. Up to 65% less. DeepSeek’s new API prices take effect on August 16. Marathon lets latency-tolerant workloads trade wait time for lower inference costs without switching models. ▷ Choose NOW when every second matters. ▷ Choose SOON or LATER when a few minutes are acceptable. ▷ Choose... show more
15,113 просмотров • 1 месяц назад •via X (Twitter)
Комментарии: 2

AI BIRD1 месяц назад
Hi Marathon Team, I’ve been using Marathon Delayed Inference and found that tasks are constantly stuck in queued state. How should I fix this? I can supply more details if required, thanks.

BlockWeb1 месяц назад
The ability to trade latency for cost is a pretty smart model.
