正在加载视频...
视频加载失败
Same model. Up to 65% less. DeepSeek’s new API prices take effect on August 16. Marathon lets latency-tolerant workloads trade wait time for lower inference costs without switching models. ▷ Choose NOW when every second matters. ▷ Choose SOON or LATER when a few minutes are acceptable. ▷ Choose... show more
15,113 次观看 • 1 个月前 •via X (Twitter)
2 条评论

AI BIRD1 个月前
Hi Marathon Team, I’ve been using Marathon Delayed Inference and found that tasks are constantly stuck in queued state. How should I fix this? I can supply more details if required, thanks.

BlockWeb1 个月前
The ability to trade latency for cost is a pretty smart model.
