Video wird geladen...
Video konnte nicht geladen werden
Same model. Up to 65% less. DeepSeek’s new API prices take effect on August 16. Marathon lets latency-tolerant workloads trade wait time for lower inference costs without switching models. ▷ Choose NOW when every second matters. ▷ Choose SOON or LATER when a few minutes are acceptable. ▷ Choose... show more
15,113 Aufrufe • vor 1 Monat •via X (Twitter)
2 Kommentare

AI BIRDvor 1 Monat
Hi Marathon Team, I’ve been using Marathon Delayed Inference and found that tasks are constantly stuck in queued state. How should I fix this? I can supply more details if required, thanks.

BlockWebvor 1 Monat
The ability to trade latency for cost is a pretty smart model.
