Video wird geladen...
Video konnte nicht geladen werden
Here is how to deploy and serve any LLM on HF with a single command in less than 3 minutes with llama.cpp $ bash -c "$(curl -s
143,140 Aufrufe • vor 2 Jahren •via X (Twitter)
8 Kommentare

Georgi Gerganovvor 2 Jahren
More info

Fernando Vidalvor 2 Jahren
Would be cool to have some kind of auth system, where you can have it check against a list of auth tokens before serving the request.

Donnekervor 2 Jahren
thanks for the demo, nice done

bornjrevor 2 Jahren
Does server binary supports LLaVA/BakLLaVA models ?

wwwwgvor 2 Jahren
@memdotai mem it

Memvor 2 Jahren
@ggerganov Saved! Here's the compiled thread: 🪄 AI-generated summary: "This thread provides instructions on how to deploy and serve any LLM on HF with a single command in less than 3 minutes using llama.cpp. More information can be found at the...

Tim Wuvor 2 Jahren
Time to fill up my runpod credits. 😁 It would be lovely if supports llava.

Filippo Brogginivor 2 Jahren
They should have made you CEO of OpenAI 😅😇💪

