正在加载视频...
视频加载失败
Here is how to deploy and serve any LLM on HF with a single command in less than 3 minutes with llama.cpp $ bash -c "$(curl -s
143,145 次观看 • 2 年前 •via X (Twitter)
8 条评论

Georgi Gerganov2 年前
More info

Fernando Vidal2 年前
Would be cool to have some kind of auth system, where you can have it check against a list of auth tokens before serving the request.

Donneker2 年前
thanks for the demo, nice done

bornjre2 年前
Does server binary supports LLaVA/BakLLaVA models ?

wwwwg2 年前
@memdotai mem it

Mem2 年前
@ggerganov Saved! Here's the compiled thread: 🪄 AI-generated summary: "This thread provides instructions on how to deploy and serve any LLM on HF with a single command in less than 3 minutes using llama.cpp. More information can be found at the...

Tim Wu2 年前
Time to fill up my runpod credits. 😁 It would be lovely if supports llava.

Filippo Broggini2 年前
They should have made you CEO of OpenAI 😅😇💪

