正在加载视频...

视频加载失败

Ollama 0.2 is here! Concurrency is now enabled by default. This unlocks 2 major features: Parallel requests Ollama can now serve multiple requests at the same time, using only a little bit of additional memory for each request. This enables use cases such as: - Handling multiple chat sessions...

219,429 次观看 • 2 年前 •via X (Twitter)

11 条评论

Augustinas Malinauskas 的头像
Augustinas Malinauskas2 年前

This is super cool! Ollama will go in history for making AI easily accessible!🔥

ollama 的头像
ollama2 年前

Let's make the future together! Look back at history later. 🤣 Time to build!! ❤️❤️❤️ Love your work!

Robert Avram 的头像
Robert Avram2 年前

Game changer you guys are the best in the game

ollama 的头像
ollama2 年前

Here to serve! 😍 Thank you for supporting Ollama

hidE ☢️ 的头像
hidE ☢️2 年前

Coincidentally today I was going to add concurrency for my agents in golang. And I didn't know if ollama would support it or not. Now I know it will.

ollama 的头像
ollama2 年前

so awesome! ❤️ Let us know how it goes for you. We're here to listen to feedback.

King 👑 Cyril 🧠 的头像
King 👑 Cyril 🧠2 年前

LFG 🚀

ollama 的头像
ollama2 年前

Go go go!! 🚀🚀🚀

Mike Bird (Hiring) 的头像
Mike Bird (Hiring)2 年前

Way to go Ollama team!

Jush@1Ghz 的头像
Jush@1Ghz2 年前

finally thank you

ollama 的头像
ollama2 年前

😍 please support Ollama!

相关视频

Small Language Models (SML) are the future of AI. "Small" (SML) instead of "Large" (LLM). These small models are highly specialized models with superhuman abilities on specific tasks. Here are two techniques to build these models: • Spectrum • Model Merging I give you a short introduction in the attached video, but here is a quick summary: Spectrum helps us identify the most relevant layers to solve one specific task. We can ignore everything else and focus on fine-tuning these layers. Using Spectrum, we can fine-tune models in a heartbeat. Model Merging combines multiple models into a unique, much better model than any of the individual input models. You can also combine models specialized in different tasks and get a model with multiple abilities. This is the state of the art of productizing models. It's what Arcee.ai's platform does behind the scenes. Arcee collaborated with me on this post and is sponsoring it. There are three main steps to produce a model for your particular use case: 1. You create a dataset by uploading your data. 2. You train a model. At this step, Arcee uses Spectrum and Model Merging to produce a highly specialized model for your task. 3. You can deploy that model to any environment you want. Three important notes: • Training process is 2x faster and 2x cheaper than regular fine-tuning. • Resultant models are smaller and have higher accuracy. • They create these specialized models from open-source models. Check this site so you can fully appreciate how this works: If you want to fine-tune an open-source model, consider Arcee's platform. This is the state of the art.

Santiago

164,162 次观看 • 2 年前