Video yükleniyor...

Video Yüklenemedi

Ana Sayfaya Dön

✨ Managed to load my custom GPT script into an open source LLM called OpenHermes 2.5 Mistral 7B I used which makes it super easy to run open source LLMs on your laptop without any cloud servers or OpenAI or ChatGPT, just your own GPU in your MacBook! I...

745,285 görüntüleme • 2 yıl önce •via X (Twitter)

10 Yorum

@levelsio profil fotoğrafı
@levelsio2 yıl önce

Sometimes I get this, it breaks into ### Response: inside the sentence and it all gets rekt, any fixes? @MaxRovensky

Rodrigo Rocco 👨‍💻📈📗 from JobBoardSearch 🔎 profil fotoğrafı
Rodrigo Rocco 👨‍💻📈📗 from JobBoardSearch 🔎2 yıl önce

I can't believe all this stuff, I installed on my windows laptop and is faster than GPT4, the OpenHermes model is just mind blowing, he also said they are building the Hermes 2.5 vision model

𝑨𝒓𝒕𝒊𝒇𝒊𝒄𝒊𝒂𝒍 𝑮𝒖𝒚 profil fotoğrafı
𝑨𝒓𝒕𝒊𝒇𝒊𝒄𝒊𝒂𝒍 𝑮𝒖𝒚2 yıl önce

Hey @Teknium1 talking about u

@levelsio profil fotoğrafı
@levelsio2 yıl önce

@Teknium1 Oh damn the guy that made it follows me W T F

Alvin De Cruz profil fotoğrafı
Alvin De Cruz2 yıl önce

Have you tried using LM Studio as a server with your own frontend. Then you can access it via localhost then connect to OpenAI’s DALL-E 3 to create images.

Brandon Braner profil fotoğrafı
Brandon Braner2 yıl önce

Any chance you could make a tutorial on how you trained it on your data. I’ve been looking into Lora but still not fully following it. Doesn’t help I’m an Ilm newb

Duncan — e/zucc profil fotoğrafı
Duncan — e/zucc2 yıl önce

Great success! Now all you need is custom Actions support (HTTP endpoints with auth) and embedded image generation can be done Image generation serving model could be hosted on the same laptop (slow) or some minimal maintenance cheap private cloud

Atlas3D profil fotoğrafı
Atlas3D2 yıl önce

If you don't know about @Teknium1 you two teaming up would be legendary.

codejake 🇺🇸 profil fotoğrafı
codejake 🇺🇸2 yıl önce

I’ve been using on my M3 Max and like it a lot. Modelfiles are really useful as GPTs-like. Any clue how compares?

Dr. Daniel Bender profil fotoğrafı
Dr. Daniel Bender2 yıl önce

The good old text adventure is back and is smarter than ever. Love the idea!

Benzer Videolar

Elon exposes OpenAI -- OpenAI execs betrayed the founding mission as soon as there was revenue and profits to be had -- If it started as a for-profit, Elon would own 50% Elon Musk: “There's a mountain of evidence that shows that OpenAI was created as an open source nonprofit.” “That's the exact description in the incorporation documents. They have completely violated that.” “And they tried to change the definition of OpenAI to mean open to everyone instead of open source, even though it always meant open source.” “ I came up with the name. That’s how I know.” “ I mean, essentially, since I came up with the idea for the company, named it, provided the A, B, and C rounds of funding, recruited the critical personnel, and told them everything I know, if that had been a commercial corporation, I'd probably own half the company.” “It was totally at my discretion. I could have done that.” “But I created it as an open source nonprofit for the world.” Chamath: “ Do you think the right thing to do is to take those models and just open source them today?” Elon: “ Yeah, I think that is what it was created to do, so it should.” “Try using any of the recent so-called OpenAI open source models, they don't work.” “They open sourced a broken, non-working version of their models as a fig leaf.” “I mean, do you know anyone who's running OpenAI's open source models? Jason: “No.” Elon: “Exactly.”

The All-In Podcast

387,223 görüntüleme • 11 ay önce

Here is how you can install an open-source, enterprise-grade RAG system on your server (with the best document understanding I've seen.) First, something obvious to anyone trying to sell RAG in the market: You are crazy if you think companies will let their data travel to a hosted model. No one wants to send their data anywhere (those who do haven't found an alternative.) Every single company would rather have an air-gapped system with no internet access. GroundX is an open-source RAG system that you can run on your servers (or any cloud provider, as long as you have access to GPUs) and works without a network. (If the military wants to do RAG, this is precisely what they will be looking for.) I installed GroundX on my AWS account and recorded a video to show you how to use it. There are two services you can use: 1. Ingest: This service uses a pretrained vision model to ingest and understand your knowledge base. 2. Search: This service combines text and vector search with a fine-tuned re-ranker model to retrieve information from your knowledge base. A quick note about the Ingest service: 99% of people think they need better "retrieval" mechanisms. I think they need better "ingestion." That's where this service comes in! Ingest "understands" your documents in a way I haven't seen before. After you try it, you'll realize why showing your LLM your raw documents is a bad idea. In the video, I use a free tool called X-Ray to test a document and understand how the Ingest service breaks it down. You can access this tool by signing up for a free GroundX cloud account and uploading your documents. You'll see a bit more about this in the video.

Santiago

89,664 görüntüleme • 1 yıl önce