Video yükleniyor...

Video Yüklenemedi

Ana Sayfaya Dön

✨ Managed to load my custom GPT script into an open source LLM called OpenHermes 2.5 Mistral 7B I used which makes it super easy to run open source LLMs on your laptop without any cloud servers or OpenAI or ChatGPT, just your own GPU in your MacBook! I...

745,229 görüntüleme • 2 yıl önce •via X (Twitter)

10 Yorum

@levelsio profil fotoğrafı
@levelsio2 yıl önce

Sometimes I get this, it breaks into ### Response: inside the sentence and it all gets rekt, any fixes? @MaxRovensky

Rodrigo Rocco 👨‍💻📈📗 from JobBoardSearch 🔎 profil fotoğrafı
Rodrigo Rocco 👨‍💻📈📗 from JobBoardSearch 🔎2 yıl önce

I can't believe all this stuff, I installed on my windows laptop and is faster than GPT4, the OpenHermes model is just mind blowing, he also said they are building the Hermes 2.5 vision model

𝑨𝒓𝒕𝒊𝒇𝒊𝒄𝒊𝒂𝒍 𝑮𝒖𝒚 profil fotoğrafı
𝑨𝒓𝒕𝒊𝒇𝒊𝒄𝒊𝒂𝒍 𝑮𝒖𝒚2 yıl önce

Hey @Teknium1 talking about u

@levelsio profil fotoğrafı
@levelsio2 yıl önce

@Teknium1 Oh damn the guy that made it follows me W T F

Alvin De Cruz profil fotoğrafı
Alvin De Cruz2 yıl önce

Have you tried using LM Studio as a server with your own frontend. Then you can access it via localhost then connect to OpenAI’s DALL-E 3 to create images.

Brandon Braner profil fotoğrafı
Brandon Braner2 yıl önce

Any chance you could make a tutorial on how you trained it on your data. I’ve been looking into Lora but still not fully following it. Doesn’t help I’m an Ilm newb

Duncan — e/zucc profil fotoğrafı
Duncan — e/zucc2 yıl önce

Great success! Now all you need is custom Actions support (HTTP endpoints with auth) and embedded image generation can be done Image generation serving model could be hosted on the same laptop (slow) or some minimal maintenance cheap private cloud

Atlas3D profil fotoğrafı
Atlas3D2 yıl önce

If you don't know about @Teknium1 you two teaming up would be legendary.

codejake 🇺🇸 profil fotoğrafı
codejake 🇺🇸2 yıl önce

I’ve been using on my M3 Max and like it a lot. Modelfiles are really useful as GPTs-like. Any clue how compares?

Dr. Daniel Bender profil fotoğrafı
Dr. Daniel Bender2 yıl önce

The good old text adventure is back and is smarter than ever. Love the idea!

Benzer Videolar

Elon exposes OpenAI -- OpenAI execs betrayed the founding mission as soon as there was revenue and profits to be had -- If it started as a for-profit, Elon would own 50% Elon Musk: “There's a mountain of evidence that shows that OpenAI was created as an open source nonprofit.” “That's the exact description in the incorporation documents. They have completely violated that.” “And they tried to change the definition of OpenAI to mean open to everyone instead of open source, even though it always meant open source.” “ I came up with the name. That’s how I know.” “ I mean, essentially, since I came up with the idea for the company, named it, provided the A, B, and C rounds of funding, recruited the critical personnel, and told them everything I know, if that had been a commercial corporation, I'd probably own half the company.” “It was totally at my discretion. I could have done that.” “But I created it as an open source nonprofit for the world.” Chamath: “ Do you think the right thing to do is to take those models and just open source them today?” Elon: “ Yeah, I think that is what it was created to do, so it should.” “Try using any of the recent so-called OpenAI open source models, they don't work.” “They open sourced a broken, non-working version of their models as a fig leaf.” “I mean, do you know anyone who's running OpenAI's open source models? Jason: “No.” Elon: “Exactly.”

The All-In Podcast

387,223 görüntüleme • 9 ay önce

Here is how you can install an open-source, enterprise-grade RAG system on your server (with the best document understanding I've seen.) First, something obvious to anyone trying to sell RAG in the market: You are crazy if you think companies will let their data travel to a hosted model. No one wants to send their data anywhere (those who do haven't found an alternative.) Every single company would rather have an air-gapped system with no internet access. GroundX is an open-source RAG system that you can run on your servers (or any cloud provider, as long as you have access to GPUs) and works without a network. (If the military wants to do RAG, this is precisely what they will be looking for.) I installed GroundX on my AWS account and recorded a video to show you how to use it. There are two services you can use: 1. Ingest: This service uses a pretrained vision model to ingest and understand your knowledge base. 2. Search: This service combines text and vector search with a fine-tuned re-ranker model to retrieve information from your knowledge base. A quick note about the Ingest service: 99% of people think they need better "retrieval" mechanisms. I think they need better "ingestion." That's where this service comes in! Ingest "understands" your documents in a way I haven't seen before. After you try it, you'll realize why showing your LLM your raw documents is a bad idea. In the video, I use a free tool called X-Ray to test a document and understand how the Ingest service breaks it down. You can access this tool by signing up for a free GroundX cloud account and uploading your documents. You'll see a bit more about this in the video.

Santiago

89,664 görüntüleme • 1 yıl önce

Introducing the Clips chrome extension - the easiest way to send bug reports to agents with video, transcript, and browser debug info captured automatically. 100% free and open source. If you are like me and get tired of manually typing instructions to agents, attaching screenshots, pasting debug logs, and all of that, this might be your new favorite tool. With the Clips chrome extension, you can just click the Clips icon, hit record, and start talking. Visually demonstrate your issue, go through the flow, point out what’s broken. Clips will capture everything on your screen, plus network requests, browser logs, client errors, and all the details around them. And it redacts sensitive information. Then it gives you a link you can send to humans so they can play it and take a look. Or, more importantly, just give it to your agents by just pasting the URL to them. The link has special metadata for agents so just from the URL, the agent can pull all information from the clip automatically. No plugin or MCP server required. That means it can "see and hear" what’s in the video - read the transcript, grab snapshots at any timestamp, and inspect the logs and network requests that were shared with it. So whether you want to quickly demo an issue and send all that context to an agent, or get better bug reports from teammates, recording and sending Clips makes that super easy. Unlike expensive apps like Loom, this is all 100% free and open source. The framework that powers this, plus a bunch of other free applications, is open source too. You can just sign up and use it, or fork it and customize it to your needs. This, in my opinion, is the future of software. Rather than bloated SaaS that charges you a ton of money and still doesn’t even have the things you need, we get free open source canonical apps that you can fork and customize in any way you want. I'll link to all this stuff in the replies. If you try it, let me know your feedback.

Steve (Builder.io)

60,574 görüntüleme • 1 ay önce