正在加载视频...

视频加载失败

✨ Managed to load my custom GPT script into an open source LLM called OpenHermes 2.5 Mistral 7B I used which makes it super easy to run open source LLMs on your laptop without any cloud servers or OpenAI or ChatGPT, just your own GPU in your MacBook! I...

745,229 次观看 • 2 年前 •via X (Twitter)

10 条评论

@levelsio 的头像
@levelsio2 年前

Sometimes I get this, it breaks into ### Response: inside the sentence and it all gets rekt, any fixes? @MaxRovensky

Rodrigo Rocco 👨‍💻📈📗 from JobBoardSearch 🔎 的头像
Rodrigo Rocco 👨‍💻📈📗 from JobBoardSearch 🔎2 年前

I can't believe all this stuff, I installed on my windows laptop and is faster than GPT4, the OpenHermes model is just mind blowing, he also said they are building the Hermes 2.5 vision model

𝑨𝒓𝒕𝒊𝒇𝒊𝒄𝒊𝒂𝒍 𝑮𝒖𝒚 的头像
𝑨𝒓𝒕𝒊𝒇𝒊𝒄𝒊𝒂𝒍 𝑮𝒖𝒚2 年前

Hey @Teknium1 talking about u

@levelsio 的头像
@levelsio2 年前

@Teknium1 Oh damn the guy that made it follows me W T F

Alvin De Cruz 的头像
Alvin De Cruz2 年前

Have you tried using LM Studio as a server with your own frontend. Then you can access it via localhost then connect to OpenAI’s DALL-E 3 to create images.

Brandon Braner 的头像
Brandon Braner2 年前

Any chance you could make a tutorial on how you trained it on your data. I’ve been looking into Lora but still not fully following it. Doesn’t help I’m an Ilm newb

Duncan — e/zucc 的头像
Duncan — e/zucc2 年前

Great success! Now all you need is custom Actions support (HTTP endpoints with auth) and embedded image generation can be done Image generation serving model could be hosted on the same laptop (slow) or some minimal maintenance cheap private cloud

Atlas3D 的头像
Atlas3D2 年前

If you don't know about @Teknium1 you two teaming up would be legendary.

codejake 🇺🇸 的头像
codejake 🇺🇸2 年前

I’ve been using on my M3 Max and like it a lot. Modelfiles are really useful as GPTs-like. Any clue how compares?

Dr. Daniel Bender 的头像
Dr. Daniel Bender2 年前

The good old text adventure is back and is smarter than ever. Love the idea!

相关视频

Elon exposes OpenAI -- OpenAI execs betrayed the founding mission as soon as there was revenue and profits to be had -- If it started as a for-profit, Elon would own 50% Elon Musk: “There's a mountain of evidence that shows that OpenAI was created as an open source nonprofit.” “That's the exact description in the incorporation documents. They have completely violated that.” “And they tried to change the definition of OpenAI to mean open to everyone instead of open source, even though it always meant open source.” “ I came up with the name. That’s how I know.” “ I mean, essentially, since I came up with the idea for the company, named it, provided the A, B, and C rounds of funding, recruited the critical personnel, and told them everything I know, if that had been a commercial corporation, I'd probably own half the company.” “It was totally at my discretion. I could have done that.” “But I created it as an open source nonprofit for the world.” Chamath: “ Do you think the right thing to do is to take those models and just open source them today?” Elon: “ Yeah, I think that is what it was created to do, so it should.” “Try using any of the recent so-called OpenAI open source models, they don't work.” “They open sourced a broken, non-working version of their models as a fig leaf.” “I mean, do you know anyone who's running OpenAI's open source models? Jason: “No.” Elon: “Exactly.”

The All-In Podcast

387,223 次观看 • 9 个月前

Here is how you can install an open-source, enterprise-grade RAG system on your server (with the best document understanding I've seen.) First, something obvious to anyone trying to sell RAG in the market: You are crazy if you think companies will let their data travel to a hosted model. No one wants to send their data anywhere (those who do haven't found an alternative.) Every single company would rather have an air-gapped system with no internet access. GroundX is an open-source RAG system that you can run on your servers (or any cloud provider, as long as you have access to GPUs) and works without a network. (If the military wants to do RAG, this is precisely what they will be looking for.) I installed GroundX on my AWS account and recorded a video to show you how to use it. There are two services you can use: 1. Ingest: This service uses a pretrained vision model to ingest and understand your knowledge base. 2. Search: This service combines text and vector search with a fine-tuned re-ranker model to retrieve information from your knowledge base. A quick note about the Ingest service: 99% of people think they need better "retrieval" mechanisms. I think they need better "ingestion." That's where this service comes in! Ingest "understands" your documents in a way I haven't seen before. After you try it, you'll realize why showing your LLM your raw documents is a bad idea. In the video, I use a free tool called X-Ray to test a document and understand how the Ingest service breaks it down. You can access this tool by signing up for a free GroundX cloud account and uploading your documents. You'll see a bit more about this in the video.

Santiago

89,664 次观看 • 1 年前

Introducing the Clips chrome extension - the easiest way to send bug reports to agents with video, transcript, and browser debug info captured automatically. 100% free and open source. If you are like me and get tired of manually typing instructions to agents, attaching screenshots, pasting debug logs, and all of that, this might be your new favorite tool. With the Clips chrome extension, you can just click the Clips icon, hit record, and start talking. Visually demonstrate your issue, go through the flow, point out what’s broken. Clips will capture everything on your screen, plus network requests, browser logs, client errors, and all the details around them. And it redacts sensitive information. Then it gives you a link you can send to humans so they can play it and take a look. Or, more importantly, just give it to your agents by just pasting the URL to them. The link has special metadata for agents so just from the URL, the agent can pull all information from the clip automatically. No plugin or MCP server required. That means it can "see and hear" what’s in the video - read the transcript, grab snapshots at any timestamp, and inspect the logs and network requests that were shared with it. So whether you want to quickly demo an issue and send all that context to an agent, or get better bug reports from teammates, recording and sending Clips makes that super easy. Unlike expensive apps like Loom, this is all 100% free and open source. The framework that powers this, plus a bunch of other free applications, is open source too. You can just sign up and use it, or fork it and customize it to your needs. This, in my opinion, is the future of software. Rather than bloated SaaS that charges you a ton of money and still doesn’t even have the things you need, we get free open source canonical apps that you can fork and customize in any way you want. I'll link to all this stuff in the replies. If you try it, let me know your feedback.

Steve (Builder.io)

60,574 次观看 • 1 个月前