Video yükleniyor...

Video Yüklenemedi

Ana Sayfaya Dön

“don’t train your own model” is common ai advice. it's wrong. your token bill's the proof. today, we’re excited to launch castform into open preview. castform is the easiest way for you to train your own model, on your own data. open-weights models are performant and much cheaper. when...

462,182 görüntüleme • 3 ay önce •via X (Twitter)

89 Yorum

girish profil fotoğrafı
girish3 ay önce

1/4 train your first model today. new users get $50 in free credits 🙂

girish profil fotoğrafı
girish3 ay önce

2/4 you can train a model on your production agent traces so the model can learn from mistakes it made in product with castform.

girish profil fotoğrafı
girish3 ay önce

3/4 you can train an agent to do accurate and fast citation-grounded search over documentation

girish profil fotoğrafı
girish3 ay önce

4/4 you can train a model to red-team & find bugs with castform

girish profil fotoğrafı
girish3 ay önce

train your first model today at

Shreyas 🦷🦷🦷 profil fotoğrafı
Shreyas 🦷🦷🦷3 ay önce

Me right now…. Congrats!!! Can’t wait to try it soon

girish profil fotoğrafı
girish3 ay önce

🤣 you can try it at -> $50 in free credits for new users :) pls give us feedback!!

bonnie profil fotoğrafı
bonnie3 ay önce

@sparab22 Can’t wait to try! Is it suitable for non technical peeps?

girish profil fotoğrafı
girish3 ay önce

@sparab22 give it a go! exporting the starter template and letting claude code go wild has worked surprisingly well - let me know if you need help with setup :)

RG profil fotoğrafı
RG3 ay önce

Love this Girish. We are building infra where you can use Prediction markets to do continuous RL workflows . We call this RLEF - Reinforcement Learning through Economic Feedback. Have some datasets and would love to train own model! @ChronisKod @HotTubLee

girish profil fotoğrafı
girish3 ay önce

@ChronisKod @HotTubLee yes would love to! if you can drop me a note at [email protected] we can get on a call

RG profil fotoğrafı
RG3 ay önce

@ChronisKod @HotTubLee 🙏

brett goldstein profil fotoğrafı
brett goldstein3 ay önce

@berman66 so good!

girish profil fotoğrafı
girish3 ay önce

@berman66 thanks brett! big fan of your launch video commentary :)

brett goldstein profil fotoğrafı
brett goldstein3 ay önce

@berman66 yall won

girish profil fotoğrafı
girish3 ay önce

@berman66 high praise - thanks so much!!

Tanner Le profil fotoğrafı
Tanner Le3 ay önce

This is SICK

girish profil fotoğrafı
girish3 ay önce

was great working with y'all on this!

Militant Hitchhiker ♥ profil fotoğrafı
Militant Hitchhiker ♥3 ay önce

54 times cheaper than a hundred billion dollars still seems like a considerable barrier to entry. But seriously, anything that makes model training more accessible is important. I've spent enough of my own time doing it to know the potential value of such services.

girish profil fotoğrafı
girish3 ay önce

the 54 times cheaper is inference cost not training cost :) small models are surprisingly cheap to post-train, on the order of 100s of $$. and yes, our goal here's to make this as accessible as possible

Vivek Kalyan profil fotoğrafı
Vivek Kalyan3 ay önce

we need more videos with castie!!!

girish profil fotoğrafı
girish3 ay önce

yes more content with castie incoming - goal's the hopefully use castie as a way to better communication post-training research with devs :)

Jonathan profil fotoğrafı
Jonathan3 ay önce

this is huge! why use a generalized model, when you can train your own with ease!

girish profil fotoğrafı
girish3 ay önce

yup exactly! we make training easy so it's a no-brainer

Cassie profil fotoğrafı
Cassie3 ay önce

Please make Castie merch!!!

girish profil fotoğrafı
girish3 ay önce

@thitrangmam 👀👀

Travis profil fotoğrafı
Travis3 ay önce

CASTIEEEEEE IS SO CUTE

girish profil fotoğrafı
girish3 ay önce

yessss 100%

Tuan Le profil fotoğrafı
Tuan Le3 ay önce

wow, this is huge

girish profil fotoğrafı
girish3 ay önce

was great working with y'all on this!

Entryism Enjoyer🐵🧦🏗️ profil fotoğrafı
Entryism Enjoyer🐵🧦🏗️3 ay önce

Are you sure the Nintendo people can't come after you for the pokemon ref?

girish profil fotoğrafı
girish3 ay önce

hahhaha

Vatsl Goswami profil fotoğrafı
Vatsl Goswami3 ay önce

can these models perform complex tasks like image segmentation from limited data? what kind and how much data would one need for a self-trained model to be able to compete with the foundational ones?

girish profil fotoğrafı
girish3 ay önce

@vatslgoswami currently, we support text models only -> multimodal coming by end of the summer :) on data, depends on the task but have seen pretty good results from 100s to 1000s of samples

Vatsl Goswami profil fotoğrafı
Vatsl Goswami3 ay önce

this is really cool stuff. can't wait to play around with it! we needed image segmentation on food dishes. i'm wondering what amt of data something like that would need. on a side note, the blog mentions something about tools. you're telling me i can train my model to be more efficient with tools, etc.?

girish profil fotoğrafı
girish3 ay önce

@vatslgoswami ahh my intuition for something segmentation would be mid 100s but could be wrong. and yes you can train models to be much more efficient with tool calls/usage

Vatsl Goswami profil fotoğrafı
Vatsl Goswami3 ay önce

that is so doable!! maybe we will be on your multimodal customer stories sometime haha

Jason profil fotoğrafı
Jason3 ay önce

small oss models!! >>

girish profil fotoğrafı
girish3 ay önce

yes 100% for specialized tasks!

Khosla Ventures profil fotoğrafı
Khosla Ventures3 ay önce

Congrats on the launch! 🚀

girish profil fotoğrafı
girish3 ay önce

thanks guys!!

Sky profil fotoğrafı
Sky3 ay önce

yo what the heck, animation go crazyyyy

girish profil fotoğrafı
girish3 ay önce

yesss

Denise Teng profil fotoğrafı
Denise Teng3 ay önce

this is awesome😍

girish profil fotoğrafı
girish3 ay önce

thanks denise! goal's to make model training as simple as prompt engineeing :)

Joshua wong profil fotoğrafı
Joshua wong3 ay önce

congrats! which open weight models do you currently support?

girish profil fotoğrafı
girish3 ay önce

thanks joshua!! currently qwen, gemma's coming very shortly :)

sweekiat profil fotoğrafı
sweekiat3 ay önce

Cute launch video! 🐙 Any details on rough pricing?

girish profil fotoğrafı
girish3 ay önce

new users get $50 in free credits & pricing is token-based (and based on the size of the model)

Alchemist Quant profil fotoğrafı
Alchemist Quant3 ay önce

What would be your recommendations for using this w/ trading?

girish profil fotoğrafı
girish3 ay önce

ooh good question, not something we've looked into actually but i wonder if you can train a model to be good at forecasting -> e.g. given newspaper headlines, twitter discourse, economist speeches etc. predict if the fed will hike/cut rates

Alchemist Quant profil fotoğrafı
Alchemist Quant3 ay önce

And yeah I was thinking about analyzing it just the way you said. BTW, can we choose the model we want to train? Or we'd be training your specific model?

girish profil fotoğrafı
girish3 ay önce

we currently have support for qwen -> gemma's gonna be out shortly :)

Chirag profil fotoğrafı
Chirag3 ay önce

Very cool. I have something similar in the works. But for a different application.

girish profil fotoğrafı
girish3 ay önce

nice! what's the application?

Joel Lee profil fotoğrafı
Joel Lee3 ay önce

Congrats! What’s the next big feature the team has lined up?

girish profil fotoğrafı
girish3 ay önce

more use-cases! custom llms are very very much under-tapped and we want to show developers more of the cool/fun stuff you can do with custom models

Andy Berman profil fotoğrafı
Andy Berman3 ay önce

Can’t wait to try it

girish profil fotoğrafı
girish3 ay önce

:)

Dante profil fotoğrafı
Dante3 ay önce

I fear we all know who the star is 🐙

Nicholas Blanchard profil fotoğrafı
Nicholas Blanchard3 ay önce

I really want to build a version of Zaxy for open weight models. It would be so interesting to move beyond "amnesiac with a great notebook", or at least try it. Lots more research to do.

girish profil fotoğrafı
girish3 ay önce

zaxy looks cool! wonder if one could post-train a model to be effective at retrieving the right stuff from memory

Nicholas Blanchard profil fotoğrafı
Nicholas Blanchard3 ay önce

That is what I'm dying to investigate!

girish profil fotoğrafı
girish3 ay önce

give castform a shot then :) $50 in free credits for new users and i'm happy to get on a call to get you set up

Alex Klarfeld profil fotoğrafı
Alex Klarfeld3 ay önce

This is awesome - congrats Girish!

girish profil fotoğrafı
girish3 ay önce

thanks alex!!

pdawg profil fotoğrafı
pdawg3 ay önce

sick!

girish profil fotoğrafı
girish3 ay önce

thanks!! please give it a shot and let us know what you think :) new users get $50 in free credits

Khushboo Verma profil fotoğrafı
Khushboo Verma3 ay önce

cool video :)

girish profil fotoğrafı
girish3 ay önce

thanks khushboo!!

Sodan profil fotoğrafı
Sodan3 ay önce

This is actually cool, man. I'm actually training an Omni model from scratch. It is going to be a 1 billion parameter model, and it took me three months just to get to this level right now.

girish profil fotoğrafı
girish3 ay önce

ooh very cool! why from scratch?

Sodan profil fotoğrafı
Sodan3 ay önce

Because at the beginning I thought of using QWEN OMNI but it was a minimum of 3 billion parameters and 1. I wanted to make it as small as possible. that's why I made it like 1 billion 2. also I trained it on regional Indian languages which qwen does not have.

girish profil fotoğrafı
girish3 ay önce

ahh gotcha that makes sense!

Sodan profil fotoğrafı
Sodan3 ay önce

By the way I am really looking for any feedbacks from you I have dm'ed you

girish profil fotoğrafı
girish3 ay önce

great will respond!

Matt Henderson profil fotoğrafı
Matt Henderson3 ay önce

Woohoo congrats! Love the Pokemon reference 😁

girish profil fotoğrafı
girish3 ay önce

hahaha hope nintendo doesn't sue us 👀

Vandos ❓ profil fotoğrafı
Vandos ❓3 ay önce

The real unlock is agent traces as training data. Every session Fable 5 runs for you is a dataset. Castform turns your best AI outputs into a cheaper model that does the same thing.

girish profil fotoğrafı
girish3 ay önce

🚀🚀🚀

Daanish Khazi profil fotoğrafı
Daanish Khazi3 ay önce

Congrats @googrish !

girish profil fotoğrafı
girish3 ay önce

thank daanish! really enjoyed our convo on rubrics when u came over to the office!

Wen Qing 文青 profil fotoğrafı
Wen Qing 文青3 ay önce

👏🏻 Audio models when?

girish profil fotoğrafı
girish3 ay önce

soon! what would you post-train an audio model for?

Wen Qing 文青 profil fotoğrafı
Wen Qing 文青3 ay önce

Lotsa music use cases suffer from having low data volume (e.g. duet vocals, non-western music) so post training would help boost their performance using small but high quality datasets 🎵

girish profil fotoğrafı
girish3 ay önce

ahh very interesting! cc @YingHang7

sudarshan profil fotoğrafı
sudarshan3 ay önce

So excited!!!

girish profil fotoğrafı
girish3 ay önce

thanks sudz!

Yuma Tanaka profil fotoğrafı
Yuma Tanaka3 ay önce

compute is scarce and will stay scarce. smaller, specialized models is key to getting around that constraint. Good to see a platform built around that bet instead of fighting it

Benzer Videolar

Small Language Models (SML) are the future of AI. "Small" (SML) instead of "Large" (LLM). These small models are highly specialized models with superhuman abilities on specific tasks. Here are two techniques to build these models: • Spectrum • Model Merging I give you a short introduction in the attached video, but here is a quick summary: Spectrum helps us identify the most relevant layers to solve one specific task. We can ignore everything else and focus on fine-tuning these layers. Using Spectrum, we can fine-tune models in a heartbeat. Model Merging combines multiple models into a unique, much better model than any of the individual input models. You can also combine models specialized in different tasks and get a model with multiple abilities. This is the state of the art of productizing models. It's what Arcee.ai's platform does behind the scenes. Arcee collaborated with me on this post and is sponsoring it. There are three main steps to produce a model for your particular use case: 1. You create a dataset by uploading your data. 2. You train a model. At this step, Arcee uses Spectrum and Model Merging to produce a highly specialized model for your task. 3. You can deploy that model to any environment you want. Three important notes: • Training process is 2x faster and 2x cheaper than regular fine-tuning. • Resultant models are smaller and have higher accuracy. • They create these specialized models from open-source models. Check this site so you can fully appreciate how this works: If you want to fine-tune an open-source model, consider Arcee's platform. This is the state of the art.

Santiago

164,162 görüntüleme • 2 yıl önce

We’re launching Optima. Now anyone can create a custom benchmark for their use case, leveraging Artificial Analysis’ leading research and platform Building and running benchmarks is difficult. We have distilled Artificial Analysis’ research and experience developing benchmarks into Optima, a new platform for benchmarking models on your own workloads and comparing performance, speed and cost efficiency. Optima allows you to find the best model for your task, or an equally performant alternative to your current setup at 10x lower cost or time per task. We’ve integrated Artificial Analysis' research and experience in benchmarks across the Optima workflow: ➤ Build benchmarks based on your own data and use cases: There are three ways to build a benchmark with Optima. Upload an existing evaluation dataset from your own files or Hugging Face, or import agent traces from platforms including Arize AI, Braintrust and langfuse.com. Install the Optima skill to build a benchmark using context from your coding environment and previous sessions. Or simply describe your use case and provide example inputs and outputs, and Optima will build the benchmark for you ➤ Run across the latest models: Run the same benchmark across leading models in a single click, and keep your leaderboard up to date as soon as new models are released ➤ Bring Artificial Analysis grading to your own benchmark: Evaluate responses against objective rubric criteria or using the same pairwise judging approach used for Artificial Analysis benchmarks including GDPval-AA and AA-Briefcase. For pairwise judging, select your preferred responses from a sample and Optima uses those preferences to rank models across your test set ➤ Compare performance, cost and time efficiency: Optima measures more than model performance. Cost per Task and Time per Task are tracked alongside benchmark scores, with category-level results and support for custom metrics, allowing you to compare the tradeoffs between models for your specific use case Ahead of launch, here are examples questions our beta testers answered with Optima: ➤ Which model can save me 10x the cost without a meaningful decrease in quality for my finance & accounting agent? ➤ Which model best matches the writing style of lawyers for my legal agent? ➤ Which model can best identify different elements in my custom image dataset? Optima is available today. Build your own benchmark at

Artificial Analysis

133,068 görüntüleme • 1 ay önce