Video yükleniyor...

Video Yüklenemedi

Ana Sayfaya Dön

My cofounder and I are making an FPGA-accelerated server for high-memory high-bandwidth workloads. We're looking for one or two companies to partner with; we'll do the work to port your application to our hardware. Please DM me if you've got a tricky workload!

79,906 görüntüleme • 11 ay önce •via X (Twitter)

36 Yorum

Peter Schmidt-Nielsen profil fotoğrafı
Peter Schmidt-Nielsen11 ay önce

You can DM me, email me at [email protected], or schedule a call here: I'd love to chat about what you're working on, or if you know anybody else I should reach out to, please let me know!

FFmpeg profil fotoğrafı
FFmpeg11 ay önce

Video encoding is a tricky workload. There are only a handful of FPGA encoders left.

Stewart Alsop - Host of Crazy Wisdom Radio Show profil fotoğrafı
Stewart Alsop - Host of Crazy Wisdom Radio Show11 ay önce

Id like to interview you about the importance of FPGAs for our future, what do you say?

Peter Schmidt-Nielsen profil fotoğrafı
Peter Schmidt-Nielsen11 ay önce

I'll DM you!

Mohand profil fotoğrafı
Mohand11 ay önce

I’ve always believed FPGAs could be a strong fit for the future of AI computing, I even almost started a thesis on it. Why do you think they haven’t seen broader adoption?

𝐷𝑟. 𝐼𝑎𝑛 𝐶𝑢𝑡𝑟𝑒𝑠𝑠 profil fotoğrafı
𝐷𝑟. 𝐼𝑎𝑛 𝐶𝑢𝑡𝑟𝑒𝑠𝑠11 ay önce

Going to be at SC25?

Peter Schmidt-Nielsen profil fotoğrafı
Peter Schmidt-Nielsen11 ay önce

I'm considering it! Will you be there?

𝐷𝑟. 𝐼𝑎𝑛 𝐶𝑢𝑡𝑟𝑒𝑠𝑠 profil fotoğrafı
𝐷𝑟. 𝐼𝑎𝑛 𝐶𝑢𝑡𝑟𝑒𝑠𝑠11 ay önce

Always. One of the hottest events of the year.

sasuke⚡420 profil fotoğrafı
sasuke⚡42011 ay önce

actually, i gave up right away on moving the code to the gpu, because the memory bandwidth of gpus is not that different from cpus *if you miss to main memory all the time*

🕷️ profil fotoğrafı
🕷️11 ay önce

Not a company but Leskovec is the Graph ML guy at Stanford and I think he regularly works with 10TB datasets and non-standard algos.

Juan profil fotoğrafı
Juan11 ay önce

do you have some docs or a blogpost on the workload porting part of the operation? i'm super curious about it

Peter Schmidt-Nielsen profil fotoğrafı
Peter Schmidt-Nielsen11 ay önce

Not yet, but we probably will at some point in the upcoming months! I think it would be fun to talk more about that.

Juan profil fotoğrafı
Juan11 ay önce

for sure. thanks

JPEL profil fotoğrafı
JPEL11 ay önce

Hey, that looks really interesting! (Tiny unsolicited advice: you might want to either unpin or release a playable version, because potential partners might worry you're the kind of founder who never ships 😅 not that I’d know anything about that!)

ren profil fotoğrafı
ren11 ay önce

Why not include HBM through the relevant Xilinx or Altera parts?

Peter Schmidt-Nielsen profil fotoğrafı
Peter Schmidt-Nielsen11 ay önce

Yes, very reasonable question. It's a complicated calculation whether HBM pencils out as better for any given workload. The design is extremely modular, and I think we are likely to add some HBM to some of the modules at some point, for workloads that benefit from it.

chandra profil fotoğrafı
chandra11 ay önce

I love this idea.

Meet profil fotoğrafı
Meet11 ay önce

@jaysidd fpga servers

s profil fotoğrafı
s11 ay önce

What kind of processor does it use? Is the bandwidth before or after hardware compression? Is rand read as fast and what is the optimal block size?

Peter Schmidt-Nielsen profil fotoğrafı
Peter Schmidt-Nielsen11 ay önce

The host Linux system is an AMD Epyc platform. The FPGAs are from Xilinx. These are raw speeds of the various interfaces, not compressed speeds. The flash speed is that we have ~160 separate NVMe interfaces, so re: rand reads, picture the performance characteristics of ~160 SSDs.

oFFMeta profil fotoğrafı
oFFMeta11 ay önce

so somewhere in between positron atlas and titan. 2u 2-3kw aircooled, ~$100k per server? they have a simple huggingface/openai endpoint mapping setup. between stuff like emfasys & fpgas it will be cool to see all of upcoming the inference workload hardware optimizations.

Rohan makes ASICs 🛠️ profil fotoğrafı
Rohan makes ASICs 🛠️11 ay önce

Ignoring FPGA is no longer an option Mark my words - Kintex Ultrascale+ will become the new ICE40 Ultra Plus.

sdmat profil fotoğrafı
sdmat11 ay önce

Hope you find success with a niche for this, but "a little bit lighter on FLOPs" is one of the statements of all time.

Luigi Cruz profil fotoğrafı
Luigi Cruz11 ay önce

Impressive! We might have a use for it depending on how fast we can get data in and out of the box. Any info regarding that? My DMs are open!

HDP profil fotoğrafı
HDP11 ay önce

Hell yes!

rvi profil fotoğrafı
rvi11 ay önce

Check DM my dude

Muthu / முத்து profil fotoğrafı
Muthu / முத்து11 ay önce

What's your Moat bro ? Programming FPGA sucks still whereas ASIC reprogram in few micro s

Chuck Petras profil fotoğrafı
Chuck Petras11 ay önce

@BrianRoemmele

Sendao Haz profil fotoğrafı
Sendao Haz11 ay önce

please be mindful if you process graphics at high speeds :) reality is less hypnotic than a chariot made of LED cascades

Michael F. profil fotoğrafı
Michael F.11 ay önce

Could be fun for radar signal processing. Best of luck and keep us updated!

Jerry Giant profil fotoğrafı
Jerry Giant11 ay önce

Some ideas: 1, RFSoC Polyphase Filterbank, FFT, signal categorization,MUSIC, for advanced passive DF for Drone detection, all HDL 2, find a Apache TVM example for quick demo

Sunoumanu profil fotoğrafı
Sunoumanu11 ay önce

Isn't FPGA have limited frequency in comparison to other types of general purpose hardware?

Happy Pirate | ソラナ | Triton.One profil fotoğrafı
Happy Pirate | ソラナ | Triton.One11 ay önce

@alessandrod wen can also get this instead of good software.

Donkey Kong profil fotoğrafı
Donkey Kong11 ay önce

A single Chrome tab will still eat up all that RAM 😂

miles profil fotoğrafı
miles11 ay önce

You should check out @solana.

Nikolai Aleichem profil fotoğrafı
Nikolai Aleichem11 ay önce

would you port cuda/cudnn/cublas projects?

Benzer Videolar