Partiful logo
Get the app
Login

vLLM & llm-d Meetup: High performance inference, at scale

Thursday, Oct 8
10:00am – 2:00pm
Hosted by
◉‿◉
´◡`
^◡^
•‿•
❛o❛
Join AMD, Cerebras, MoonMath AI, Red Hat, and the vLLM project for a half day on open source inference, ending with a hands-on workshop you can follow along with on your own laptop. Self-hosting open models is now a real option for teams that want high performance without per-token costs. We will cover how vLLM and llm-d make distributed inference fast at scale, and how being smarter about which model serves which request cuts your inference bill without giving up quality.

Guest List

51 on the list
ᵔ▾ᵔ
´◡`
:)

Explore More

9ntNLUrNni4tg0aG42PLe
Chinatown Autumn Moon Festival
KsXwI4Sq3o02k8Pik2GwQ
Bean-Up
j4Z4oHK5uFjGtqshOT8QI
A Very Odysserious Run

Restricted Access

Must be on the list to view event activity & see list details
Already RSVP'd? Sign in
Partiful logo
Home
Explore
Create
Send a card
Log in
Partiful logo
Explore eventsCreate a free event
HelpBlogCareersAboutGet the app