AMD x RedHat AI-Infra Workshop: High performance inference, at scale
About This Event
Join AMD, Cerebras, MoonMath AI, Red Hat, and the vLLM project for an in-person half day on open source inference, ending with a hands-on workshop you can follow along with on your own laptop.
Self-hosting open models is now a real option for teams that want high performance without per-token costs. We will cover how vLLM and llm-d make distributed inference fast at scale, and how being smarter about which model serves which request cuts your inference bill without giving up quality.
Hosted by
We email you a private link showing how many people find this event through Mimetic. Tick the box if you are open to sponsors for events like this one and we will get in touch when there is interest. Anything from there is yours to decide.
Want the other 53 SF AI events this week? One email, Monday morning.
Unsubscribe anytime.
Tell us who you are and what you sell. A person at Mimetic reads it and gets back to you by email. Nothing is charged here and we do not sell attendee data.
Get a free growth analysis for your company
See how your website, messaging, and go-to-market strategy stack up, in minutes.
Get My Free AnalysisMore SF events like this
AWS Startup Day SF 2026 — Go Agent & Go Global
The AI Conference 2026 - Shaping The Future of AI
Agents on a Leash, Products on Trial | SF | AI Show and Tell
Degrees of Freedom: Robotics Builders Night @ Mission Robotics