Schedule6 sessions · 10 speakers
Pick a day to see what's on, then tap a session for the full description.
11:15 AM
InfrastructureTalk
Serving four hundred models on one GPU fleet
Multi-tenant inference from the operator's seat: scheduling, memory packing, cold-start mitigation, noisy-neighbour isolation, and the observability you need before you can safely oversubscribe anything.
11:15 AM - 12:00 PMMain Stage
Speaker
- LFLiam Ferguson
Staff Site Reliability Engineer · Orbit Cloud
Format: TalkTrack: Infrastructure
Tracks: AI Engineering · Product · Infrastructure