Lead story
Ask your AI
Top stories
Models & availability
Latest
Lead story
Ask your AI
Top stories
Models & availability
Latest
Runway built a capacity controller called deckard that reallocates GPUs between production inference and research based on daily demand cycles, using queueing theory to optimize allocations so that production tracks demand without over-provisioning and more GPUs are available for research overnight.
From the source
We built a lightweight capacity controller to do this reallocation (internally called deckard, after the Blade Runner protagonist who relentlessly reclaims replicants).
runway.com