The article discusses optimization of LLM performance, specifically how long prompts block other requests, and describes a vLLM update that implements parallel prefills with a limit on long prompt…
Featherless AI is now a supported Inference Provider on the Hugging Face Hub, offering serverless inference with a large catalog of open-source models.
Hugging Face announced the Kernel Hub, a platform for loading pre-compiled, optimized compute kernels directly from the Hub, simplifying the process of using custom kernels for model acceleration.
NVIDIA announced the availability of Isaac GR00T N1.5, the first major update to the open foundation model for humanoid robots, and provided a tutorial on fine-tuning it for the LeRobot SO-101 arm.
Hugging Face and NVIDIA announced Training Cluster as a Service, a new offering that provides accessible GPU clusters for research organizations to train AI models.
OpenAI and Mattel are partnering to integrate AI into iconic brands such as Barbie and Hot Wheels, aiming to enhance creative development, streamline workflows, and create new ways for fans to engage.
OpenAI announces a new Outbound Coordinated Disclosure Policy for responsibly reporting vulnerabilities in third-party software, focusing on integrity, collaboration, and proactive security.
This is a guide on how to write effective prompts for Google's Veo 3 video generation model, covering visual elements, audio prompting, and character consistency.