Hugging Face announces a custom text-generation pipeline class for the Llama 2 family of models (7b, 13b, 70b) on Intel Gaudi 2 AI Accelerator using Optimum Habana, providing easy-to-use scripts and…
LG AI Research proposes a new framework for 3D human pose and shape estimation from video, incorporating Spatial Alignment Module, Space2Batch, and uncertainty-guided attention re-weighting,…
Hugging Face announces the TTS Arena, a tool for benchmarking text-to-speech models via human voting, with an initial set of six models including both open-source and proprietary ones.
This blog post provides an overview of AI watermarking techniques, including visible and invisible watermarks, data poisoning, and signing methods, as well as tools available on the Hugging Face Hub.
Opening Up New Possibilities for AI, Lab Leader Soonyoung Lee of the Multimodal Lab
LG AI Research's Multimodal Lab, led by Soonyoung Lee, focuses on multimodal AI research including image generation, image editing, image understanding, and medical data analysis.