SegMoE is a framework for creating Mixture-of-Experts Diffusion models. The blog announces the release of three SegMoE models (SegMoE-4x2, SegMoE-2x1, SegMoE-SD4x2) and the segmoe package for…
Hugging Face introduces the NPHardEval leaderboard, a dynamic benchmark for evaluating LLM reasoning abilities using complexity classes, with 900 algorithmic questions updated monthly.
Blog post providing a tutorial on getting started with the PatchTST model for time series forecasting, including installation, training on the Electricity dataset, and zero-shot transfer learning on…
Hugging Face announces general availability of Text Generation Inference (TGI) on AWS Inferentia2 and Amazon SageMaker, enabling deployment of open LLMs like Zephyr 7B on AWS Inferentia2.
Hugging Face presents an end-to-end recipe for Constitutional AI using open models, releasing a new tool called llm-swarm for scalable synthetic data generation on GPU Slurm clusters, along with…
Hugging Face and Intel present optimizations for StarCoder-15B on 4th gen Xeon, achieving over 7x inference acceleration through 8-bit and 4-bit quantization combined with assisted generation…
Hugging FaceStory of the weekResearch paper· 29 Jan
Hugging Face announces the Hallucinations Leaderboard, a comprehensive platform evaluating LLMs on hallucination-related benchmarks using in-context learning, with tasks including closed-book QA,…
The blog post discusses multimodal generation, highlighting LG AI Research's paper on music loop generation and mentioning Google DeepMind's Lyria and Meta's Audiobox.
LG AI Research summarizes key AI technology trends observed at CES 2024, focusing on on-device AI, Korean companies' activities, and the integration of generative AI into existing services.
Replicate announces that Code Llama 70B is now available via its API, with code examples for JavaScript, Python, and cURL, and mentions three variants: Base, Python, and Instruct.
Sakana AI announces it is one of seven institutions in Japan selected by the Japanese government to receive a supercomputing grant, organized by NEDO under METI, to support development of foundation…