Hugging Face announces Falcon-Edge, a series of 1.58-bit language models based on the BitNet architecture, available in 1B and 3B parameter sizes with base and instruction-tuned variants.
Hugging Face announces that the Transformers library will act as a central standard for model definitions across the ML ecosystem, aiming to increase interoperability among training frameworks,…
Hugging Face and Kaggle announce an integration that allows users to navigate between Hugging Face model pages and Kaggle notebooks, automatically generate Hugging Face model pages from Kaggle…
Hugging Face announced a new deployment option for OpenAI Whisper on Inference Endpoints, achieving up to 8x speed improvement and leveraging community open-source technologies.
This blog post surveys recent developments in Vision Language Models (VLMs) over the past year, covering new model trends such as any-to-any models, reasoning models, small yet capable models, and…
OpenAI describes Codex, a cloud-based coding agent powered by codex-1, a version of o3 optimized for software engineering through reinforcement learning on real-world coding tasks.
OpenAI introduces HealthBench, a new evaluation benchmark for AI in healthcare that evaluates models in realistic scenarios, built with input from over 250 physicians to provide a shared standard for…
Replicate announces that NVIDIA H100 GPUs are now available on their platform, along with multi-GPU configurations of A100 and L40S GPUs that were previously only available in deployments.
The blog announces that LoRAs can now run directly on the Hugging Face Hub using Replicate for inference, via a small update to Hugging Face's inference client that routes requests to Replicate's…