OpenAI announces 4o image generation, a new image generation approach that is more capable than DALL·E 3, can create photorealistic output, and can take images as inputs and transform them.
Hugging Face announces native integration of Intel Gaudi hardware support into Text Generation Inference (TGI), enabling deployment of LLMs on Gaudi accelerators with features like multi-card…
An updated version of DeepSeek-V3 (DeepSeek-V3-0324) has been released, featuring improvements in instruction following, code and math capabilities, and now using an MIT license.
Hugging Face published a blog post detailing how to use the Sentence Transformers library to finetune reranker (cross-encoder) models, including dataset loading, loss functions, training arguments,…
Hugging FaceStory of the weekCapability change· 24 Mar
Hugging Face announced a major update to Gradio's Dataframe component, adding ten new features such as multi-cell selection, row numbers, column pinning, copy and full-screen buttons, scroll-to-top,…
Alibaba's Qwen team officially released the first version of QVQ-Max, a visual reasoning model that can understand and reason with content from images and videos.
Alibaba releases Qwen2.5-Omni, a new end-to-end multimodal model that processes text, images, audio, and video and generates real-time streaming responses via text and speech.
This blog post by Replicate highlights recent AI model releases and updates available on the platform, including ShieldGemma 2 for NSFW detection, Hunyuan3D 2Mini for 3D generation, CSM-1B and…