Google DeepMind has released the Gemma 4 family of multimodal models on Hugging Face. The models are open-weight under the Apache 2 license and support text, image, and audio inputs with long context…
TII introduces Falcon Perception, a 0.6B-parameter early-fusion Transformer for open-vocabulary grounding and segmentation, and Falcon OCR, a 0.3B-parameter OCR model.
Hugging Face announces gradio.Server, a new server class that extends FastAPI to allow developers to build custom frontends with any framework (React, Svelte, vanilla HTML/JS) while leveraging…
IBM and Hugging Face announce Granite 4.0 3B Vision, a compact vision-language model for enterprise document understanding, featuring table extraction, chart understanding, and semantic KVP…
Hugging Face releases TRL v1.0, marking a shift from a research codebase to a stable library for post-training methods, now implementing over 75 methods with a stable/experimental split and semantic…
Together AI, Stanford University, the University of Wisconsin–Madison, and Bauplan collaborated to test whether LLMs can optimize database query execution plans.
Together AI announces the availability of the Wan 2.7 video model suite, beginning with text-to-video and soon including image-to-video, reference-to-video, and video edit.
Deepgram's speech-to-text and voice models (Nova-3, Nova-3 Multilingual, Flux, Aura-2) are now available natively on Together AI's Dedicated Model Inference platform, enabling teams to run the full…
The blog post tells the story of Together AI's kernels team, highlighting their work on FlashAttention, the ThunderKittens library, and their rapid optimization of kernels for NVIDIA's Blackwell…
is an open-source, RL-based framework for adaptive speculative decoding that learns from live inference traces and continuously updates the speculator without interrupting serving, achieving an…
OpenAI raises $122 billion in new funding to expand frontier AI globally, invest in next-generation compute, and meet growing demand for ChatGPT, Codex, and enterprise AI.
Google Research introduces a systematic evaluation framework that converts psychological questionnaires into situational judgment tests to measure how closely LLMs' behavioral dispositions align with…
Google Research introduces an evaluation framework for ML models that optimizes the trade-off between the number of items and raters per item to build reproducible AI benchmarks that capture human…
Google Research publishes a whitepaper with updated quantum resource estimates for breaking elliptic curve cryptography, showing a 20-fold reduction in physical qubits needed, and proposes a…
Google posted an overview of their AI updates from March 2026, but no specific details about models, products, or other announcements were provided in the source text.
Alibaba announces Qwen3.6-Plus, a new model with enhanced agentic coding capabilities, multimodal reasoning, and a 1M context window, available immediately via API.
Cursor 3 introduces a new interface with an Agents Window for running multiple agents in parallel across different environments, a Design Mode for annotating UI elements in the browser, Agent Tabs…