Amazon announces custom reward functions for multi-turn reinforcement learning with Amazon Nova Forge, including a serverless multi-turn RL option now generally available, and describes how to design…
Amazon shows how to combine OpenAI-compatible endpoints on SageMaker AI with Bedrock AgentCore runtime to build multi-agent workflows, using Qwen 3.5 9B, Claude Haiku 4.5, and Claude Sonnet 4.6.
Amazon Bedrock AgentCore Browser Tool is a fully managed, cloud-based browser service that enables AI agents to automate legacy web applications through secure, isolated sessions, integrating with…
Amazon Quick for Microsoft 365 extensions bring agentic AI capabilities directly into Word, Excel, PowerPoint, and Outlook, allowing users to access connected data and perform document editing…
Amazon Bedrock now supports granular cost attribution via CUR 2.0 with IAM principal data, enabling per-user and per-application cost tracking through Amazon Athena queries and CUDOS dashboards.
OneAdvanced deployed over 50 AI agents on UK-sovereign AWS by self-hosting Llama 4 Maverick and Llama Guard 4 on Amazon SageMaker AI, using vLLM on p5.48xlarge instances in the London region.
Amazon Bedrock AgentCore payments, a capability for AI agents to make governed payments, was introduced in May 2025 in partnership with Coinbase and Stripe.
Amazon SageMaker HyperPod introduces a tiered KV cache architecture using Curvine, a distributed cache filesystem, to extend KV cache across GPU, CPU, and shared NVMe, enabling cross-replica cache…
OpenAI's Daybreak Red (GPT-5.6 Cyber) and Daybreak Blue (GPT-5.6 Sol) cybersecurity models are now available on Amazon Bedrock to eligible customers, with security features like zero-operator access…
ONESTRUCTION, with advisory from AWS GenAIIC, built Ishigaki-IDS, a foundation model specialized for construction industry BIM workflows, using a three-stage training pipeline (CPT, SFT, RLVR) on top…
First Orion uses Amazon Nova Act to accelerate QA automation by allowing QA analysts to describe tests in plain English instead of writing and maintaining code.
Amazon announces a production reference deployment of the Claude apps gateway for AWS, providing a self-hosted governance layer for Claude Code and Claude Desktop, with details on architecture,…
Amazon announces the SageMaker AI Spaces add-on for Amazon EKS, which allows data scientists to run managed JupyterLab and Code Editor environments directly on their existing EKS cluster, reducing…
nOps transitioned to Amazon Bedrock AgentCore to accelerate product delivery, improve response quality, and reduce operational complexity in building their FinOps AI agent, Clara.
Announcement of a guide for developers on using GPT-5.6 to build faster, more cost-efficient AI agents with smarter model selection and new Responses API capabilities.
OpenAI previews Ultrafast, a new API service tier for GPT-5.6 Sol that offers up to 14× faster inference and up to 750 output tokens per second, powered by Cerebras hardware.
OpenAI publishes two studies on enterprise AI adoption, showing a shift from assistance to execution, with frontier firms generating 8.3× more output tokens per active user than typical firms, and…
RingCentral uses ChatGPT Work and Codex to enable AI-native development across engineering and operations, including an AI-Native Challenge and PMO workflow automation, accelerating product feature…
OpenAI announces that Daybreak models (Daybreak Blue and Daybreak Red) are now available on AWS through Amazon Bedrock, building on the earlier general availability of OpenAI frontier models and…
OpenAI's CFO shares five lessons from building an AI-native finance function, including giving broad AI access, redesigning workflows around decisions, and working toward a zero-day close and…
Model ML uses GPT-5.6 Sol in its agents to automate finance workflows, creating editable PowerPoint and Excel files with fewer tokens and higher professional readiness compared to other models like…
OpenAI expands Daybreak with two access tiers (Daybreak Blue and Daybreak Red) and introduces GPT-5.6-Cyber, a cybersecurity-specific model built on GPT-5.6 Sol, with improved performance on…
OpenAI is expanding the Daybreak Cyber Partner Program to bring its frontier cyber models (Daybreak Blue and Daybreak Red) to security partners, allowing them to integrate these models into their…
Zapier's enterprise marketing team uses ChatGPT Work to automate lead funnel optimization, campaign asset building, and reporting, reducing manual effort and enabling focus on strategic work.
OpenAI is introducing Premium seats for ChatGPT Business, offering 5x more usage than Standard seats, no five-hour usage limit, and predictable weekly usage resets.
In the Hugging Face Hub, the top 25 models by downloads and the top 25 by likes share only one repository, indicating that attention (likes) and adoption (downloads) measure different things.
The article describes a new streaming data loop feature in Strands Robots that allows recording robot demonstrations, training on them by streaming directly from Hugging Face Hub, and deploying the…
Hugging Face reports on a community hackathon that reproduced 2,226 papers from ICML 2026 using coding agents, finding 51% of examined papers had at least one verified claim and 23% had at least one…
OlmoEarth Studio now supports custom embedding exports, allowing users to compute and download embedding vectors from OlmoEarth foundation models for downstream tasks like similarity search and…
LFM2.5-VL-3B is our most capable vision-language model you can run on your own hardware. It understands documents and screens alike, grounds objects, and can call tools.
IBM Research introduces ALTK-Evolve, an agentic memory system that learns from agent trajectories and delivers guidelines selectively, achieving better accuracy and lower token cost compared to ACE…
NVIDIA released Magpie TTS Multilingual, an open-weights 364M-parameter text-to-speech model supporting 12 languages, including three new languages (Modern Standard Arabic, Korean, Brazilian…
MultiverseComputingCAI researchers publish a paper on efficient knowledge distillation, introducing offline top-K logit caching and a fused chunked KL loss to reduce VRAM usage, enabling long-context…
Meta released Muse Glimmer, a 30B parameter multimodal model distilled from Muse, licensed under Apache 2.0, designed for local agentic use cases like coding, document analysis, and personal…
xAI's Grok 4.6 reasoning model is rolling out in GitHub Copilot, designed for agentic coding and complex multi-step workflows, with strong results in terminal-based coding tasks.
GitHub Copilot weekly updates include new models (Kimi K3 and MAI-Code-1.1-Flash), general availability of Agent Plugins 1.0, improved plugin management, side chat for agent questions, subagent…
GitHub announced Agent Plugins 1.0, an open standard that packages agent skills and MCP servers into one installable plugin, with support across VS Code, Copilot CLI, GitHub Copilot SDK, and the…
GitHub Copilot for JetBrains is updated with persistent memory across chat sessions, the ability to use Ollama as a BYOK provider, expanded Codex workflows, easier Copilot CLI setup, and additional…
GitHub announces the deprecation of MAI-Code-1-Flash across all GitHub Copilot experiences on September 10, 2026, with MAI-Code-1.1-Flash as the suggested alternative.
Microsoft's MAI-Code-1.1-Flash, an updated small coding model with vision support and improved performance, is rolling out in GitHub Copilot at a reduced price.
Indonesia launched its first university-based AI center, UGM Indosat NVIDIA AI Technology Center, powered by NVIDIA's full-stack AI platform and Indosat's GPU Merdeka, to develop local AI talent and…
NVIDIA announced partnerships with Apollo, BlackRock, Blackstone, Brookfield, Goldman Sachs and KKR to establish independent financing platforms designed to mobilize over $500 billion of third-party…
NVIDIA announces the 800 VDC power architecture for AI factories, developed with Google and Microsoft through OCP, enabling more efficient power delivery for next-generation AI compute.
NVIDIA highlights multiple open-source models and tools from the community, including Cosmos 3 Edge, MiniMax-H3, Laguna S 2.1, DeepSeek-V4-Flash, Inkling-Small, Wan-Animate-2, LTX-2.5, Muse Glimmer,…
NVIDIA announced Nemotron 3.5 Lightning, a 30-billion-parameter mixture-of-experts open model for high-volume agentic AI tasks, and NeMo Switchyard, an open-source library for intelligent model…
Google DeepMind announces Gemini 3.7 Flash, an improved workhorse model for coding and agents, with better performance on benchmarks like FrontierCode 1.1 Main (43.6% vs 34.4%) and DeepSWE v1.1…
Google DeepMind announced Gemini Omni, a multimodal model, with its first release Gemini Omni Flash focused on video generation and editing via conversational interaction.
Google DeepMind announced SL2T, a massively multilingual sign-language-to-text translation model, and its integration into Gboard and Live Transcribe on Pixel 11 for ASL-to-English dictation, with…
Google DeepMind and Google Research announce a study in which AMIE, a research medical AI system built on Gemini and Project Astra, demonstrates real-time clinical video consultation capabilities,…
xAI releases Grok 4.6, an update to Grok 4.5 focused on long-running agents, interactive and visual work, achieving frontier intelligence on several benchmarks, available in Cursor, Grok Build, API,…
Grok Bot is a team of always-on AI agents with their own computer, launched in early beta. Available for SuperGrok Plus and Heavy, Cursor Pro+ and Ultra, and Cursor Teams Standard and Premium…
Google Antigravity introduces Gemini 3.7 Flash, a new model for coding and agents with improved performance and introductory pricing at half the cost of 3.6 Flash.
Google Antigravity announces Custom Agents, a new feature for Antigravity 2.0 and the CLI that allows users to create specialized, file-based agent configurations with scoped instructions, tools, and…
Google Research announces AMIE (Video), a real-time video configuration of its research medical AI system AMIE, built on Gemini and Project Astra, which conducts synchronous clinical video…
MindTopo is a new benchmark for testing topological reasoning in AI, evaluating whether multimodal models can understand concepts such as connectivity, enclosure, order, separation, and knots.
CARE-X is a research model from Microsoft Research that combines generative and discriminative capabilities for chest X-ray interpretation, using reinforcement learning (DAPO) and tool-augmented…
Tencent's CarbonX program supports CERT Systems, which uses Direct CO2 Electrolysis to convert captured carbon dioxide, water, and clean electricity into ethylene, aiming to replace fossil-fuel-based…
Anthropic announces that future Claude models will include a text watermark to comply with the EU AI Act, using a method based on SynthID-Text that does not affect output quality, readability, or…
Cursor introduces Builds for Cloud Agents, providing pre-built environments that agents boot into instead of setting up from scratch each session, resulting in 3x faster time to first token and…
LG AI Research introduces Segment-based Topic Allocation (SBTA), a method that assigns text segments to topics rather than entire documents, reducing topic contamination.
WhatsApp is introducing Scam Alert, an optional on-device ML feature that detects scam messages without sending any message content to servers, using a transparent model that can be independently…
Microsoft announces a four-part blog series on optimizing AI agent costs using Microsoft Foundry, focusing on financial discipline and managed investment.
MiniMax introduces Music 3.0, a next-generation music generation model that composes, arranges, performs, and produces a complete song from a creative concept and optional lyrics in a single…
Mistral AI announces three steps for sovereign AI: general availability of Mistral Regional Endpoints for in-region inference, public preview of Mistral Priority Tier with committed service levels…