Amazon announced the Agentic Data Operations Platform (ADOP), a reference architecture on AWS that uses AI agents to accelerate data engineering from weeks to hours.
Amazon announces Amazon Bedrock AgentCore Gateway, a managed service capability for governing AI agent tool access with authentication, authorization, and policy enforcement.
Amazon Bedrock introduces a query-aware compression pattern for RAG: a smaller, lower-cost model filters retrieved chunks against the user's query before the primary model generates the answer,…
The post describes how Panasonic Avionics Corporation built an agentic AI system on AWS to accelerate IFEC diagnostics, using Amazon Bedrock, Amazon SageMaker, and AWS Glue.
Amazon Bedrock now offers OpenAI GPT-5.6 models with cross-Region inference, supporting three variants (Sol, Terra, Luna) via geographic and global inference profiles to improve throughput and…
Amazon announces a no-code ML workflow integrating Snowflake, Amazon SageMaker Canvas, and Amazon Quick, with a focus on data preparation and model building using SageMaker Canvas and the XGBoost…
Amazon announces the integration of Amazon SageMaker Canvas predictions with Amazon Quick Sight (now part of Amazon Quick) to build interactive dashboards for fraud detection, including generative BI…
Amazon announced an expansion of the Policy Authoring capability in Amazon Bedrock AgentCore, which now allows users to convert natural language policy documents into Dogwood formal specifications…
This post discusses architectural principles for scaling agentic AI systems in multi-framework, multi-model, multi-provider enterprise environments, emphasizing the separation of control and…
Amazon announces a multi-agent framework on Amazon Bedrock AgentCore that automates cloud migration tasks including discovery, IaC generation, governance, and operations, reducing IaC development…
AWS announces vector solutions including search and retrieval capabilities that can be added to existing AWS data stores (Amazon OpenSearch Service, S3, Aurora PostgreSQL, DynamoDB, ElastiCache for…
Amazon Bedrock is used to build intelligent security for healthcare APIs, providing context-aware security monitoring, anomaly detection, data sensitivity classification, and compliance report…
Amazon announces runtime domain and published-date filtering for Web Search on Amazon Bedrock AgentCore, shipping in web-search connector version 1.2.0.
Amazon announces two AWS solutions, the GAIIC IDP Accelerator and Quick Automate, that together automate document processing for mortgage lending and other industries.
Fanatics Betting and Gaming built a multi-agent customer support system on AWS using Amazon Bedrock. The system includes a Responsible Gaming classification agent powered by Amazon Nova 2 Lite and a…
Amazon introduced KnowledgeForge, a system that mines resolved ITSM incident tickets to generate and curate knowledge base articles using generative AI and RAG, running on AWS infrastructure…
Amazon announced the general availability of AgentCore payments, a service that enables agents to autonomously pay for APIs, MCPs, and content, with integrations with Coinbase and Stripe wallets.
This post describes how to customize Amazon QuickSight embedded chat with visual theming (container CSS, SDK frame options, brand removal) and chat persona/tone configuration to match an…
Amazon Bedrock is used in a multi-agent document classification solution that combines Anthropic's Claude Haiku 4.5 for textual reasoning and Amazon Titan Multimodal Embeddings G1 for visual pattern…
Amazon introduced AI-Driven Annotation (AIDA) solution on Amazon Bedrock to improve contract search accuracy using implicit and explicit filtering with metadata-enriched chunking.
NVIDIA Nemotron 3.5 Lightning is now available on Amazon SageMaker JumpStart. It is an open model with a hybrid MoE architecture (30B total, 3B active), offering up to 4x higher throughput and 30%…
Amazon Bedrock AgentCore payments integrates with OpenClaw agents to enable programmatic, bounded payments for autonomous agents, using wallet providers like Coinbase or Stripe Privy and protocols…
OpenAI announced a new blog called 'Intelligence Age' that will explore how transformative AI could reshape power, governance, the economy, and individual freedom.
OpenAI announced a new blog called AI Futures that will explore how transformative AI could reshape power, governance, the economy, and individual freedom.
OpenAI reaffirms zero data retention policy for eligible API customers and previews a new private safety processing feature for advanced AI safety that maintains data privacy.
OpenAI launches an initiative to strengthen democratic oversight of AI in national security, providing tools, training, and expertise to government institutions.
OpenAIStory of the weekProduct launch· 12 days ago
OpenAI discusses the cybersecurity threat from AI models and its own defensive measures, including the use of GPT-5.6 Sol and ChatGPT Work for security assessments and fixes.
OpenAI announces its participation in the PORTS-Pike project in Pike County, Ohio, including a partnership with SB Energy, NVIDIA, and the U.S. Department of Energy to develop a data center expected…
OpenAI is awarding $1 million in grants plus up to $1 million in API credits to 14 independent projects focused on economic opportunity and societal resilience related to AI advances, following a…
Hugging Face describes how they use Inference Endpoints, Jobs, and Storage Buckets to build a hybrid search system for Papers with Code, using the Qwen/Qwen3-Embedding-0.6B model for embeddings.
Hugging Face introduced held-out sets in three speech recognition leaderboards to better measure real-world performance and address benchmark optimization.
Liquid AI releases DSpark draft model checkpoints for three LFM2.5 family models, providing speculative decoding for faster inference (up to 3.18x throughput improvement on GPU, up to 2.87x…
ALTK-Evolve is a method for agentic memory that distills guidelines from an agent's past trajectories and injects them at inference time without weight updates.
Sentence Transformers library version 6.0 introduces a new model type called MultiVectorEncoder, which implements ColBERT-style late interaction retrieval.
Hugging Face blog post describes a constraint-aware GPU allocator that improves GPU utilization by up to 33 percentage points and priority-weighted output by up to 105% compared to FIFO scheduling.
Together AI compares GLM-5.3 and Claude Fable 5 on DeepSWE, finding near-identical pass@1 accuracy but GLM-5.3 dominating cost and multi-attempt metrics.
Together AI compares GLM-5.3 and GPT-5.6 Sol on the DeepSWE benchmark, showing that a cascade strategy (use GLM-5.3 first, escalate to Sol on failure) solves 85.9% of tasks at a lower cost than using…
The post compares DeepSeek V4 Pro 0813 and GPT-5.6 Sol on the DeepSWE benchmark. GPT-5.6 Sol has higher single-shot accuracy (72.7% pass@1 vs 62.8%) but is 35x more expensive ($8.37 per rollout vs…
Together AI introduces native A/B testing for LLM endpoints, allowing traffic splits between a control and up to 20 variants with fixed percentages, etag-guarded updates, and blue-green promotion.
The blog post compares DeepSeek V4 Pro 0813 and Claude Fable 5 on the DeepSWE benchmark, finding that DeepSeek V4 Pro 0813 is 90x cheaper and achieves comparable or better performance with multiple…
xAI announces that Grok 4.6 is now available on Google Enterprise Agent Platform via Model Garden, with a 500k context window, configurable reasoning efforts, and per-token pricing for input, cached…
Grok 4.6 is now generally available on Amazon Bedrock, with a 500k context window and configurable reasoning efforts, at the stated pricing per million tokens.
Tencent describes how Weixin Mini Games have become a social gaming platform within the Weixin app, with 500 million monthly players, a broad demographic, and a developer ecosystem of small studios.
The source text is Tencent's Carbon Neutrality Mid-Term Report, which discusses progress toward carbon neutrality and the role of AI in supporting lower-carbon growth.
GitHub announces a public preview of the new GitHub Copilot experience in Slack, which brings agentic capabilities of GitHub Copilot CLI and the GitHub Copilot app into Slack.
GitHub announces a public preview of shared agentic work with GitHub Copilot in Microsoft Teams, allowing users to start collaborative agent sessions, create code channels, and manage tasks…
GitHub Copilot for JetBrains now supports enterprise managed settings, enabling administrators to enforce consistent controls for plugin governance, MCP server access, OpenTelemetry, and permission…
Antigravity announces a remote control feature that lets users connect to and drive their Antigravity sessions from any machine with a web browser, enabling multi-instance management, untethered…
Google Antigravity announces IDE extensions for Visual Studio Code, Visual Studio, JetBrains, Zed, and Xcode, enabling developers to access Antigravity's agentic features within their preferred IDEs.
Google Antigravity is now bundled with Gemini Enterprise subscriptions, providing developers with autonomous agentic capabilities including IDE extensions, sandbox controls, budget caps, and audit…
Google Research introduces the Biomarker Discovery Framework, a multi-agent system that prioritizes candidate biomarkers from wearable sensor data through iterative hypothesis generation, statistical…
Google Research introduces a dynamic, mobility-informed framework called Mobility-Embedded POIs (ME-POIs) that enriches language model place representations by incorporating aggregated and anonymized…
Google Research introduces PhotoScan, a deep learning framework that estimates body composition metrics (body fat percentage, A/G ratio, V/S ratio) from standard 2D smartphone photos, achieving…
Cursor released improvements to cloud agents, including subscriptions to events (PRs, Slack, scheduled tasks), custom modes, subagents on their own VMs, a /goal command for long-lived objectives, and…
Cursor announces Origin, a new code hosting feature in early beta for paid plan users, allowing them to create and host repos, sync GitHub repos, manage pull requests, and integrate with apps like…
Google DeepMind published a blog post explaining the concept of full-stack AI, featuring an interview with engineering lead Paige Bailey who breaks down the layers of infrastructure, security,…
Agent API presets now automatically use stable prompt cache keys, allowing reuse of shared prompt prefix across requests, which can reduce costs by about 5%.
Black Forest Labs launched FLUX Upscale, a standalone tool and API endpoint that upscales video up to native 4K, fixing imperfections like smudged faces and gridded artifacts.
Cohere's analysis of LLM training data reveals a 'culture funnel' where cultural diversity narrows from pretraining to post-training, as post-training datasets focus more on technical tasks and lose…
Microsoft was named a Leader in the 2026 Gartner Magic Quadrant for Cloud-Native Application Platforms, marking its third consecutive year in this position.
Microsoft Research released Skala-1.1, a deep-learning DFT functional trained on 2.5× more data than the first public version, achieving a weighted average error of 2.8 kcal/mol on GMTKN55.
Mistral AI announces Agentic Search, a retrieval layer that uses a multi-step search loop to improve accuracy and efficiency when navigating complex documents, reducing latency and token consumption…
NVIDIA announces a partnership with SB Energy to secure LPS (land, power, shell) capacity at the PORTS-Pike Technology Campus in Ohio, where NVIDIA compute will be hosted and OpenAI will be the…
Z.ai releases GLM-5.3, an update to its model series, with improved coding capabilities (50% gain over GLM-5.2 on Z.ai Code Bench) and emergent cybersecurity capabilities matching Mythos 5 in…