The post is a technical deep dive into the Codex agent loop, explaining how Codex CLI orchestrates models, tools, prompts, and performance using the Responses API.
OpenAI and the Gates Foundation announced Horizon 1000, a $50 million pilot initiative to advance AI capabilities for healthcare in Africa, aiming to reach 1,000 clinics by 2028.
OpenAI announced Stargate Community plans, a community-first approach to AI infrastructure with locally tailored plans shaped by community input, energy needs, and workforce priorities.
ServiceNow is expanding access to OpenAI's frontier models to enable AI-driven enterprise workflows, summarization, search, and voice capabilities on its platform.
ChatGPT is rolling out age prediction to estimate whether accounts belong to users under or over 18, applying safeguards for teens and refining accuracy over time.
IBM Research introduces AssetOpsBench, a benchmark and evaluation system for AI agents in industrial asset lifecycle management, featuring 2.3M sensor telemetry points, over 140 scenarios across 4…
The article reflects on the one-year anniversary of DeepSeek's R1 model release, analyzing how it catalyzed the growth of China's open source AI ecosystem by lowering technical, adoption, and…
Microsoft introduces DIFF V2, a differential attention mechanism that doubles query heads without increasing key-value heads, enabling faster decoding and eliminating the need for custom attention…
Alibaba's Qwen team announces Qwen3-Max-Thinking, a new flagship reasoning model with adaptive tool-use capabilities and a test-time scaling strategy, now available in Qwen Chat and via API.
Alibaba's Qwen team has open-sourced the Qwen3-TTS family of speech generation models, including two sizes (1.7B and 0.6B parameters), supporting voice clone, voice design, and natural language-based…
Google Research introduces GIST, a novel algorithm for data subset selection that balances diversity and utility with provable guarantees, presented at NeurIPS 2025.
Sakana AI published an unofficial guide on what they look for when interviewing research candidates, emphasizing understanding over implementation, asking questions that distill the problem space,…
Moonshot AI is open-sourcing the Kimi Vendor Verifier (KVV), a project that provides benchmarks and tools to verify the accuracy of inference implementations for open-source models, addressing issues…
Together AI shares lessons learned from large-scale deployments on optimizing inference speed and costs, covering quantization, distillation, regional inference proxies, reducing compute stalls,…
Z.ai launched GLM-4.7-Flash, a lightweight and efficient free-tier version of GLM-4.7, with strong performance on coding, reasoning, and generative tasks, low latency, and high throughput.