August 05, 2026 09:27 AM
Anthropic reportedly agreed to buy six years of cloud capacity from AI infrastructure startup Volta. The planned 133-megawatt Norway data center would be developed with Bitdeer and powered by NVIDIA Vera Rubin systems.
Read MoreAugust 05, 2026 09:27 AM
Model routing on the Google Cloud API Gateway is now available in public preview. The API gateway provides a lightweight, serverless ingress layer that accepts OpenAI-compatible requests and dynamically routes them to Gemini, Claude, or OpenAI OSS-GPT. It can be used standalone for simple rate limiting and token tracking or paired seamlessly with the Gemini Enterprise Agent Platform. This article contains a step-by-step guide on how to configure router logic.
Read MoreAugust 05, 2026 09:27 AM
Cloudflare Wallets is a system designed to give AI agents stable identities and controlled access to payments for APIs, MCP tools, and online content. Virtual Wallets will support spending limits, allow lists, and transaction caps for safer agentic commerce.
Read MoreAugust 05, 2026 09:27 AM
ChatGPT Work is OpenAI's agent product for knowledge work. Its current form is an amalgamation of ChatGPT, the Codex app, the Codex harness, the original Codex cloud agent, ChatGPT agent, Atlas, OpenClaw, and more. This article decodes the complex product lineup and explains what Work is, where it fits in OpenAI's lineup, the many interesting choices in its design, the tensions underneath, and where the product is likely headed. OpenAI plans to merge Chat and Work, so Work is really a preview for how ChatGPT's billions of users will soon use the app.
Read MoreAugust 05, 2026 09:27 AM
A developer recorded the requests Codex generated for a 16-character prompt by pointing it at a custom local server. They measured what changed as it loaded instructions, exposed tools, read files, ran commands, received images, and compacted its history. The experiment did not call an external model. This post details the findings from the experiment.
Read MoreAugust 05, 2026 09:27 AM
Shieldstral features a 3B open-weights multimodal safety classifier that outperforms models up to 7x its size. It accepts plain-language policies at inference time, unifying text and image safety evaluation without retraining. Shieldstral delivers calibrated safety scores across diverse benchmarks while running efficiently on a single 16GB NVIDIA GPU. It is a step toward moderation that adapts to context instead of forcing every product through one frozen taxonomy.
Read MoreAugust 05, 2026 09:27 AM
NVIDIA has released Alpamayo 2 Super under a commercial license for robotaxis and autonomous vehicles. The reasoning model was designed to handle rare driving scenarios with inspectable decisions and broad multitask capabilities.
Read MoreAugust 05, 2026 09:27 AM
Kiro Crew is a persistent workspace for development work that self-improves and continues beyond one session. It can run locally or remotely on personal hardware. Kiro Crew allows developers to start work from within a desktop app, web dashboard, or CLI, and continue the same work through connection tools like Slack and Discord. It can run multistep tasks and recurring jobs on a schedule and also monitor systems until something needs attention.
Read MoreAugust 05, 2026 09:27 AM
DiffusionGemma adapted Gemma 4 into a discrete diffusion model that refined 256-token blocks in parallel, reaching roughly 1,500 output tokens per second on a single NVIDIA H100.
Read MoreAugust 05, 2026 09:27 AM
Computer-use verification lets coding agents reproduce bugs, validate implementations against specs, and attach screenshots or videos to pull requests. Integrated into triage, implementation, and review workflows, cloud-based subagents reduce human review burden and enable iterative self-debugging.
Read MoreAugust 05, 2026 09:27 AM
SpaceX's capital expenditures hit $18.4 billion in the most recent quarter. The bulk of that figure is tied to the company's ongoing AI build-out. The company is plowing money into terrestrial computing infrastructure, signing AI deals, and working toward launching orbital data centers. SpaceX is on track to have $100 billion in annualized recurring revenue by December, most of it coming from data center deals.
Read MoreAugust 05, 2026 09:27 AM
NemotronLabs VoiceChat is an 11B end-to-end speech model that handles streaming understanding, speech generation, and tool calling within one architecture.
Read MoreAugust 05, 2026 09:27 AM
Backflip AI developed a model that converts physical parts into digital CAD files in minutes for around $10.
Read MoreAugust 05, 2026 09:27 AM
Cursor's Mixture-of-Kittens (MoK) is an open-source optimized Mixture-of-Experts (MoE) megakernel that improves efficiency on NVL72s GPUs, which significantly boosts performance for models like Composer by addressing computation and communication bottlenecks.
Read MoreAugust 05, 2026 09:27 AM
LFM2.5-2.6B, a 2.6B parameter on-device agentic model, enables free inference, low latency, and robust privacy by running locally on hardware like phones or CPUs.
Read More