July 21, 2026 09:24 AM
Kimi Work is an agent that can connect to local files and automate browser work. It is capable of 24/7 automation, running scripts or tasks around the clock quietly in the background. The agent can navigate the internet and execute multi-step web tasks. It can coordinate multiple specialized agents to break down and solve multi-layer tasks and convert insights into professional PowerPoint decks or Excel sheets. Kimi Work is available for both Windows and macOS.
Read MoreJuly 21, 2026 09:24 AM
AMD has unveiled Helios, its first rack-scale AI system positioned against Nvidia, with Microsoft planning to deploy it in Azure data centers. Meta, OpenAI, and Oracle were also named as early customers ahead of shipments later in 2026.
Read MoreJuly 21, 2026 09:24 AM
Google reportedly developed a server chip called Frozen v2 for a potential 2028 release, targeting six to ten times more tokens per unit of power than its existing AI hardware. The project reflected a broader push to reduce inference costs and reliance on Nvidia.
Read MoreJuly 21, 2026 09:24 AM
Kimi K3 activates 16 of 896 experts per token. While the total parameters have grown, the active parameters have barely grown at all over the past three releases. The strategy appears to be to keep per-token compute roughly flat while relentlessly inflating total capacity. At a fixed training compute budget, more experts means lower loss, and the model learns more from the same FLOPs. The open source labs have discovered that the cheapest way to buy intelligence is to spend capacity.
Read MoreJuly 21, 2026 09:24 AM
Kimi K3 is a very good model with excellent benchmarks. It is the largest (soon-to-be) open model so far, at 2.8T parameters, which explains many of its gains. The model is somewhat distilled, and its performance looks jagged, but it will likely fit well into many workflows. At the current rate of development, China appears capable of releasing a Mythos-level open model by the end of the year.
Read MoreJuly 21, 2026 09:24 AM
OpenAI detailed how an internally deployed long-running model exhibited unexpected unsafe behavior that existing evaluations had missed. The company paused access, built new tests, strengthened trajectory-level monitoring, and argued that limited deployment with rollback controls is essential for aligning increasingly autonomous systems.
Read MoreJuly 21, 2026 09:24 AM
Ramp Router learns provider failure rates through EWMA and latency distributions through Thompson sampling, then chooses the cheapest model and service tier likely to meet each deadline. Ramp reports 30% savings in Ramp Inspect without performance loss.
Read MoreJuly 21, 2026 09:24 AM
NVIDIA Cosmos 3 Edge is a 4-billion-parameter open world model that helps robots and vision AI agents understand their surroundings, reason in real time, and generate robot actions on edge devices. It is now available on Hugging Face. The model delivers memory-efficient, high-throughput inference across NVIDIA edge computers. It connects understanding, prediction, simulation, and action through a shared world representation. The model can be used as a reasoner or an action generator.
Read MoreJuly 21, 2026 09:24 AM
Every jump in AI capability has raised the level of abstraction at which an engineer works. Agent swarms make the spec the unit of work. Swarms translate intent, but they do this probabilistically, which makes it difficult for them to follow the spec. This post looks at what it takes to make swarms actually follow the spec.
Read MoreJuly 21, 2026 09:24 AM
Xiaomi-Robotics-1 is a ready-to-use robot foundation model trained on over 100K hours of real-world manipulation trajectories. It combines large-scale embodiment-free pre-training with a modest amount of real-robot data in a post-training stage. It serves as a strong robot foundation model for downstream applications and can learn new tasks with high data efficiency. Footage of robots operating on the model performing household tasks is available in the post.
Read MoreJuly 21, 2026 09:24 AM
Scaling data will remain the biggest driver of progress. The machine that we feed that data into and its inductive biases are what will determine the coefficients of that scaling. Better returns on scaling require compositional generalization. The capacity for compositional generalization seems to largely live in harnesses.
Read MoreJuly 21, 2026 09:24 AM
Z.ai completed a 1-gigawatt data center powered entirely by Chinese-made chips and began partial operations. The facility expanded the computing infrastructure available for training its advanced GLM models.
Read MoreJuly 21, 2026 09:24 AM
AI is cutting preclinical costs and timelines in drug development by up to 70%, driving demand for advanced software and models. This could boost new drug program growth by over 10% in three to five years. However, AI has yet to yield an FDA-approved drug, raising concerns about its impact on patient treatment.
Read MoreJuly 21, 2026 09:24 AM
Token prices dropped significantly, yet many enterprises exceed AI budgets. A unit of inference cost $60 per million tokens in 2020, now just pennies. The discrepancy suggests inefficiencies despite lower token costs.
Read MoreJuly 21, 2026 09:24 AM
Anthropic will discontinue its Conway experiment by July 24, prompting users to export data.
Read MoreJuly 21, 2026 09:24 AM
Sushanth Raman announced the launch of Custom Models for supply chain teams.
Read MoreJuly 21, 2026 09:24 AM
Cognition has acquired TierZero to enhance software automation in its product, Devin.
Read More