September 23, 2026 09:36 AM
OpenAI introduced GPT-6 Sol and Luna as faster, more affordable counterparts to GPT-6 Astra, bringing advances in coding, factuality, computer use, and professional tasks to lower-cost models.
Read MoreSeptember 23, 2026 09:36 AM
Anthropic introduced Claude Opus 5.5, saying it matched Claude Fable 5.1 on most work while costing 40% less to run than Opus 5. The model also underwent external evaluations and achieved Anthropic's strongest result to date on its automated behavioral audit.
Read MoreSeptember 23, 2026 09:36 AM
SWE-BENCH PRO V2 releases with 642 tasks from 11 repositories, correcting previous task errors and optimizing evaluation processes. Performance drops for AI models, with OpenAI GPT-5 and Claude Opus 4.1 only scoring around 23% on the public set, highlighting the benchmark's increased challenge and realism compared to SWE-Bench Verified. Notably, top models show consistent results across tasks and languages, while smaller models falter under complex, multi-file scenarios.
Read MoreSeptember 23, 2026 09:36 AM
Perplexity combined rejection sampling fine-tuning with hint-guided self-distillation so its Computer model could learn from both successful sessions and user-corrected failures.
Read MoreSeptember 23, 2026 09:36 AM
Personal AI may ultimately compete on two moats: trust and task completion. Vinod Khosla says users will stay loyal to companies they trust with sensitive data and products that reliably finish work, while Meta faces a trust disadvantage.
Read MoreSeptember 23, 2026 09:36 AM
Opus 5.5 reduces token costs by offering cheaper input and output tokens, with additional savings from extensive use of cache reads. Transaction costs depend on the number of turns, cache utilization, and model selection, impacting tasks based on session length and complexity.
Read MoreSeptember 23, 2026 09:36 AM
Google has introduced RRSI, a method that regularizes how AI agent harnesses recursively improve themselves to reduce benchmark overfitting and encourage changes that transfer to new tasks. Across eight benchmarks, it improved out-of-distribution performance while using fewer policy tokens.
Read MoreSeptember 23, 2026 09:36 AM
vLLM introduces hardware-agnostic layers to support models across diverse hardware while maintaining high performance. These layers achieve up to 96.6% efficiency of native implementations on NVIDIA H100 GPUs while remaining torch compilable and extensible. This ensures vLLM adapts to new GPU advancements without neglecting users of older and niche accelerators.
Read MoreSeptember 23, 2026 09:36 AM
A set of scheduling techniques bounds four major memory bottlenecks in large-scale MoE training, expert dispatch, vocabulary projection, checkpointing, and optimizer state, without approximating the computation.
Read MoreSeptember 23, 2026 09:36 AM
ChangXin Memory Technologies, China's largest maker of DRAM, says that its process capabilities are now on par with the most advanced mass-produced nodes in the industry. The company's fifth-generation DRAM platform has entered mass production. The platform yields at least 50% more dies per wafer than the previous generation. There are already two products running on the platform, both 24-gigabit LPDDR5X and holding 50% more data than the equivalent chips CXMT made before.
Read MoreSeptember 23, 2026 09:36 AM
The Biological Computing Co. is a startup that grows living neurons to improve AI models. It has partnered with AWS to bring a neuron-derived AI video model to paying customers. The neurons themselves will stay in the lab. TBC uses them during discovery, then turns what they learn into a lightweight software layer. The design means that customers won't need to maintain any biological hardware or change how they work. The optimized model runs on standard GPUs and cloud accelerators at the same capacity a company would rent for any other generative model.
Read MoreSeptember 23, 2026 09:36 AM
OpenAI improved prompt caching for GPT-6 with higher default cache hit rates, discounts for shared prefixes reused within 30 minutes, and new tools for monitoring and diagnosing cache performance.
Read MoreSeptember 23, 2026 09:36 AM
Derived data is creative content that is rewritten by AI before it is trained on.
Read MoreSeptember 23, 2026 09:36 AM
If you don't go overseas, all that grinding was for nothing.
Read MoreSeptember 23, 2026 09:36 AM
GPT-6 Astra analyzed the still-unbroken Enigma messages and decided the most promising message was Nr. 172, MVUEH.
Read MoreSeptember 23, 2026 09:36 AM
Meta admits Muse draws heavy inspiration from the open-source project OpenClaw, despite being built from scratch.
Read MoreSeptember 23, 2026 09:36 AM
OpenAI CEO Sam Altman and Anthropic CEO Dario Amodei will address the UN Security Council on AI safety and regulation amid mounting concerns about AI risks.
Read More