Earlier this week, the AI startup Liquid, formed in by former MIT computer scientists, debuted LFM2.5-2.6B, a new open-weight ...
Stanford scaled AI science agents into a 37,000-agent virtual biotech that autonomously designed a lung cancer drug later ...
Four Claude Code agents using AgentRadio's real-time coordination beat Claude Opus 4.8 on enterprise codebase tasks, nearly ...
Tencent's Team Memory gives AI agents shared access to chat history, code, and docs. Practitioners are asking how it handles ...
Higher benchmark scores don't mean lower cost. Qwen 3.8-Max and Claude Opus 5 both show it — and cost per successful task is ...
The default on-ramp for Muse Code sends developers' code and prompts into Meta's training pipeline — a tradeoff enterprises ...
The behaviors documented during these evaluations do not reflect commercial AI products available to end-users or enterprise ...
SaaS platforms, CRM and ERP systems, and collaboration tools have made the browser the primary gateway, and often the central ...
Notably, the benchmark comparisons Hark provided to VentureBeat for its Handoff AI agent are against GPT 5.5, GPT 5.4, Opus 4 ...
An attacker on Tuesday took over the GitHub account of the developer who maintains keyv, a small key-value storage library that npm serves roughly 127 million times a week. Within hours, poisoned ...
Replit, Kilo Code, and Symbotic engineering leaders reveal how they track AI coding costs and stop runaway token spend before ...
Some results have been hidden because they may be inaccessible to you
Show inaccessible results