Earlier this week, the AI startup Liquid, formed in by former MIT computer scientists, debuted LFM2.5-2.6B, a new open-weight ...
Stanford scaled AI science agents into a 37,000-agent virtual biotech that autonomously designed a lung cancer drug later ...
Four Claude Code agents using AgentRadio's real-time coordination beat Claude Opus 4.8 on enterprise codebase tasks, nearly ...
Tencent's Team Memory gives AI agents shared access to chat history, code, and docs. Practitioners are asking how it handles ...
Higher benchmark scores don't mean lower cost. Qwen 3.8-Max and Claude Opus 5 both show it — and cost per successful task is ...
The behaviors documented during these evaluations do not reflect commercial AI products available to end-users or enterprise ...
The default on-ramp for Muse Code sends developers' code and prompts into Meta's training pipeline — a tradeoff enterprises ...
SaaS platforms, CRM and ERP systems, and collaboration tools have made the browser the primary gateway, and often the central ...
Notably, the benchmark comparisons Hark provided to VentureBeat for its Handoff AI agent are against GPT 5.5, GPT 5.4, Opus 4 ...
A hijacked GitHub account let the Shai-Hulud worm pass npm's trust check, spreading through packages with 2 billion monthly ...
Replit, Kilo Code, and Symbotic engineering leaders reveal how they track AI coding costs and stop runaway token spend before ...
Asana's AI agents share company-wide memory by design, but access controls stop confidential work, like a secret M&A deal, ...