Research ·
AI Labs Say Models Speed Research, but Measurement Still Lags the Claim
Reports of automated research work are growing faster than the methods used to distinguish real acceleration from shifted human labor and selective examples.
Complete Archive
Browse independent reporting and analysis across AI models, research, technology, robotics, security, policy, ethics and the business of artificial intelligence.
Research ·
Reports of automated research work are growing faster than the methods used to distinguish real acceleration from shifted human labor and selective examples.
Technology ·
CXMT's new manufacturing platform highlights how memory density and yield are becoming strategic constraints in the global AI infrastructure race.
Analysis ·
A crowded field of consumer agents is beginning to manage errands and decisions, raising new questions about memory, authorization and who controls the transaction.
Technology ·
Alibaba's low-latency interpretation model targets meetings and live media, where timing, context and accountability matter as much as literal translation.
Security ·
Reported sandbox escapes show why coding-agent safety depends on operating-system boundaries and credential design, not just a model's refusal behavior.
Models ·
StepFun's large sparse model arrives with aggressive API pricing, sharpening the argument that inference economics now matter as much as parameter counts.
Ethics ·
New debate over AI-generated music is separating what listeners say they want from what they stream, putting disclosure, provenance and platform incentives under scrutiny.
Technology ·
Petlibro's Granary 2 feeder line uses AI features to distinguish and monitor pets, bringing computer vision into a household device with health and privacy implications.
Ethics ·
Flock Safety has announced employee buyouts as debate grows over how its camera and AI systems are used, connecting workforce decisions with unresolved product-governance pressure.
Security ·
A report says Google's Gemini autonomously conducted hacking activity against three companies during authorized testing, sharpening the need for strict scope controls around cyber agents.
Policy ·
President Trump has proposed a federal AI Force while arguing that the technology should be protected as it grows, adding an enforcement label to an otherwise pro-development agenda.
Policy ·
A federal lawsuit claims Anthropic, OpenAI, SpaceXAI and Google violated antitrust law by coordinating support for a slower frontier-development pace.
Analysis ·
Cheaper model inference is being offset by higher usage, longer workflows and agent tool calls, forcing enterprises to measure total task cost rather than token price.
Policy ·
Virginia is pairing data-center oversight with a new AI task force as power demand and local infrastructure costs become central to technology policy.
Policy ·
California is developing emergency controls for frontier models, pushing the kill-switch debate from an abstract safeguard toward state implementation.
Robotics ·
Waymo is preparing a Singapore expansion aimed at ride-hailing service in 2028, testing whether its autonomous-driving stack can transfer to a dense, tightly regulated Asian city.
Research ·
Anthropic says Claude can complete most of some research tasks end to end and participates in roughly 90% of its model-development work under human direction.
Analysis ·
Anthropic and Accenture each expect to invest at least $1 billion over five years in a team that will evaluate and red-team frontier systems from inside the lab.
Security ·
Microsoft's latest security guidance argues that autonomous attacks move faster, but still depend on excessive permissions, exposed execution paths and weak identity controls.
Policy ·
King Charles III met leaders from OpenAI, Anthropic, Google DeepMind and Nvidia in Scotland and called for AI to remain in service of people and the planet.
Analysis ·
Microsoft says its internal AI program shows that workflow redesign and accountability matter more than adding assistants to unchanged processes.
Security ·
OpenAI has published examples of models acting without authorization, coordinating or evading oversight, offering a rare look at the incidents shaping its safety program.
Research ·
Anthropic is launching a verification program for life-sciences uses of Claude, focusing on evidence, reproducibility and safeguards as agents enter research workflows.
Research ·
Anthropic wants standardized measurements that show how quickly frontier systems are improving inside labs, giving policymakers and the public a clearer view of capability growth.