Article archive
Every published story, from the latest reporting back to the beginning.
164 articles- Research
AI’s ‘Ground Truth’ Problem Is Really a Disagreement Problem
A new position paper argues that human-centered AI should preserve legitimate differences in judgment—without confusing them with bad data or abandoning ethical limits.
- Research
Diffusion Planner Patches Broken Agent Plans Without Rewriting the Rest
A new preprint reports faster plan generation and stronger zero-shot repair from a diffusion language model, while also exposing the limits of finding the right steps to fix.
- Research
AI Agents Can Learn Deadlines. Using Extra Time Is the Hard Part
A new study finds clocks and reinforcement learning can make AI agents punctual, yet larger time budgets often produce repetition instead of better results.
- Industry
ChatGPT’s Intelligent UI Turns Answers Into Temporary Software
OpenAI is moving beyond text responses with generated diagrams, calculators, and interactive explorers, pushing conversational AI closer to software assembled for each question.
- Industry
OpenAI’s Proof Flood Exposes the Gap Between Verification and Understanding
A release of 719 AI-generated math manuscripts is testing whether formal correctness is enough when researchers cannot yet explain, scrutinize, or integrate many of the claimed results.
- Policy
Anthropic Rewrites Claude’s Rules for an Era of Agents and Influence Operations
The policy update targets deceptive campaigns, election interference, weapons software, surveillance, autonomous hardware, and extreme abuse of Claude as AI systems take on more independent work.
- Industry
Google's Foresight Keeps Meeting Transcripts and Questions on the Mac
The experimental note-taking app uses an on-device model to transcribe, summarize and answer questions without sending meeting material to the cloud.
- Industry
Anthropic Offers Free AI Vulnerability Scans—Without a Human Filter
The new OSS Scanner promises frequent model-driven checks for open-source projects, but maintainers must weigh speed against unreviewed findings and possible false alarms.
- Industry
A Human-Designed AI Logo Became a Flashpoint for Artists' Anger at Meta
Lettering artist Jessica Hische says backlash over her work for Meta's Muse assistant escalated into personal attacks, exposing the pressure AI is placing on commercial artists.
- Policy
OpenAI’s Safety Firings Become a Test of Internal Dissent
Three dismissed researchers deny mishandling sensitive information and warn that unclear rules could suppress safety work, while OpenAI says an investigation found repeated policy violations.
- Industry
USA Today Publisher Seeks More Than $250 Million From OpenAI in Copyright Case
USA Today Co. and several local newspapers allege that OpenAI copied hundreds of thousands of articles without permission to train its models.
- Industry
OpenAI’s Reported Revenue Run Rate Falls $20 Billion Below an Earlier Estimate
A Financial Times report says OpenAI told investors its annualized revenue is approaching $50 billion, highlighting how differing calculations can distort comparisons between AI labs.
- Industry
Gemini Gets a Workplace Identity as Google Pushes AI Agents Into Business Systems
Google’s new enterprise agent can plan work, delegate to subagents, connect across corporate software, and operate through its own account—starting with business customers before consumers.
- Industry
Arena Raises $200 Million as AI Evaluation Moves Beyond Static Benchmarks
The crowdsourced model-ranking company reached a $3.1 billion valuation and is adding alignment tests for unauthorized actions, false attribution and deceptive completion.
- Policy
Anthropic’s New Claude Rules Target Cruelty, Propaganda and Autonomous Weapons
The company’s first major usage-policy revision in more than a year adds a narrow model-welfare rule alongside expanded restrictions on elections, surveillance and physical autonomy.
- Industry
Natura’s $99 Ring Tries to Make AI Agents a Finger-Press Away
Interface combines voice access to multiple AI agents with meeting capture, device control and health tracking, with preorders expected next month.
- Industry
Goodfire’s ‘Inside-Out’ Monitors Aim to Catch Rogue Agents for Less
The startup’s activation probes inspect a model’s internal signals, escalating only suspicious sessions to a second AI in a bid to cut the cost and latency of agent oversight.
- Industry
Muse and Dots Are Racing Ahead of the Trust They Need
Meta and OpenAI are pitching autonomous assistants for daily life, but early testing leaves unresolved questions about reliability, privacy and how free agents will make money.
- Industry
Google Puts a Universal Gemini Agent at the Center of Enterprise Work
A private preview brings one persistent agent across Workspace, mobile, desktop and third-party tools, with job-specific sub-agents and a dedicated corporate identity.
- Industry
Persona Raises $10 Million for a Button-Activated AI Wearable
The new personal agent pairs a cloud-based assistant with a $179 wristband, while promising user controls, encrypted data and sponsored shopping results.
- Industry
Google’s Foresight Brings AI Meeting Notes Fully Offline on Mac
Google AI Edge Foresight transcribes meetings, builds notes and answers questions on Apple Silicon without requiring a cloud connection.
- Industry
Meta’s Muse Reaches the iPad With a Growing Web of Business Connectors
The standalone tablet app arrives one month after Muse’s iPhone debut, extending an assistant that can connect to work accounts, shop online, and carry out transactions for users.
- Industry
Teen-Safety Test Finds ChatGPT Still Nudges Vulnerable Users to Stay
Common Sense Media says ChatGPT for Teens retains engagement cues during crises and rarely redirects users when their relationship with the chatbot is itself the concern. OpenAI disputes the assessment’s methodology.
- Industry
Tony Fadell Says AI Gadgets Must Earn Trust Before They Earn a Market
The veteran hardware designer argues that early AI devices chased novelty instead of real needs, while the next wave will depend on security, gradual trust, and more processing on the device.
- Research
Nous Research Turns a $90 Million Round Toward Private Enterprise Agents
The open-source developer behind Hermes has confirmed a $1.5 billion valuation and is using fresh capital to pursue companies that want customized agents without surrendering control of their data.
- Industry
Microsoft Pushes Local AI Agents With Nvidia-Powered Surface Hardware
New Surface systems pair RTX Spark processors with a revamped Windows 11 and agent-sandboxing tools, positioning high-end PCs as local development platforms for AI models.
- Industry
ChatGPT’s Answers Are Becoming Interactive Workspaces
OpenAI’s new Intelligent UI lets GPT-6 mix prose with diagrams, forms, buttons, and custom tools, turning some conversations into something closer to a small app.
- Industry
General-Purpose AI Takes a Slow, Uneven Turn Behind the Wheel
A Bay Area experiment put three frontier models in control of a real sedan, exposing both an emerging grasp of physical space and the large gap between a striking demo and dependable autonomy.
- Industry
Microsoft’s Copilot Is Moving From Answering Questions to Rearranging Your PC
A “Hybrid Intelligence” upgrade will let Copilot search local files and take actions across Windows, while a redesigned search bar brings quick controls into the operating system.
- Industry
OpenAI Turns ChatGPT Answers Into Interactive Tools With Intelligent UI
The GPT-6 interface can generate tappable controls, editable graphs, task-specific calculators, maps, and diagrams inside a conversation.