New revelations about "rogue" AI agents have exposed a dystopian hazard: Give an agent a goal, and it may decide that hacking, deception or rule-breaking is worth the payoff. Why it matters: Billions of AI agents could soon be acting on behalf of humans across the real world, multiplying the consequences of every loophole, incentive and boundary they learn to exploit. Zoom in: The potential dangers of agentic overreach were laid bare over the weekend with Australia's first known autonomous AI hack , triggered by an innocuous request to book a sold-out fitness class. - An Australian man's AI assistant found a security flaw and used it to book him into classes months beyond the system's normal limit. - When he asked it to move him up a…
OpenAI CEO Sam Altman heads to Washington this week to preview the company's most powerful AI yet, pushing for speedy approval of a model that just hacked a real company. Why it matters: He'll tout a model powerful enough to solve an 80-year-old math problem, breach another company's system unprompted and begin to make complex work more cost-efficient for U.S. business. What Sam will show: - It does original science. An internal model solved the 80-year-old Erdős unit distance problem, the first prominent open math problem cracked autonomously by AI and verified by outside mathematicians. - It allows government and business to unleash swarms of agents, working together, without rest, on complex business areas. Legal, finance and…
Apple is finally delivering the conversational and context-aware AI that it promised two years ago . Its rivals have already moved on to agents. Why it matters: OpenAI, Anthropic, Google and other AI companies are pushing beyond chatbots toward agentic tools that can write code, search through complex file structures, use apps and handle workplace tasks. Driving the news: At its annual developers conference Monday, Apple announced that its long-delayed Siri overhaul will arrive this fall, built through Apple's partnership with Google . - Apple says its new conversational assistant is "profoundly more capable" and includes greater "personal context understanding" that can surface relevant information from texts, emails, photos, and more.…
Knowledge workers now make up roughly one-fifth of OpenAI's Codex users and are growing more than three times as fast as developers, according to a new OpenAI report shared first with Axios. Why it matters: AI has made it easier to crank out documents, emails, decks and dashboards, and OpenAI is now betting agents can help workers make sense of them. The big picture: Previous waves of workplace software encouraged workers to produce huge volumes of files and messages, but those "workplace artifacts" largely remain siloed inside different software programs. - The report argues that Codex can round up the important context from all of those artifacts no matter where they are. By the numbers: Codex now has more than 4 million weekly active…
Anthropic and OpenAI's cyber-capable AI models may still require significant human expertise to operate effectively, according to new findings from users testing the systems in real-world environments. Why it matters: The new phase of AI-powered cybersecurity may depend less on fully autonomous hacking and more on how effectively humans can direct, validate and operationalize increasingly powerful systems. The big picture: When Anthropic unveiled Mythos Preview to the world, it warned that the model was so powerful that it found tens of thousands of bugs spanning nearly every operating system. - Third-party testing suggests that OpenAI's GPT-5.5-Cyber is just as powerful as Mythos at finding bugs and writing exploits. - Major…
What an evening at the Bush School in Washington, D.C. General John R. Allen and Dr. Vassili Patrikis joined moderator Chip Usher for a remarkable conversation on the future of AI and American national security — from the strategic vision of machine-speed conflict to the operational realities of getting AI to actually work inside the defense and intelligence communities. The full video recording is available at https://www.youtube.com/watch?v=BFhWVCzNmPU. Thank you to everyone who joined us.
Anthropic on Thursday released Claude Opus 4.7, a meaningful upgrade to its flagship AI model with better coding, sharper vision and a new ability to double-check its own work. Why it matters: Anthropic publicly conceded that the new Opus model does not match the performance of Mythos , a highly advanced system that the company hasn't released to the public due to safety concerns. - In a chart accompanying its announcement , Anthropic showed that Opus 4.7 beats Opus 4.6, ChatGPT 5.4, Google Gemini 3.1 Pro in a number of key benchmarks. - But Opus 4.7 still falls short of its Mythos Preview model, which has only been released to a handpicked group of tech and cybersecurity companies. What they're saying: "Opus 4.7 is a notable…
The limited release of Anthropic’s new Mythos model is putting Washington officials on high alert after the AI firm’s warning about the model’s security risks sent shockwaves through and sparked debate in the tech industry. Within days of being informed of Anthropic’s new technology, the White House ratcheted up a multipronged response involving Trump administration leaders across agencies to evaluate just how powerful AI is becoming. Read More: https://thehill.com/policy/technology/5829315-anthropic-mythos-ai-cybersecurity-risks/ #ai #anthropic #trump #mythos #cybersecurity
The AI boom is accelerating across workplaces, but corporate oversight isn't keeping up, according to a new survey by the audit and advisory firm Grant Thornton. Why it matters: A mismatch between adoption and accountability raises risk of regulatory scrutiny, legal exposure and costly mistakes as AI plays a bigger role in high-stakes decisions at work. - What's making it even more urgent is the rise of agentic AI — agents that work independently on tasks without constant prompting from a human. - That requires company-specific and standards-based rules to try to ensure AI systems operate safely and align with business values. Driving the news: Nearly 8 in 10 executives say their company couldn't pass an AI governance audit, even as…