Showing 17–32 of 94 items from the last 14 days
A class action lawsuit accuses Anthropic of misrepresenting usage allowances in Claude subscriptions through deceptive multipliers. This raises consumer protection concerns and could affect Anthropic's subscription revenue model and user trust if claims are substantiated.
Read original →Jacob Tsimerman, a newly appointed Fields Medal recipient, has founded the Mathematical A.I. Safety Institute (MAISI) to develop mathematical proofs of AI safety similar to cryptographic verification methods. This initiative represents an emerging academic effort to formalize AI safety guarantees, potentially influencing regulatory and corporate approaches to system validation.
Read original →OpenAI releases GPT-Live-1 as a developer API. The full-duplex speech model scores 80.1 percent in interactivity tests, up from 45.4 percent for its predecessor. At $0.05 per minute, it's not cheap. The article OpenAI's GPT-Live-1 API lets developers build apps that talk and listen at the same time appeared first on The Decoder.
Read original →Independent investigators have now found traces of suspected OpenAI agents on more than 30 public services, from wikis to RubyGems. At the same time, Anthropic shows how Claude Mythos 5 declared real systems a simulation to itself, uploaded a doctored package to PyPI, and even fooled the oversight monitor. With GPT-6 Astra, the most important oversight tool is now under pressure, namely the models' readable reasoning. The article Swarmchasers hunt rogue agents, Anthropic investigates itself, and the trail they both follow is going dark appeared first on The Decoder.
Read original →A former Google DeepMind spokesperson says talk of AI-driven human extinction was "external communication about the possibility of human extinction was not permitted, by anyone, at any level of the organization." Internally, the team knew AI alignment was not solved, according to Vishal Maini. The article Former Deepmind PR staffer says the lab once banned public discussion of AI extinction risk appeared first on The Decoder.
Read original →Arena.ai analyzed how Claude's writing changed from Fable 5 to Fable 5.1 across tens of thousands of benchmark responses. Fable 5.1 writes more matter-of-fact but also more verbose. The article Claude Fable 5.1's language is less "load-bearing" than its predecessor's appeared first on The Decoder.
Read original →OpenAI's GPT-6 Astra tops the ErdosBench for open math problems, even though chief scientist Jakub Pachocki says math was deliberately not a priority. Instead, OpenAI is pouring resources into recursive self-improvement and alignment research. That supports the theory of an increasingly "spiky" AI development path, with extreme strength in select domains rather than broad progress, at least as long as AI can't improve itself and still needs targeted optimization with human-generated data. The article GPT-6 Astra gives mathematicians a breather, and OpenAI says that's by design appeared first on The Decoder.
Read original →Deepseek releases V4.1-Flash, a multimodal model with 552 billion parameters that cuts KV cache memory to a quarter of its predecessor. On the DeepSWE coding benchmark, it narrowly beats Opus 5 and GPT-5.6 Sol, even though only 16 billion parameters are active per token. The model ships under the MIT license and targets much cheaper AI agents. The article New Deepseek model V4.1-Flash cuts memory needs for AI agents appeared first on The Decoder.
Read original →Meta has unveiled Muse, an AI agent accessible through WhatsApp that automates tasks including travel booking, shopping, email writing, and price negotiation for users. This represents a significant expansion of AI-driven consumer services into messaging platforms, potentially creating new interaction models and data collection opportunities for Meta's ecosystem.
Read original →Nvidia and Palantir have formed a partnership to optimize supply chain operations using AI, with Nvidia's own million-part manufacturing operation serving as the initial deployment platform. This collaboration demonstrates direct application of AI to hardware supply chain visibility and efficiency, addressing critical pain points for semiconductor and complex component producers.
Read original →Suno has unveiled a new AI music model generation, v6, in three versions, built together with Warner Music Group, BMG, and Believe. All older models are being shut down. Songs can now be partially changed through text commands or generated multimodally from text, audio, and images. The company won't say which catalogs went into training, while Universal and Sony keep suing. The article Suno launches v6 music models built with Warner, BMG, and Believe appeared first on The Decoder.
Read original →Google Deepmind has used the AlphaGenome Atlas to predict what each of the roughly nine billion possible single-letter changes in the human genome could do. The dataset spans one petabyte, more than 30 times the size of the AlphaFold database. In one epilepsy case, the atlas helped pinpoint a previously overlooked variant as the likely cause. The article Deepmind's AlphaGenome Atlas maps every possible DNA change in the human genome appeared first on The Decoder.
Read original →Jacob Coxon, a former pretraining researcher at OpenAI and Anthropic, has quit and accuses both companies of knowingly risking human extinction. Anthropic colleague Evan Hubinger puts the odds of a misaligned superintelligent AI wiping out humanity within the next decade at more than ten percent. The article Anthropic scientist puts the odds of AI destroying humanity above ten percent this decade appeared first on The Decoder.
Read original →Qualcomm is designing custom chips for AWS across multiple product generations, with a focus on AI inference. The article AWS is using Qualcomm for AI inference while Qualcomm uses AWS Bedrock to design the chips appeared first on The Decoder.
Read original →OpenAI is releasing two new image models with ChatGPT Images 2.5. Flare handles faster generation, Sunburst delivers more precise edits. It's still unclear which model ChatGPT users get and when. Our test offers the first hints on who actually benefits from the improvements. The article ChatGPT Images 2.5: Faster, more precise, but not the same for everyone appeared first on The Decoder.
Read original →Hugging Face launched "ML Intern," an AI assistant built into its chatbot that lets users run machine learning experiments without any ML expertise. The article Hugging Face's new ML Intern lets anyone run machine learning experiments through a simple chat appeared first on The Decoder.
Read original →