Showing 81–96 of 100 items from the last 14 days
Alibaba has unveiled Qwen 3.8, a multimodal AI model with 2.4 trillion parameters that the Qwen team says rivals leading models and trails only Fable 5. A preview is available now. The article Alibaba's Qwen takes on Kimi K3 with open-weight Qwen 3.8, says model is "second only to Fable 5" appeared first on The Decoder.
Read original →Google Deepmind's GenCeption repurposes a video generator for classic vision tasks such as depth estimation and segmentation, matching state-of-the-art systems with far less training data. The model trained almost entirely on synthetic videos. Its results add to the debate over whether video generators already contain a kind of universal world model. The article Google Deepmind argues video generators already contain the world models computer vision has been missing appeared first on The Decoder.
Read original →Moonshot's Kimi K3 is the first Chinese model to top the Code Arena: Frontend rankings, beating Claude Fable 5 and GPT-5.6 Sol by a wide margin. But on advanced math, the gap is stark: Kimi K3 scores only about 39 percent on FrontierMath Tier 4, while models from OpenAI and Anthropic hit close to 90. The article Moonshot's Kimi K3 outperforms Fable 5 in frontend code but lags far behind in complex math appeared first on The Decoder.
Read original →Epoch AI tested three leading AI text detectors (Pangram, GPTZero, and Originality.ai) using style-imitated texts. Up to 18 percent of AI-generated passages went undetected. For scientific writing, the miss rate climbed as high as 48 percent, the very genre where these detectors likely see the most real-world use. The article AI text detectors struggle when language models mimic an author's style appeared first on The Decoder.
Read original →The RadLE 2.0 benchmark tests whether AI models in radiology can tell when they should leave a diagnosis to a human. Many models deliver wrong findings with full confidence, and human radiologists are still well ahead. Before AI can diagnose on its own, it needs to learn when it's better to say nothing. The article AI chatbots reading X-rays can be dangerously confident even when they're wrong appeared first on The Decoder.
Read original →I'm loyal to Google phones, but some changes need to be made.
Read original →China announced 5,000 AI training slots for Global South countries and launched the World Artificial Intelligence Cooperation Organization with cooperation centers planned for ASEAN, the African Union, and BRICS. This represents a systematic effort to establish parallel AI governance structures outside Western-led frameworks, potentially fragmenting global AI standards and creating competing technology ecosystems.
Read original →Open-weight AI models including GLM-5.2 and DeepSeek V4-Pro now match frontier closed models' cyber capabilities from just four months ago at substantially lower cost, while safety measures on open models remain largely ineffective. This acceleration compresses the development cycle advantage of proprietary systems and creates security risks as capable models become widely accessible.
Read original →The US Department of the Navy has signed a strategy to "weaponize" data and AI and build an "AI-first" fleet. Large language models would run directly on warships, and an AI war council would prioritize mission scenarios. The core message is that moving too slowly carries greater risks than "imperfect alignment." The article The Pentagon's new AI playbook treats slow adoption as a bigger risk than imperfect alignment appeared first on The Decoder.
Read original →Anthropic will include Claude Fable 5 in Max and Team Premium plans starting July 20, but at just 50 percent of regular limits, which themselves drop by a third that same day. Pro users get a one-time $100 credit, then pay API rates. The reversal from Anthropic's original plan to pull Fable from subscriptions entirely likely comes down to competitive pressure from OpenAI's cheaper GPT-5.6 Sol. The article Anthropic slashes Claude Fable 5 limits in Max and Team Premium and pushes Pro users toward API pricing appeared first on The Decoder.
Read original →Meta is reportedly in talks with Anthropic to rent out compute capacity from its data centers. The article Zuckerberg's plan to sell excess AI compute could finds its first big customer in Anthropic appeared first on The Decoder.
Read original →OpenAI's GPT-5.6 has accidentally wiped users' entire home directories in several cases, mostly in the unprotected "Full Access Mode." The model overwrites a temporary directory variable and carries out destructive actions on its own instead of asking for confirmation. OpenAI has announced extra safeguards and a detailed post-mortem. The article GPT-5.6 is deleting user files when given full access, and OpenAI says it shouldn't but did appeared first on The Decoder.
Read original →Moonshot AI's Kimi K3 model, built by a 300-person team, matches Anthropic's Opus 4.8 performance and is reigniting debate over whether computational efficiency or raw computing power determines AI capability. This challenges Western AI labs' historical competitive advantage and may reshape investment priorities toward efficiency-focused architectures over brute-force scaling.
Read original →Bunkerhill Health has raised $55 million to scale its agentic AI platform, Carebricks. The closing of the company’s Series B round, announced today, folds in continued participation from Sequoia Capital, Felicis, Optum Ventures, and Y Combinator. However, a funding total doesn’t answer the key question any hospital executive wants to know about healthcare AI: does […] The post Bunkerhill raises $55M to scale agentic AI across health systems appeared first on AI News.
Read original →Some of my favorite new features in iOS 27 are flying under the radar, but no less impactful.
Read original →