AI daily

2026-08-12

2026-08-12 AI Daily

Featured

SSignal87
NVIDIA Developer Blog

Alibaba Releases Qwen3.8-2.4T-A95B Open-Weight Model with Configurable Reasoning

Alibaba has released the open weights for Qwen3.8-2.4T-A95B (Qwen3.8-Max), its largest open-weight model, featuring 2.4 trillion total parameters with 95 billion activated per token. The model uses a fine-grained mixture-of-experts architecture with hybrid attention and supports configurable reasoning, deployable on NVIDIA GB300 NVL72, bringing near-frontier capabilities to the open ecosystem.

APriority78
SpaceXAI (X)

xAI Releases Grok 4.6 with Major Performance Boost at Same Price

xAI has officially released Grok 4.6, a frontier model that delivers significant improvements over Grok 4.5 while keeping the same price. Priced at $2 per million input tokens and $6 per million output tokens, it is half the cost of other frontier models. Grok 4.6 is now available on Grok Build, Cursor, Grok Bot, and the API, with double usage offered for the first week on Cursor and Grok Build.

SSignal89
AI前线

Major AI Paper: Distillation Defenses of Top Three Global Models Fully Broken—Small Models Extract Hidden Reasoning Chains from Large Models, Kimi-K3 Shows Anomalous Reproduction Probability

A new study reveals severe security vulnerabilities in the APIs of Anthropic, OpenAI, and Google, allowing attackers to use lightweight models as "decoders" to recover hidden reasoning chains from encrypted reasoning blocks. Led by MATS program researcher Alexander Panfilov and co-authored by the Max Planck Institute for Intelligent Systems and others, the paper has garnered over 2.1 million reads. Additionally, the study found that Kimi-K3 reproduces Claude/GPT reasoning fragments with anomalously high probability, though the authors stress this is not sufficient to prove distillation.

SSignal89
DeepTech深科技

Researchers Find Easy Way to Extract Encrypted Chain-of-Thought from OpenAI and Anthropic

Researchers have discovered a method to extract encrypted chain-of-thought reasoning from the APIs of Anthropic, OpenAI, and Google. By exploiting a vulnerability where encrypted reasoning blocks can flow across sessions and models, attackers can have weaker models decode and output hidden reasoning from flagship models. The study was released on August 10 by MATS Research, the University of Tübingen, and the Max Planck Institute for Intelligent Systems.

APriority76
TestingCatalog (X)

Claude in Chrome sessions now powered by same solution as Claude Cowork

Anthropic announced that Claude sessions in the Chrome browser are now powered by the same solution as Claude Cowork and are synced across different clients. The feature supports Skills and Connectors, and is available to Max and Team users today, with Pro rollout in the coming weeks.

SSignal86
The AI Insider

Spotify to Label AI-Generated Artists With 'AI Persona' Tags, Exclude Them From Recommendations

Spotify announced it will begin labeling AI-generated artists with 'AI Persona' badges starting in mid-September and will exclude such artists from its editorial and algorithmic recommendations. The policy aims to distinguish AI-generated artist identities from human artists, while not affecting the labeling of AI-generated music itself. This move is Spotify's latest step in the AI music space, balancing innovation and authenticity.

SSignal89
机器之心

116-Page Paper Reveals LLM API Flaw: Hidden Reasoning Traces Can Be Fully Extracted

A paper reveals a design flaw in frontier LLM APIs that allows attackers to extract fully encrypted hidden reasoning chains. The vulnerability affects major vendors like Anthropic, OpenAI, and Google, exploiting security weaknesses in weaker models within the same family to achieve cross-model reasoning extraction. The study verified extraction accuracy via API billing token counts, sparking widespread community attention.

SSignal87
量子位

DeepSeek V4 Pro Officially Released, Benchmarking Against Fable 5

DeepSeek has released the official version of V4 Pro, model ID DeepSeek-V4-Pro-0813, which performs strongly on multiple benchmarks, approaching or even surpassing Fable 5. API pricing remains low, but the chain-of-thought output has become terse, sparking discussion among users.

APriority76
Google DeepMind (X)

Google unveils SL2T sign language-to-text model for Android

Google has unveiled SL2T, a breakthrough sign language-to-text model that powers new features for Deaf and hard of hearing users on Android. Starting with American Sign Language-to-English on Pixel 11, users can sign directly into Gboard and Live Transcribe instead of typing. The model tracks body poses on-device for privacy and was developed in collaboration with the Deaf community.

SSignal86
机器之心

Meitu Imaging Research Institute Publishes 11 Papers at Top Conferences, Advancing AI Image Editing

Meitu Imaging Research Institute (MT Lab) had 11 papers accepted at top international conferences in 2026, covering video editing, portrait editing, and more. They introduced innovative technologies like MiVE and CFT to improve consistency and controllability in AI image editing. These technologies have been integrated into products like Meitu Xiuxiu, pushing AI from 'generation' to 'creation'.

More

SSignal86
量子位

Zidong Taichu Introduces GMC Core-Set Pruning: 80% Fewer Tokens, Full Multimodal Fidelity

The Zidong Taichu team at the Chinese Academy of Sciences has proposed GMC, a core-set pruning method that reduces visual tokens by over 80% while preserving multimodal performance, addressing the bottleneck of token redundancy. The training-free method works with mainstream models like Qwen and LLaVA, matching or exceeding original performance on multiple benchmarks.

SSignal86
InfoQ

Microsoft Officially Releases Agent Framework Harness and Hosted Agents

Microsoft has officially released the Agent Framework Harness and Foundry Hosted Agents, marking the framework's transition from SDK to a supported production runtime. The Harness, a single binary, runs across local, container, and hosted environments, with built-in features like function calling, persistence, context compression, tool approval, and OpenTelemetry. This release addresses the runtime question and highlights the Harness's importance in agent systems.

APriority76
阿里云开发者

Qwen AI Arena Launches: A Real Battlefield for Agents

On August 12, Alibaba's Qwen AI platform officially launched Qwen AI Arena, a challenge and evaluation platform for AI agents, designed to generate tasks from real business scenarios and validate agents' ability to solve practical problems. The first task focuses on cross-border e-commerce product listing, requiring developers to build agents that generate multilingual and multimodal listing materials, with top solutions receiving token rewards and potential entry into Alibaba Cloud's ecosystem.

APriority76
Hacker News (AI filter)

Launch HN: Discovered Materials (YC P26) - AI agents to discover new materials

Discovered Materials, a YC P26 startup, launched on Hacker News with AI agents that discover new materials for the semiconductor industry. The founders highlight the heat problem in GPUs and the potential of AI to reduce the time and cost of introducing new materials into chips.

APriority74
TestingCatalog (X)

NoimosAI Releases Version 2.0 of Its SEO Agent

NoimosAI has released version 2.0 of its SEO Agent, which can conduct SEO research and retrieve data from tools like Semrush and Google Search Console. After user approval, it can apply changes directly to websites, creating a continuous SEO loop.

SSignal87
Latent Space

How to Steal a Reasoning Trace

A new paper demonstrates how to decode and port encrypted reasoning traces from frontier reasoning models, significantly boosting open model performance, while warning that publicly shared sessions can leak sensitive data. The technique replays encrypted reasoning blocks into weaker models to induce transcription, revealing major security and alignment concerns.

SSignal86
机器之心

ByteDance Seed Team Uncovers New Scaling Variable for Text-to-Image: Caption Information, Not Length

ByteDance Seed team found that the training performance of text-to-image models depends not on caption length but on the amount of image-bound information within. They proposed Structured Prompt, which improves both diffusability and promptability, achieving significant gains on complex composition, reasoning, and world knowledge generation tasks. This research opens a new direction for scaling text-to-image models.

APriority85
THE DECODER

xAI's Grok 4.6 Matches OpenAI's Best Model and Undercuts It on Price

xAI's Grok 4.6 scores 61 points on the Artificial Analysis Intelligence Index, tying OpenAI's GPT-5.6 Sol and trailing only Anthropic's Claude Opus 5. On agentic tasks, it completes complex workflows in about 53 steps where Claude Opus 5 needs 103, at a price more than 60 percent lower.

APriority72
TestingCatalog (X)

Gemini Notebooks Now Supports Copying

Google announced that Gemini Notebooks now supports copying, allowing users to duplicate shared notebooks as their own, including all sources and artifacts. This feature aims to help students personalize study materials and teammates quickly start new projects.

APriority72
Claude (X)

Claude in Chrome sessions now sync across devices

Anthropic announced that Claude sessions in Chrome now carry over to desktop, web, and mobile, with conversations saved and skills/connectors working in the browser. The feature is available today for Max and Team users, with Pro users getting it in the coming weeks.

APriority83
NVIDIA AI Blog

NVIDIA AI Factory Compute Is Becoming an Investable Asset Class

NVIDIA announced partnerships with major financial institutions including Apollo, BlackRock, Blackstone, Brookfield, Goldman Sachs, and KKR to establish independent financing platforms aimed at mobilizing over $500 billion for AI infrastructure. This marks a shift from project-based data center construction to financing AI factories as productive infrastructure assets, highlighting the growing investment appeal of AI compute.

APriority79
NVIDIA AI (X)

NVIDIA Launches Nemotron 3.5 Lightning, Partners Showcase Domain-Specific Post-Training

NVIDIA has launched Nemotron 3.5 Lightning and showcased how partners are post-training it for their specific domains. Partners including Prime Intellect, CrowdStrike, Dream, and CodeRabbit are leveraging the model to enhance performance in areas like cybersecurity and code review.

APriority72
Google Gemini (X)

Gemini now connects to 14 new apps including OpenTable, Ticketmaster, and more

Google announced that Gemini can now connect to more third-party apps, including OpenTable, Ticketmaster, Wix, and 11 others. Users can use natural language commands to book reservations, build websites, stream music, and more directly within Gemini. This move aims to expand Gemini's ecosystem and enhance its utility as an AI assistant.

SSignal86
量子位

Fine-tune with 0.01% parameters: BEFT method accepted at ACL 2026

Researchers from Lund University and Google DeepMind propose BEFT, a bias-only fine-tuning method that trains only the value bias b(v), achieving near full fine-tuning performance with just 0.01% of parameters in low-data settings. The work has been accepted at ACL 2026 and integrated into the Hugging Face PEFT library, offering an efficient solution for resource-constrained fine-tuning.

APriority81
THE DECODER

Researchers can now reverse-engineer LLM prompts from output text with near-perfect accuracy

Researchers at IIT Bombay and Adobe Research have developed an inverse language model that reconstructs the original prompt from an LLM's output with near-perfect accuracy. The method, called 'Previous-Token Prediction,' does not require access to model weights and works across different models. This poses a serious security risk for companies relying on proprietary system prompts.

APriority79
钛媒体AGI

Belle Fashion Shares Enterprise AI Adoption Practice: From Collaborative Online to AI-Native

Ji Yanli, General Manager of the Technology Center at Belle Fashion Group, detailed a three-step strategy for enterprise AI adoption, emphasizing the importance of business process digitalization, standardized decision-making, and closed-loop task management. The article argues that successful AI implementation starts with collaboration, gradually building AI-native capabilities that integrate AI into business operations.

APriority76
Artificial Analysis (X)

Grok 4.6 Shows Big Gains on AA-Briefcase Benchmark

Grok 4.6 achieved significant improvements on the AA-Briefcase benchmark, which tests long-horizon agentic knowledge work tasks. The model performs on par with Claude Fable 5 but at a much lower cost, at $4.42 per task compared to Claude Fable 5's $22.30. This release highlights xAI's competitive edge in both performance and cost efficiency.

APriority81
DeepTech深科技

UCSD Professor Xie Pengtao's Startup: AI for Science Delivers Top-Journal-Level Results in 4 Hours, Secures Millions in Funding

Xie Pengtao, an associate professor at UC San Diego, and his team have launched AIBuildAI Science Agent, an AI for Science agent that ranked first on NatureBench, achieving or exceeding Nature paper results on 57.8% of tasks and surpassing human best methods on 25.6%. The agent can independently complete scientific research tasks within 4 hours, and the company has raised approximately $4 million.

APriority81
DeepTech深科技

Berkeley Open-Source Project SkyPilot Spins Off into Startup with $20M Seed Funding

SkyPilot, an open-source AI compute orchestration project from UC Berkeley, has spun off into a startup and raised $20 million in seed funding. The project addresses the fragmentation of GPU management across clouds and has gained wide adoption. Co-founders Zongheng Yang and Zhanghao Wu stated that commercialization will focus on organization-scale capabilities while keeping individual use open-source.

APriority81
InfoQ AI/ML/Data Eng

MCP Goes Stateless, and Developers Ask Whether That Just Makes It an API Again

The MCP 2026-07-28 specification removes the initialize handshake and session header, adding required method and tool-name headers for stateless routing. Developers are split on whether this makes MCP just an API again or if the standard's value remains.

APriority78
AI Business

Microsoft Cuts Prices for Coding Model to Stay Competitive

Microsoft has announced price cuts for its coding model to remain competitive. The upgraded model is now faster and more token-efficient, enabling users to complete tasks more quickly and at lower cost.

APriority81
InfoQ

Tencent Cloud Open-Sources Cube Sandbox: An Execution Environment for AI Agents

Tencent Cloud has officially open-sourced Cube Sandbox, an execution environment base for AI agents, evolved from its Serverless infrastructure. The article, through an interview with a Tencent Cloud expert, explores the new infrastructure demands of agents and how Cube provides an isolated, recoverable runtime.

APriority78
极客公园

Hands-on with GenOffice Alpha: AI-native office suite shows promise in formatting but lags in basics

GeekPark tested Genspark's AI-integrated office client, GenOffice Alpha. The product leverages AI to directly manipulate Office file formats, converting generated content into editable Word and PPT documents. While basic editing features still suffer from stability issues, its AI-powered formatting capabilities impressed, automatically recognizing document structures and rebuilding layouts. Released as open-source and free, the Alpha was built by one engineer in a week and will iterate rapidly based on user feedback.

APriority78
DeepTech深科技

xAI Co-founder's River AI Raises $1.1B to Push Open-Source and Personal AI

On August 11, River AI, founded by xAI co-founder Igor Babuschkin, announced a $1.1 billion funding round led by General Catalyst and AMP PBC, with participation from NVIDIA, AMD Ventures, and others. The company aims to advance open-source models and personal AI, enabling businesses and individuals to own their AI, and has launched its first product, River API.

APriority79
DeepTech深科技

Fudan Young Scientist Solves Century-Old Nuclear Material Problem, Paving Way for 'Artificial Sun'

Zhang Hongliang, a young researcher at Fudan University, has made multiple breakthroughs in nuclear materials, including discovering radiation-induced segregation in ceramics, revealing a new deformation mechanism in intermetallic compounds, and developing a process that reduces corrosion rates in nuclear fuel cladding by over 40%. His work aims to address critical safety and performance issues in nuclear energy, providing material solutions for advanced systems like fusion reactors.

APriority76
InfoQ AI/ML/Data Eng

Spotify Builds External Index to Enable Low Latency Point Queries on Its Data Lake

Spotify introduced an external indexing architecture for Apache Parquet data lakes that enables low-latency point queries without replicating datasets into operational databases. The approach maps lookup keys to Parquet files and row locations, allowing targeted reads from cloud object storage while supporting analytics, machine learning, AI applications, and online services from the same datasets.

APriority76
少数派

Quote/0 E-ink Display Update: Open APIs and AI Commands, Plus Duo Interaction

The Quote/0 e-ink display has received a major update, allowing users to curate their information feed via the app and customize content using text, images, RSS, and more. The developer platform now offers text, image, and canvas APIs, and supports natural language commands through Dot Skill to let AI generate display content. Additionally, the duo bundle enables two devices to share to-do lists and send images to each other, enhancing interactivity.

APriority76
meng shao (X)

SpaceXAI Launches 'New AI Teammate': Grok Bot

SpaceXAI has introduced Grok Bot, a new AI teammate that can log into users' tools and websites, operate software interfaces like a human, and deliver finished work directly. Grok Bot supports asynchronous 24/7 operation, parallel scaling, and can learn workflows through demonstration for automation. The product is currently in early beta, aiming to change how users interact with AI assistants.

APriority74
The AI Insider

OpenAI COO Brad Lightcap Departs; Company Launches ChatGPT Desktop App for Linux

Brad Lightcap, OpenAI's COO, announced his departure to pursue a new venture, marking the latest in a series of executive exits as the company prepares for an IPO. Meanwhile, OpenAI released a preview of its ChatGPT desktop app for Linux, extending support to major distributions and completing coverage across all major desktop operating systems.

APriority74
meng shao (X)

Manus Resumes Independent Operations, Some Users Need to Back Up Data

Manus announced it will resume operating as an independent company and, due to regulatory requirements, some users will need to back up their data by a specified deadline. The company will provide a restoration portal and welcome-back bonuses, and teased a series of new features.

APriority76
The AI Insider

Anthropic's Unreleased AI Model Makes Major Progress on 150-Year-Old Riemann Hypothesis

Anthropic announced that an unreleased AI model made significant progress on the Riemann hypothesis, a math problem unsolved for over 150 years. The model autonomously worked for about a day and a half, testing 650 approaches and coordinating 60 subagents, with two subagents making the key breakthroughs. The findings were formalized using the Lean proof assistant and have sparked debate within the mathematical community about AI's role in mathematics.

APriority78
DeepTech深科技

AI Scientists Begin Redistributing Venture Capital

AI startups like Thinking Machines Lab and Periodic Labs are launching external research grant programs, directly providing funding and compute to academic researchers. This reflects venture capital's trust in AI scientists and a new trend of AI companies reallocating research resources.

APriority74
THE DECODER

Mistral now offers EU data processing and priority access, but both come with important limits

Mistral is giving customers the option to route AI requests through servers in either Europe or the US, and selling priority queue access during peak traffic. Both come with a surcharge, and the regional routing doesn't cover all features or data. This move aims to address European data processing needs but has limitations.

APriority72
腾讯技术工程

WorkBuddy Product Manager Publishes 30,000-Word Comprehensive Guide

WorkBuddy's product manager Kong Deyuan has released a comprehensive 30,000-word guide covering everything from installation to advanced usage, aiming to help users fully master this AI office tool. The guide details three working modes (Ask, Craft, Plan) and emphasizes WorkBuddy's capability as an AI agent to directly manipulate files, generate reports, and more.

APriority76
量子位

Embodied AI data startup Yuan Dian Tech raises tens of millions in funding to build physical AI infrastructure

Yuan Dian Tech (SCALEFORCE), an embodied intelligence data infrastructure company, has completed a funding round of tens of millions of RMB, with investors including top domestic embodied intelligence industry players, Hengxu Capital, and Kailian Capital. The funds will be used for the iteration of its MatrixOS technology, construction of large-scale data production networks, and team expansion. The company aims to address data quality, scale, and generalization challenges in embodied AI, accelerating the development of physical AI.

APriority76
AI前线

As AI inference scales up, Huawei redefines the role of storage

At a data storage user forum in Wuxi, Huawei proposed that as AI inference scales, storage must evolve from traditional persistence devices into the inference path, and introduced the industry's first context memory storage (CMS) for large-scale inference, a G3.5 storage tier. Huawei emphasized a full-stack approach to integrate data, compute, models, and agents, aiming to advance AI adoption.

APriority76
InfoQ

DeepSeek + Pi Combo Beats Claude Code? Pi Founder Says He Called It

A public benchmark by Composio shows that Pi Harness with DeepSeek V4 Flash achieved the highest success rate (66.7%) among 8 agent frameworks, at a cost seven times lower than Claude Code. Earlier, developer 0xEvan processed nearly 1 billion tokens with Pi and DeepSeek V4 Flash, hitting a 99.93% cache hit rate for just $2.65. Pi founder Mario Zechner noted he had endorsed this combination back in May.

APriority74
MIT Technology Review AI

Scaling AI agents with trustworthy data

A new report from MIT Technology Review highlights that while enterprises are rapidly adopting agentic AI, legacy data systems are major blockers to realizing ROI. Based on a survey of 300 data and technology executives, it finds that only a few 'data leaders' provide ample data access to AI agents, leading to better outcomes. The report urges organizations to modernize data infrastructure to enable agents to make real-time decisions and take actions.

APriority74
The AI Insider

AI Competitive Intelligence: How to Track Rival Moves

This article explores how businesses can track rivals' model releases, pricing changes, and product updates for competitive intelligence in the AI era. It cites the Institute of AI Product Management and Tiger Tail, highlighting AI tools' ability to catch patterns humans miss, and argues that competitive intelligence is now a core competency in AI-driven business strategy.

APriority74
The AI Insider

AI Enterprise Startup June Emerges From Stealth With $20M to Simplify Agent Implementation

June, an AI enterprise platform founded by former Salesforce executives, has come out of stealth with $20 million in pre-seed funding led by Marc Benioff's Time Ventures. The company aims to help businesses implement AI agents more easily by scanning existing systems and building optimized workflows. The funding round also includes investments from Michael Dell, Aaron Levie, and George Kurtz.

APriority76
量子位

FEAGINE Launches Flexible Robots and Cross-Embodiment Foundation Model, Exploring New Path in Embodied AI

FEAGINE Technology unveiled three bionic flexible robots, FEAGINE A01, A02, and A03, along with the first-generation cross-embodiment foundation model Fi0, aiming to advance embodied intelligence through redesigned bodies. The company argues that generality should come from shared intelligence rather than a unified humanoid form, and has achieved standardized mass production of tendon-driven flexible bodies.

APriority76
机器之心

Westlake Robotics Raises 500 Million Yuan in Series A, Totaling 500 Million in Six Months

Westlake Robotics, a general brain company for embodied intelligence, has completed a Series A funding round, bringing total funding to 500 million yuan within six months. Investors include Sequoia, Xiaomiao Langcheng, and others. The funds will be used for R&D of a unified large model for humanoid robots and talent training. The company has partnered with Alibaba, Intel, JD.com, and HKUST, and secured orders worth nearly 100 million yuan.

APriority74
量子位

OpenAI loses four executives in a month, safety team nearly gutted

OpenAI has seen a wave of executive departures, including former COO Brad Lightcap, safety systems head Johannes Heidecke, chief futurist Joshua Achiam, and ethics lead Chloé Bakalar. The exits come just before a potential IPO, raising concerns about the company's stability and commitment to safety. Analysts suggest the departures reflect internal tensions between safety and commercialization.

APriority74
THE DECODER

Microsoft's New MAI Code 1.1 Flash Gets Crushed by DeepSeek on Both Price and Performance

Microsoft has released MAI Code 1.1 Flash, a code model for GitHub Copilot that is claimed to be 25% more token-efficient at a quarter of the cost of its predecessor. However, in benchmarks, it is outperformed by the cheaper DeepSeek V4 Flash on both price and performance. This move fits a pattern where Microsoft promotes open AI while integrating inferior proprietary models into its apps to protect margins.

APriority72
AI Business

IBM Signs $240M Deal for Nvidia-Powered AI Cluster

IBM has signed a $240 million deal with Together AI to provide an Nvidia-powered AI cluster, responding to the growing demand for open models. The partnership highlights the increasing investment in AI infrastructure.

APriority74
量子位

Jeff Dean's Final 48 Hours at Google Revealed: 1,500-Person Farewell, Late-Night Replies

Google legend Jeff Dean detailed his last 48 hours at the company during KDD 2026, including announcing his departure to 139 close colleagues, attending a farewell event with 1,500 people, and officially leaving on August 6. He then became co-founder and CEO of Discovery Loop, a new venture with Sanjay Ghemawat and others to explore AI-accelerated research. This marks a significant shift in AI leadership and signals new directions for research.

APriority72
THE DECODER

OpenAI launches ChatGPT desktop app for Linux

OpenAI has released a Linux version of its ChatGPT desktop application, expanding the platform coverage of its desktop client. This move allows Linux users to access ChatGPT more conveniently without relying on a web browser.

APriority72
量子位

Former Qwen Lead Lin Junyang Launches AI Agent Startup Pragmatik Labs, Backed by Tencent

Lin Junyang, former head of Alibaba's Qwen large model series, has officially announced his new startup Pragmatik Labs, which focuses on building next-generation AI agents that span digital and physical worlds. The company has already secured funding from investors including Gaorong Ventures, Sequoia China, and Tencent, with a valuation of $2 billion. Lin aims to shift AI from reasoning-centric to agentic thinking, emphasizing productivity and long-term scientific progress.

APriority72
AI前线

Ex-Qwen Tech Lead Launches Startup Valued at $2B

Lin Junyang, former head of Alibaba's Qwen technology, has officially launched his startup Pragmatik Labs, valued at $2 billion. The company focuses on next-generation agents spanning digital and physical worlds, with backing from GSR Ventures and Sequoia Capital China. This move signals a shift from foundation models to agentic AI, though it faces challenges in hardware-software integration and long development cycles.

Selected
62
Sources
28
Featured
12
More
50
All reports