Google Allows Watermark Removal on AI Content; OpenAI Disbands Safety Team; NVIDIA Secures AI Factory with SB Energy
Major developments in AI include Google's new watermark removal feature, OpenAI's safety team dissolution, NVIDIA's infrastructure partnership, and breakthroughs in world models and open-source models.
Selected
45
Sources
20
Featured
12
More
33
AI DAILY BRIEFING
The AI industry is rapidly evolving with significant moves from major players. Google announced users can disable visible watermarks on AI-generated content while retaining invisible tracking, and open-sourced Credentio. OpenAI disbanded its Preparedness safety team, raising concerns about safety oversight. NVIDIA partnered with SB Energy to secure infrastructure for AI factories, with OpenAI as tenant. Meanwhile, HiDream.ai released a native all-modal interactive world model, and Prime Intellect showed open-source models can approach top-tier closed-source performance. Alibaba's Qwen3.8-27B gained massive adoption, and Alaya Lab open-sourced a world model generating hour-long videos.
Google's watermark removal feature balances user control with content provenance, potentially impacting content authenticity debates.
OpenAI's disbanding of its Preparedness team signals a shift in safety strategy, but may increase internal and external scrutiny.
NVIDIA's infrastructure partnership highlights compute as a critical bottleneck for frontier AI development.
Watch next: Monitor the impact of OpenAI's safety team dissolution on AI governance, and the adoption of open-source models like Qwen3.8-27B in enterprise settings.
Google announced it will allow users to remove visible watermarks from AI-generated images, videos, and songs, while keeping invisible tracking methods intact. The feature will be available in Gemini, Flow, and soon Search. Google also open-sourced Credentio, a library for verifying content credentials.
OpenAI has disbanded its Preparedness team, which was responsible for assessing catastrophic risks from frontier AI, stating that safety work has been integrated into the development process. The team's lead, Dylan Scandinaro, has departed to research safety risks of recursively self-improving AI. This follows the dissolution of several other safety teams and high-profile departures, fueling internal unease.
NVIDIA announced a partnership with SB Energy to secure LPS capacity at the PORTS-Pike Technology Campus in Portsmouth, Ohio, to host NVIDIA compute, with OpenAI as the tenant. This move aims to secure critical infrastructure for AI factories, as frontier AI labs' growth is increasingly constrained by compute availability rather than algorithms or customer demand.
HiDream.ai today unveiled the world's first native all-modal interactive world model, HiDream-O1-World, built on its proprietary UiT architecture. The model supports multimodal inputs including text, images, and interactions, and can generate dynamic worlds with long-term spatiotemporal and physical consistency. In the WBench benchmark jointly launched by Meituan LongCat and Fudan University, HiDream-O1-World topped the Navi sub-leaderboard on its first attempt, scoring 73.3 in physical dimensions and 88.0 in consistency, ranking first overall. This launch marks a shift in AI content generation from one-way output to explorable and interactive world models.
Prime Intellect introduced a multi-agent harness that enables open-source models to approach top-tier closed-source models in a nanoGPT optimizer speedrun. In experiments, Kimi K3 with the harness achieved 2930 steps, surpassing GPT-5.6 Sol and coming close to Opus 5, showing that open-source models can catch up through efficient research processes.
Qwen3.8-27B, an open-source 27B-parameter multimodal model from Alibaba, outperforms Claude Opus 4.6 Max on coding and agentic tasks. Within two days of release, it surpassed one million downloads and received rapid support from hardware vendors like NVIDIA and AMD, as well as inference tools. This highlights the growing demand for capable local AI models and the power of open-source ecosystems.
Alaya Lab, together with USTC and Shanghai Innovation Institute, released Alaya-EVOKE, a video world model that can generate 120 minutes of continuous video (4800 clips, 172,800 frames) with only 3 denoising steps. By externalizing memory into a world state bank and redesigning the teacher model for long-horizon supervision, it overcomes color drift, content drift, and VRAM limits. It achieves top scores on WBench navigation subset and VBench-2.0.
Alibaba has officially launched the Qwen-Audio-3.0 series of voice models on its Qianwen AI platform, including ASR, TTS, and real-time interaction variants. The models achieved top rankings in three categories on Artificial Analysis's voice leaderboard, and developers can access them via API or Token Plan subscriptions. They are already integrated into products like the Qianwen App.
This article details the security center team's practice of rebuilding a knowledge base in the AI era, emphasizing that when agents become primary knowledge consumers, the knowledge supply system must achieve structured, automated, and dynamic elimination. The author presents a complete methodology from goal definition to flywheel operation, sharing key decisions and reusable methods.
At COLM 2026, multiple papers simultaneously question Transformer's default architectural choices, such as QK normalization, GQA, and sliding window attention, finding that these designs underperform in long-context tasks and that combining them exacerbates performance drops. Another study shows that the output projection layer loses most gradient information during backpropagation, hindering training efficiency. These findings may prompt a reevaluation of seemingly reasonable default configurations in future model designs.
OpenAI has signed a 20-year lease for an 8-gigawatt data center in Ohio, with Nvidia guaranteeing up to $105 billion for the residual value of the facilities and becoming the exclusive chip supplier. The deal underscores the massive scale of AI infrastructure investments, as nine tech companies hold around $3 trillion in AI commitments that are not on balance sheets.
Charlotte Bunne's team at EPFL, along with Stanford, Google Research, and Harvard, has introduced VirTues, a foundation model for spatial proteomics that creates virtual tissue representations. Published in Nature, the model overcomes data silos and demonstrates strong generalization across cohorts and platforms, with potential applications in triple-negative breast cancer.
NVIDIA has released the Nemotron 3.5 Lightning NVFP4 checkpoint, developed with the NVIDIA Model Optimizer, which preserves accuracy while delivering up to 4x faster throughput and compressing the model from 66 GB to 22 GB. This advancement aims to help developers deploy large language models more efficiently.
A new report from energy research firm Noreva warns that natural gas prices could triple in some U.S. regions in the coming years, driven by demand from AI data centers, slowing supply growth, and rising LNG exports. Major hyperscalers like Meta, Microsoft, Google, and Amazon have committed to building massive natural gas power plants in Texas and Louisiana to support their AI operations. Noreva projects prices could exceed $10 per million BTUs in certain areas, up from the current $2-$4.50 range, potentially raising operating costs for data centers and impacting electricity prices and token pricing.
According to Bloomberg, Stripe is acquiring AI startup OpenRouter for more than $7 billion, a significant premium over its latest valuation of $1.3 billion. OpenRouter provides access to over 400 AI models and has eight million users, with its CEO previously describing the company as "Stripe for AI." The acquisition would mark a major move by Stripe into the AI infrastructure space.
Peking University, together with StepFun and Beijing University of Posts and Telecommunications, introduced TensorCast, a unified programmable tensor lifecycle management abstraction layer for large language model infrastructure. It decouples tensor state management from compute logic, enabling flexible cross-component strategies while matching specialized system performance, achieving up to 93.2% reduction in time-to-first-token for high-concurrency multi-turn agent scenarios.
Current Robotics has unveiled CurrentWorld-0, the world's first physical world model that integrates cross-embodiment, multi-view, and force-tactile prediction. It learns physical laws from real-world data, enabling simulation, learning, and evaluation for various robot types. This model addresses the limitations of traditional physics engines in modeling soft objects and complex contacts, offering a scalable testing environment for robotics.
A team from the University of Hong Kong and Super Dynamic has developed the SMASH system, enabling two humanoid robots to autonomously play a complete table tennis match under the 11-point rule. The system integrates visual perception, trajectory prediction, motion planning, and whole-body control, and has been upgraded to version 2.0 to support a wider range of shots and autonomous serving. The achievement will be showcased at the second World Humanoid Robot Games, where the robots will also compete against famous table tennis players.
In 2026, the embodied intelligence sector is shifting from a debate between VLA and world models to a focus on practical deployment and route hybridization. At NVIDIA GTC, players like Geely, Momenta, and Huawei publicly questioned VLA's limitations, while at WAIC, vendors showcased 'world model + VLA' fusion solutions. Ant Lingbo and Face Intelligence released new VLA models emphasizing versatility and on-device deployment, with industry consensus leaning toward integration rather than replacement.
Speko (YC S26) has launched a platform that routes voice AI requests to the best combination of speech-to-text, LLM, and text-to-speech models based on user constraints like accuracy, latency, and cost. It aims to solve the problem of voice agents using outdated models due to the complexity of switching vendors. The platform provides an API and dashboard for easy model switching.
WindBorne Systems, a startup that uses AI-powered weather forecasting with long-flying balloons, has raised $37 million in Series B funding at a $250 million valuation. The round was co-led by Khosla Ventures and Galvanize, with participation from other investors. The company plans to use the funds to expand its commercial business, particularly targeting investment funds that use weather data for market predictions.
Anthropic CEO Dario Amodei pushed back against claims that his warnings about AI risks have fueled public backlash, arguing that distrust stems from broader societal issues. The company also clarified details about its Claude text watermarking system, introduced to comply with the EU AI Act, including its use of SynthID-Text and plans for a detection API.
SpaceX has officially closed its $60 billion acquisition of AI coding startup Cursor, as announced on Cursor's blog. The deal, first agreed in April, aims to leverage SpaceX's computing infrastructure and GPU fleet to scale advanced AI applications. Cursor joins Musk's portfolio of AI assets, following the earlier acquisition of xAI.
The AI video market has rebounded from the initial setback of Sora, with AI production companies like Promise setting up near Hollywood's historic studios to cut film costs using real-time backgrounds and other AI tools. Netflix already uses AI in 300 of its 1,000 titles, and startup Higgsfield now has a $5.4 billion valuation.
This issue of Import AI introduces DiG-bench, a new benchmark of 70 games designed to assess AI systems' ability to discover environmental rules through exploration and curiosity. Developed by researchers from Oxford, Princeton, and other institutions, it aims to measure AI's intuition and creativity in novel environments.
Serve Robotics announced it is adding Grubhub to its autonomous delivery network, making robot delivery available from over 100 merchants in Chicago and nearly 200 in Los Angeles, plus service in Alexandria, Va. The company also expanded to Washington, D.C., and San Jose, Calif., through its DoorDash partnership, and introduced new products like Moxi 2.0 and an AI advertising tool. This comes as Uber Eats sold its remaining shares in Serve, with their partnership set to end next year.
Gravis Robotics, an ETH Zurich spinout, has raised $200 million in Series A funding from SoftBank to expand its physical-AI platform for autonomous heavy construction equipment. The funding will support hiring, global expansion, and additional deployments of its hardware-agnostic control system, which can retrofit machinery from major manufacturers. This investment underscores the growing interest in automating construction, an industry where automation has historically been challenging to deploy.
Stripe has reportedly reached a deal to acquire AI model marketplace OpenRouter for more than $7 billion, according to Bloomberg. OpenRouter lets customers switch between hundreds of AI models to avoid vendor lock-in. The acquisition follows OpenRouter's $113 million Series B at a $1.3 billion valuation and its growth to 8 million users.
Qingmang AI, the first Chinese startup focused on self-evolving AI infrastructure, aims to use AI to optimize AI and address efficiency gaps in domestic chips. Rooted in research from Tsinghua's Professor Zhu Wenwu, the company targets providing trillion-token-scale computing power for the Agent era.
OpenAI has an internal email address, friction@openai.com, where employees can report various issues, which are then triaged by leadership and fast-tracked for resolution, with significant problems reaching CEO Sam Altman. The mechanism was institutionalized by apps CEO Fidji Simo to cut bureaucracy and maintain agility as the company scales.
Bun's creator Jarred Sumner used Claude AI agents to rewrite Bun from Zig to Rust in 11 days, costing about $165,000. Zig creator Andrew Kelley criticized the project's coding practices, citing a lack of quality control. The event sparks debate on AI-assisted large-scale code rewrites.
Noiz AI, in collaboration with researchers from HKUST, Tsinghua, CMU, and Google DeepMind, has released HelixWorld 1.0, the first real-time interactive audio-visual world model that generates synchronized visuals and audio at 24FPS and 48kHz stereo. The model uses a native Transformer architecture to unify audio and video generation from the start, ensuring that user actions affect both modalities simultaneously. The team plans to open-source the model weights and code in the coming weeks, which could accelerate progress in interactive world models.
CodeRabbit, an AI-powered code review platform, has secured $143 million in funding, valuing the company at $1.5 billion. The round was co-led by Atomico and Smash Capital, with participation from BMW i Ventures, Datadog, and Hirtle Callaghan. The investment underscores growing confidence in AI coding tools and the need for oversight of AI-generated code. The company plans to expand internationally and invest in free tools for open-source projects.
Symbiosis Robotics released a new demo showing a humanoid robot driving a go-kart, and introduced Direct Perception Control (DPC), which eliminates intermediate motion representations and directly generates joint actions from visual and task information, aiming to bridge the gap between perception and whole-body control in hierarchical systems. The team compiled 15,010 hours of embodied data for training.
AI and data centers have become prominent topics in US campaigns, appearing in nearly 40% of races, ahead of Israel, racism, and manufacturing. The impact of data centers on electricity costs and local resources drives most of the conversation.
Alibaba is selling its gaming subsidiary Lingxi Interactive to CITIC Capital, reportedly for over $1.5 billion. This move is part of Alibaba's strategy to focus on AI and AGI by divesting non-core assets. Lingxi's CEO Zhou Bingshu confirmed the deal in an internal letter, and the existing management team will continue to operate the company.
openJiuwen has upgraded its swarm agents to WorkSwarm, debuting on the HarmonyOS PC app market and supporting multi-agent collaborative office work. WorkSwarm organizes multiple agents to work in division of labor through coordination engineering, enabling complex tasks like music creation, document writing, and programming, transforming AI from a single assistant into a team.
On August 17, RoboScience officially launched its first wheeled humanoid general-purpose robot, REX G1, equipped with its self-developed VLOA architecture embodied AI model Visics. Targeting logistics, factories, retail, and home scenarios, the robot features 22 degrees of freedom and ±0.1mm repeat positioning accuracy, capable of operating in 0.75-meter narrow aisles. It aims to adapt to real production environments and advance robots from single tasks to complete task execution.
Fields Medalist Timothy Gowers observes that AI's most prominent mathematical breakthroughs often come from finding counterexamples rather than direct proofs. Citing cases like the Jacobian conjecture and the Erdős unit distance problem, he argues that AI excels at searching vast spaces for special objects, leveraging cross-domain knowledge and low-cost trial and error.
Grab is using AI agents to automate analytics workflows, cutting mechanical analyst work from 44% in February to 30% in June. Its approach combines agent autonomy, certified data, context management and human oversight, with self-service analytics increasingly handling metric, data, and SQL requests without analyst intervention.
This article explores the emotional bond between children and AI robot companions like Moxie, and the potential grief when the robot's functionality declines or ceases. Through the interaction of 10-year-old Xander and Moxie, it illustrates the changing relationship and raises questions about the impact of AI companionship on children's emotions.
Amazon has been exposed for buying large quantities of printed books, scanning them as AI training data, and then destroying them, including rare books. This practice was revealed through AirTag tracking, raising concerns about copyright and cultural heritage preservation.
Several Chinese robotics and embodied AI startups have announced new funding rounds, spanning tactile sensing, physical AI data, force sensors, humanoid robots, and foundation models. ModuTech completed a Pre-A round, Mifeng Technology secured hundreds of millions of yuan led by China Telecom, and Bluepoint Touch closed a Series D. These investments indicate that Chinese capital is heavily investing in the infrastructure layer of embodied AI.
Anthropic CEO Dario Amodei posted on X, acknowledging that AI companies have failed to deliver on promises, leading to a trust crisis. He discussed regulation and power concentration, advocating for institutional constraints on frontier AI firms. LeCun criticized his views, attributing the crisis to power concentration.
OpenAI researcher Tibo announced on X that users can now enable a 1M context window for GPT-5.6 Sol in Codex with a three-line configuration, previously API-only. The setup involves adding model and context parameters to the config.toml file or using CLI flags. Tibo cautioned that long context can double token consumption and reduce performance, but OpenAI is giving users the choice.