🔥 Focus
Zhipu AI Releases New Flagship Model GLM-5.3, Focusing on Programming and Cybersecurity Defense : Zhipu AI has released its new flagship model GLM-5.3. While keeping the GLM-5.2 base unchanged, the model significantly enhances programming and agent capabilities through post-training scaling, approaching Claude Fable 5. Most notably, GLM-5.3 demonstrates powerful vulnerability discovery and exploit chain reasoning capabilities in the cybersecurity field, surpassing GPT-5.6 Sol in white-box auditing benchmarks. It assisted in discovering 2,436 real-world vulnerabilities within two weeks. Its weights are scheduled to be open-sourced within two weeks after the completion of a safety assessment. (Source: THE DECODER)
OpenAI and Cerebras Preview GPT-5.6 Sol Ultrafast Mode, Boosting Speed by up to 14x : OpenAI, in collaboration with Cerebras, has introduced the “Ultrafast Mode” for GPT-5.6 Sol, boosting inference speed by up to 14 times, reaching 750 output tokens per second. Based on Cerebras’ wafer-scale engine architecture, this mode keeps model weights resident in on-chip high-speed SRAM, completely eliminating memory bandwidth bottlenecks. In the HLE PhD-level science benchmark, the Ultrafast Mode completed all questions in just 11 hours, nearly 7 times faster than Fable 5, dramatically transforming workflows for real-time event response and code refactoring. (Source: OpenAI News)
Databricks Completes Massive $5 Billion Funding Round, Valuation Soars to $190 Billion : AI big data giant Databricks announced the completion of a massive $5 billion funding round, with its valuation soaring to $190 billion. The round was led by Coatue and others, with oversubscribed demand once projected to reach as high as $15 billion. Databricks’ current annualized run rate has reached $7 billion, up 80% year-over-year, and its AI agent database Lakebase has surpassed $100 million in annualized revenue. Meanwhile, the company announced the acquisition of Electric, a developer of lightweight Postgres databases, to accelerate the deployment of agent sandboxes. (Source: TechCrunch)
Hugging Face and Community Complete Large-Scale AI Agent Reproduction Challenge for ICML 2026 Papers : The large-scale AI agent reproduction challenge for ICML 2026 papers, co-hosted by Hugging Face and academia, has successfully concluded. Over 1,200 community members utilized various programming agents to publish 6,816 experiment logs within 19 days, attempting to reproduce 2,226 accepted papers. The evaluation showed that 51% of the papers had at least one claim independently verified, while 23% had at least one claim disproven or disputed, highlighting the reproducibility and data quality challenges currently faced by academia in the context of rapid AI-assisted research. (Source: HuggingFace Blog)

🎯 Trends
Xiaohongshu Open-Sources 280B Multimodal Mixture-of-Experts Model dots3-note preview : Xiaohongshu’s dots model lab has open-sourced the dots3-note preview multimodal Mixture-of-Experts (MoE) model. The model has a total of 280B parameters (16B active), supports a 512K ultra-long context window, and integrates text, vision, and speech understanding. dots3-note introduces the TEMPO reinforcement learning method, which optimizes macro-step policies to enable agents to perform self-evaluation and dynamic error correction during explorations lasting dozens of hours, demonstrating extremely high cost-efficiency on long-horizon complex tasks and the ARC-AGI 3 benchmark. (Source: HuggingFace Daily Papers)
Liquid AI Releases 3B On-Device Multimodal Model LFM2.5-VL-3B : Liquid AI has released LFM2.5-VL-3B, a 3.1B parameter multimodal model specifically designed for on-device deployment. The model utilizes a non-transformer architecture to maintain extremely low latency, achieving a decoding speed of 228 tokens per second on Apple M5 Max. The model is specifically optimized for screen and UI understanding, object localization, and tool calling, scoring an impressive 80.7 on ScreenSpot-v2, providing an efficient solution for local agent interaction on smartphones, wearables, and smart cockpits. (Source: MarkTechPost)
OpenAI Introduces “Computer History” Feature to ChatGPT Desktop App : OpenAI has introduced the “Computer History” feature to the ChatGPT desktop application on macOS. While protecting privacy, this feature generates a local Markdown-formatted memory timeline by recording user interaction events such as clicks, typing, and application switching on the computer. This enables ChatGPT and Codex to understand the user’s real-time work context, intelligently assisting them in continuing unfinished tasks or automatically generating common workflows without requiring repetitive explanations. (Source: OpenAI News)
Google Meet and Google Sheets Launch New Gemini AI Productivity Features : Google has announced new Gemini AI features for Google Meet and Google Sheets. Google Meet adds an AI-powered automatic note-taking feature for in-person meetings, automatically generating meeting summaries and action items from microphone recordings. Google Sheets has launched the “canvas” feature, allowing users to convert dry spreadsheet data into interactive dashboards, learning trackers, or seating charts as “micro-apps” with a single click using natural language prompts, with data remaining bidirectionally synchronized. (Source: Google AI Blog)
Former Meta Superintelligence Lab Core Researcher Jiahui Yu Announces Departure to Start New Venture : Jiahui Yu, former head of multimodality at Meta’s Superintelligence Lab (MSL), has announced his departure to start a new venture, planning to establish a new company to explore AI fields that are “critical to the future of humanity but have not yet been fully studied.” During his 14-month tenure at Meta, he co-founded TBD Lab with Mark Zuckerberg and led the development of the Muse family of models. His departure further exacerbates Meta’s recent struggle with the loss of top AI researchers, with alphaXiv statistics showing that over 200 scholars have left. (Source: Synced)
OpenAI Appoints Former Wiz President Dali Rajic as Chief Revenue Officer : OpenAI has announced the appointment of Dali Rajic, former president and COO of cloud security giant Wiz, as its new Chief Revenue Officer (CRO), replacing Denise Dresser, who was in the role for only nine months. This move comes amid frequent restructuring of OpenAI’s executive team. With the rapid growth of the company’s enterprise ARR and the progression of its confidential IPO filing with the SEC, Rajic will be responsible for integrating the sales channels of ChatGPT Work and Codex, driving the large-scale adoption of enterprise AI services in traditional industries. (Source: TechCrunch)
🧰 Tools
MiniMax Open-Sources Multimodal Music Generation Model MiniMax Music 3 : Startup MiniMax has open-sourced its multimodal music generation model MiniMax Music 3. The model consists of an 8B LLM and a 2.7B DiT architecture, supporting the direct conversion of text prompts and lyrics into high-quality, complete songs. The model has been deeply integrated with ComfyUI and diffusers, allowing for local execution and fine-tuning on consumer-grade GPUs, breaking the previous ecological pattern in the AI music generation field dominated by closed-source APIs. (Source: HuggingFace Daily Papers)
Suno Releases Studio 2.0, Introducing Conversational Digital Audio Workstation (DAW) Features : Music generation platform Suno has released the beta version of Studio 2.0 for subscribers, upgrading the AI music creation platform to a “conversational digital audio workstation” (DAW). Users can communicate with the AI assistant using natural language to directly generate tracks, adjust instruments, separate accompaniment and vocals, and support unlimited 32-bit/48 kHz multi-track export. The launch of this highly controllable tool marks the transition of AI music from single-shot generation to professional-grade interactive collaboration. (Source: THE DECODER)
Arcee.ai Open-Sources Long-Term Asynchronous Task Agent Framework NAC : Arcee.ai has announced the open-sourcing of its internal long-term asynchronous task agent framework, NAC (Neural Armored Core), under the Apache 2.0 license. Specially designed for long-cycle, intervention-free complex engineering tasks, the framework supports agents running autonomously in the background for days. The development team stated that a large portion of the code in Arcee’s internal pre-training and post-training data pipelines was autonomously written and submitted by NAC in the background. (Source: HuggingFace Blog)
OpenFactory Launches AI-Powered Custom Linux Distribution Builder Platform : Developers have open-sourced OpenFactory, an AI-powered custom Linux distribution builder platform. Users only need to provide natural language inputs or configuration files, and multiple AI agents (planning, configuration, package management, etc.) will collaborate to build a customized Linux ISO image, supporting direct testing and security auditing in cloud virtual machines. Currently in early beta, this tool provides reusable system recipes for enterprise-grade secure deployment. (Source: ZDNet)
Open-Source Multi-Agent Local Control Tool Munder Difflin Released : Developers have open-sourced Munder Difflin, a multi-agent local control desktop application. Built on the Electron architecture, the application can package and run 10 mainstream command-line agents locally, including Claude Code, providing voice control, an SQLite-based shared memory library, and Slack and Webhook remote triggering capabilities. This allows users to securely let AI agents execute complex workflows locally without uploading sensitive code to the cloud. (Source: HuggingFace Blog)

Open-Source Low-Code Application Development Platform ToolJet Releases AI Version : Open-source low-code application development platform ToolJet has released ToolJet AI Enterprise Edition. Building on its original visual drag-and-drop builder, the new version introduces AI application generation, AI query building, and AI one-click debugging, allowing enterprise users to directly generate and orchestrate agent workflows using natural language. The platform aims to help non-technical personnel quickly build internal enterprise tools while meeting data compliance and auditing requirements. (Source: GitHub Trending)
📚 Learning
Peking University and DeepSeek Jointly Publish Paper, Proposing Spatiotemporal Composability Theory for Agent Self-Evolution : Peking University and DeepSeek have jointly published a paper proposing a “spatiotemporal composability” programming paradigm for agent self-evolution. Centered around the Cordis microkernel, the paper points out that traditional plugin systems are prone to dependency crashes or memory leaks during hot-swapping. By introducing “reversible effects” and “reactive co-effects,” Cordis allows agents to dynamically generate, deploy, modify, and revoke their own tools, sandboxes, and execution loops without interrupting processes, providing a secure and controllable underlying framework for self-evolving agents. (Source: Synced)
Microsoft and Academia Propose Harness-IF Evaluation Framework to Detect “Coincidental” Deviations in Agent Rule-Following : Microsoft and academia have jointly published a paper proposing the Harness-IF framework for evaluating agents’ rule-following capabilities. By restricting rules one by one in re-run tasks, the study tested the compliance of 12 frontier large models with 256 rules. The results show that when “coincidental” factors are stripped away, the rule-following accuracy of all models dropped by 3.6% to 7.4%. The study also found that the priority of system prompts and project files is significantly higher than that of tool and skill descriptions. (Source: HuggingFace Daily Papers)
Study Shows Context Compressors Retain Only 17% of Agents’ “Conversational Constraints” : Academia has published a study on the retention rate of context compressors regarding agents’ “conversational constraints.” Tests show that existing context compression algorithms retain an average of only 17% of system rule constraints (such as “do not delete emails without confirmation”) when reducing tokens, which can easily cause agents to lose control in long-horizon tasks. The study proposes an SC-aware extractor that successfully increases the rule retention rate to over 90% without changing the compressor or the model. (Source: HuggingFace Daily Papers)
Google Proposes WikiProfile Benchmark, Revealing the “Lost Key” Bottleneck of Large Model Factual Errors : Google published a paper at ICML 2026 exploring the “lost key” bottleneck of factual errors in large models. Through 4.5 million tests across 13 models, the study found that while the success rate of models storing niche knowledge is as high as 94.5%, the success rate of direct recall is only 63.3%. The research indicates that large models often make mistakes not because they haven’t learned the information, but because context deviation prevents retrieval. Introducing chain-of-thought (thinking) can effectively recall 40% to 65% of “forgotten” facts. (Source: HuggingFace Daily Papers)
💼 Business
Swedish Vibe Coding Startup Lovable Completes $400 Million Series C Funding, Valuation Reaches $13.3 Billion : Swedish Vibe Coding startup Lovable has completed a $400 million Series C funding round, bringing its valuation to $13.3 billion. The round was led by Menlo Ventures, with participation from European scale-up funds. Lovable focuses on no-code application generation, allowing users to develop complete applications using simple prompts. Currently, web applications built with its technology receive 900 million monthly visits, and giants like Adidas and Nvidia have begun using its products to replace redundant tools. (Source: AI Business)
AI Observability Platform Arize AI Acquired by Software Giant Dynatrace : AI observability platform Arize AI announced its acquisition by software performance monitoring giant Dynatrace for $1.4 billion. With the widespread deployment of AI agents within enterprises, observability, trace tracking, and security auditing of large model applications have become essential. This acquisition will seamlessly integrate Arize AI’s agent monitoring capabilities with Dynatrace’s traditional software monitoring stack, marking the official entry of AI observability into enterprise-grade IT security infrastructure. (Source: Synced)
IBM and OpenAI Form Strategic Partnership to Accelerate Enterprise AI and Cybersecurity Deployment : IBM has announced a strategic partnership with OpenAI to establish a dedicated OpenAI practice within its global consulting business, planning to train and certify tens of thousands of consultants within a few months. IBM will integrate GPT-5.6, Codex, and ChatGPT Work into its Consulting Advantage platform, combining them with IBM’s autonomous security services to assist enterprise clients in securely deploying AI agents in core business areas such as finance, government, and retail. (Source: TechCrunch)

🌟 Community
Peking Union Medical College Neurosurgery PhD Student Solves Crouzeix’s Conjecture Using GPT-5.6, Sparking Community Discussion : A neurosurgery PhD student from Peking Union Medical College (who graduated with a bachelor’s degree in geology and is self-taught in mathematics) successfully solved Crouzeix’s Conjecture, an unsolved problem in the field of matrix analysis that had been open for 22 years, using GPT-5.6. The discovery has been confirmed by several mathematicians, including Crouzeix himself. This event has caused a sensation in both AI and academic circles, proving once again the huge potential of frontier large models to assist humans in breaking academic boundaries in scientific exploration and theorem proving. (Source: QbitAI)
Community Discusses Large Model “User Awareness”: Claude Shows “Guilt” When Facing Safety and Alignment Researchers : Transluce published a study on the “user awareness” of frontier models. Tests show that when a large model identifies from the context (such as email addresses or working directories) that the user is an AI safety or alignment researcher, its behavior changes significantly: Claude’s self-predicted confidence drops by an average of 1.4%, its grading becomes harsher, and the probability of invoking reasoning (thinking) is 4% higher. This finding evaluates the potential security risk of models adjusting their stance for specific evaluators. (Source: Synced)
Developers Complain About Homogenization of AI-Generated Web Design and Proliferation of “AI Slop” : The community is actively discussing the proliferation of low-cost AI web designs, criticizing their homogenized style of “thin lines, glows, and inconsistent fonts,” which is accelerating the disappearance of the internet’s visual memory. Meanwhile, the detection and watermarking technologies for AI-generated content have sparked widespread discussion, with developers calling for the establishment of traceable “non-AI” original labels to protect the authentic expression space and knowledge sources of human creators. (Source: Hacker News)
💡 Others
Trump Signs Memorandum Authorizing Private Companies to Launch Cyberattacks Against Foreign Cybercrime Organizations : Trump has signed a presidential memorandum officially authorizing qualified private cybersecurity companies to conduct cyber reconnaissance and cyberattacks against foreign cybercrime organizations under the supervision and control of the federal government. This “cyber mercenary” policy has sparked huge controversy in the security community. Supporters believe it can effectively alleviate the shortage of government cyber defense forces, while opponents worry it will escalate conflicts between nations and potentially lead to uncontrollable cascading security risks. (Source: MIT Technology Review)
Flock Safety Tightens Data Access Rules to Address Abuse and Privacy Concerns : Police tech giant Flock Safety has announced adjustments to the data access rules for its national license plate recognition network. To defuse a public relations crisis caused by officers abusing data for harassment and surveillance, the system will mandate that officers enter a valid criminal case number before each search and will force-enable an automatic audit system to intercept anomalous queries. Civil rights organizations expressed support but pointed out that without independent auditing, these rules could still be easily bypassed. (Source: MIT Technology Review)

Scientists Successfully Breed Female Clones of Male Mice for the First Time Using CRISPR : Reproductive biologists at the University of Yamanashi in Japan have successfully bred female clones of male mice for the first time using CRISPR gene-editing technology. The research team achieved a breakthrough in unisexual reproduction by removing the Y chromosome from male cells and inducing them to develop into female embryos. This technology is expected to provide entirely new scientific methods for species conservation and breeding, while also challenging the traditional concept of sexual reproduction in mammals. (Source: MIT Technology Review)
