EASYHUB JOURNAL / 2026-09
AI news in September 2026
64 stories, ordered by their original dates. Open a title for its summary, image and source link.
- Microsoft
Copilot brings Home, Code and Autopilot into a staged rollout
Copilot is bringing chat, Cowork and editable Office files into Home, adding Code for building apps in managed environments and Autopilot for persistent cloud tasks. Home and Code roll out through Frontier, while Autopilot expands in private preview at month-end. The announcement is not universal availability, and intensive agent work uses consumption-based billing.
- IT之家
DeepSeek Harness gets a desktop developer preview for Windows and Apple Silicon
The V0.1.7-rc.2 desktop developer preview supports Windows x64 and Apple Silicon Macs without first installing Node.js or launching the WebUI from a terminal. Standard mode handles files and code, PTC supports batch tool-result processing, and create mode customizes plugins and interfaces. It remains a preview rather than a stable or unsupervised execution guarantee.
- Meituan LongCat
LongCat-2.5-Preview adds image understanding for coding and extended tasks
LongCat’s dated release note introduces 2.5-Preview with image understanding, cross-modal questions and support for coding and multi-step tasks. It integrates with environments such as Claude Code, OpenClaw and OpenCode. API and web access remain a preview service, not an announcement of downloadable weights or unlimited free usage.
- Anthropic
Claude Science performs a nine-loop amplitude calculation checked by researchers
An invited Anthropic article describes Claude Science organizing code and compute for a nine-loop amplitude calculation in a specific N=4 super-Yang–Mills setting, with Lance Dixon checking the result. It also acknowledges related AI-assisted work by Song He’s team and the use of known methods. This is not a new physical law or a blanket claim of autonomous science.
- LangChain
LangSmith smithtune connects agent traces to fine-tuning and evaluation
The public beta turns successful agent trajectories into datasets and connects curation, train/test splits, supervised fine-tuning and evaluation through smithtune. Training runs on Fireworks or Baseten and requires the relevant accounts and API keys. It targets repeated specialized tasks; the announcement also shows that poorly selected training data can make results worse.
- IT之家
Tencent Hy Translation launches voice, camera and offline translation apps
Tencent packages Hy-MT2 into a translation app with voice, camera and 33-language translation plus selected minority languages and dialects. Offline use requires downloading the local model first. The app initially launches in 12 markets, alongside mini-program, plugin and PC experiences, rather than in every app-store region.
- IT之家
QClaw announces December 24 shutdown and migration to WorkBuddy
QClaw stops new registrations and subscription purchases or renewals on September 24, with its main service scheduled to close on December 24, 2026. Existing users can back up settings, memories and conversations or authorize migration to personal WorkBuddy accounts. Backup access remains until March 24, 2027.
- IT之家
WorkBuddy 5.6.1 adds WeChat mini-program generation and publishing
From version 5.6.1, WorkBuddy can generate and preview WeChat mini programs with optional databases, authentication and storage. Users without an account can create a 14-day trial. Long-term operation still requires a linked mini-program account and WeChat review before a production release.
- Google
Introducing Gemini 3.8 Live with Live Avatar
Google pairs native live dialogue with low-latency video to give conversational avatars synchronized expressions and speech. The announcement places Live Avatar in Gemini Enterprise for uses such as interactive walkthroughs, rather than promising availability to every consumer account.
- Sarvam AI
Sarvam Vision 2.1: Pushing the Pareto frontier of document intelligence
Sarvam AI updates its document vision model with multipage table handling, key-value extraction and Indic handwriting recognition, alongside serving improvements. The article separates global and Indic evaluations and discusses weaknesses; English benchmark performance alone is not a guarantee of local-language usability.
- Anthropic
Project Swap studies whether agents understand the preferences they represent
Anthropic ran a controlled book-swapping market with 201 employees and Claude-powered agents, comparing centralized allocation with agent-to-agent negotiation. Understanding what people actually wanted was a larger bottleneck than bargaining. The research explores preference representation and market rules, rather than announcing a consumer shopping or trading service.
- 人人都是产品经理
A community Qwen Audio Studio brings scripts, voices and previews together
Creator Ai学习的老章 presents a community Qwen Audio Studio with templates for podcasts, ads and audiobooks, controls for characters and reference voices, and preview, download and version comparison. A companion automation skill is also described. It is a community interface to Qwen audio services, not an official Alibaba client or proof of offline local-model inference.
- 每日经济新闻
Yuanbao reaches HarmonyOS with expert mode, voice and document tools
Yuanbao’s HarmonyOS app combines Hy4 preview and expert mode with writing, voice calls, image recognition, translation and document reading. It also exposes Hy Image3.5 preview and uses WeChat articles and video content as search sources. WeChat and QQ sign-in continue account history; phone-only accounts need to link the relevant login first.
- Meta AI
Meta announces Ray-Ban Meta Audio while Muse integration remains ahead
At Connect, Meta introduced the roughly 43-gram Ray-Ban Meta Audio for open-ear listening, calls and voice-assistant access, alongside updated Gen 3 glasses. Audio opened for preorder with shipping planned for October 13. Bringing the Muse personal agent to the glasses remains a forward-looking addition, not a capability already available on every pair.
- Anthropic
Claude discovers a novel enzyme system with CRISPR-like repeats
Anthropic reports early work in which agents identified an enzyme system associated with CRISPR-like DNA repeats. Laboratory research is continuing, and the system's function remains unknown. This is an early discovery report, not a validated treatment or gene-editing product.
- Cursor
Bots for the last mile: Rollouts, Security Review
Cursor introduces Rollouts and Security Review to help follow changes through deployment, inspect regressions and review security concerns. They address work after code generation, without replacing a team's responsibility for testing, review and release decisions.
- Cursor
Improved token efficiency for longer agent runs
Cursor explains changes to assembling and managing context for longer agent runs, reducing repeated processing overhead. The post reports observations from production traffic and focuses on harness and caching efficiency rather than introducing a new foundation model.
- Google
Gemini 3.8 text-to-speech says hello
Google introduces Gemini 3.8 Flash TTS and Flash-Lite TTS with more flexible voice, emotion and dialogue direction. The release discusses developer access and integration into creative products, alongside safeguards for generated speech.
- Google
A new wave of Connected Apps is rolling out to Gemini.
Google expands Gemini's Connected Apps across design, project management and everyday planning. Integrations are rolling out and require the relevant user connection or authorization; they do not automatically grant access to every external account.
- Epoch AI
Epoch AI tests furniture-assembly reasoning, with sample size and latency limits
Epoch AI evaluated assembly-error detection using manuals and 60 photos from three furniture builds. Its September 23 report gives GPT-6 Astra an 80% score with a median of about three minutes per photo. The authors emphasize limited coverage and latency; these results do not establish reliable real-time guidance for arbitrary repair tasks.
- IT之家
Ant's Ming models generate designs and editable layers, with companion skills
The two 6B Ming models target text-rich visual design and decomposition into editable transparent layers. Companion Design and PPT skills connect the output to page and presentation workflows. Complex occlusion, reflections and ambiguous layer boundaries remain limitations; the right layer granularity also depends on the intended design task.
- Alibaba Cloud
Qwen-Audio-3.1 spans transcription, soundscape generation and live dialogue
Qwen announced five audio models covering transcription cleanup, speaker-and-timestamp recognition, controlled speech synthesis, combined dialogue and soundscape generation, and duplex conversation with tools. The verified Qwen account says the APIs are available on its platform. Individual model access and usage quotas depend on the service, rather than a universal free allowance.
- GitHub
Copilot for JetBrains 1.18.0 adds rewinds, shared skills and assisted approvals
Version 1.18.0 lets users revise earlier prompts while rewinding conversation and file changes, and brings shared organizational skills into local and agent sessions. Assisted approvals are in public preview, with higher-risk actions still requiring a decision. The update also adds Codex planning and persistent MCP controls, while hiding inline chat in remote-development environments.
- Anthropic
Introducing Claude Opus 5.5
Anthropic introduces Claude Opus 5.5 with changes to codebase-wide work, computer use, professional tasks, efficiency and safeguards. Performance and savings discussed in the announcement are evaluation results from the publisher, not guaranteed outcomes for every workload.
- IT之家
Hy Image3.5 preview adds five-image references and multi-turn editing
Hy Image3.5 preview supports text-to-image generation, reference-image editing and multi-turn revisions with up to five input images and 2K output. It reaches Yuanbao, WorkBuddy, ima, Miora, WorkRally and OnSolo, alongside a cloud API. These are application entry points for the model, not separate model releases.
- Hugging Face
Transformers now runs llama.cpp quants
Hugging Face adds efficient GGUF inference to Transformers so developers can use familiar Python interfaces with local quantized models. Initial support has architecture and hardware limitations; the announcement does not promise identical performance across all models or devices.
- OpenAI
Better prompt caching for GPT-6
OpenAI describes prompt-caching improvements for long conversations and agents, including diagnostics, explicit breakpoints and preserving reusable context when reasoning effort or tool availability changes. The post explains integration choices rather than promising cache hits for every request.
- OpenAI
Introducing GPT-6 Sol and Luna
OpenAI adds Sol and Luna to the GPT-6 family for different cost and latency requirements. The announcement covers ChatGPT Work, Codex and API access, while explicitly noting that these models were not yet available in ordinary Chat.
- Alibaba Cloud
Alibaba outlines agent-cloud and context services while Qwen 4 remains in training
Alibaba outlined an AI stack spanning chips, models, agent deployment and enterprise context. Agent Native Cloud addresses operation and management, while Agent Context connects business data and long-term memory. The same announcement says Qwen 4 is still training and Qwen-Image 3.1 is planned for later in the year; the roadmap is not a blanket general-availability release.
- Xiaomi MiMo
HySparse2 studies two-level KV sharing for long-context agent inference
Xiaomi's HySparse2 paper combines two-level KV sharing, token-level sparse selection and a shorter prefill path to reduce the overhead of long tool-rich agent histories. Experiments use an 80B-total, 3B-active MoE model to compare retrieval and inference costs. These are architecture research results, not a release of a complete MiMo-V3 model.
- Xiaomi MiMo
Xiaomi opens MiMo-V2.6 models and training resources for multimodal long tasks
MiMo-V2.6 introduces Pro and Flash with native text, image, video and audio inputs for long coding, design and tool-based tasks. Xiaomi is publishing the technical report, training environments and reinforcement-learning code, alongside a staged UltraSpeed rollout. Open weights do not imply ordinary-laptop deployment; total parameters, active parameters and memory needs remain distinct.
- Hugging Face
tokenizers v1: encode, decode and scaling, measured
Hugging Face previews tokenizers v1 performance work across encoding, decoding and scaling to reduce CPU tokenization bottlenecks. The article discusses release-candidate progress and should not be read as confirmation that every described change has shipped in a stable release.
- OpenAI
Expanding OpenAI Academy with new learning paths
OpenAI Academy adds learning paths for developers, leaders, educators and students. Courses emphasize practical tasks, context and checking results, with badges awarded after assessments. These course badges should not be confused with academic degrees or general professional credentials.
- Strands Agents
Strands Harness packages a general agent for local and cloud deployment
Strands Harness packages file, shell and web tools with context management, session resumption, memory and subtask delegation. Developers choose a model, use Python or TypeScript and deploy locally or to Linux-container platforms. It is an agent framework, not a new model; the publisher's cost reductions come from particular benchmarks rather than a universal workload guarantee.
- IT之家
Yuanbao’s AI recorder adds photos, device audio and WorkBuddy sharing
From Yuanbao 2.85.0, the AI recorder can capture photos alongside audio, keeping images, transcripts and summaries in one record, and record internal device audio. Users can correct transcripts, manage recordings in batches and pass results to Tencent Docs or WorkBuddy for further document and slide work.
- Anthropic
Partnering with Accenture on embedded evaluation
Anthropic announces an evaluation partnership with Accenture involving Faculty, covering model assessments, red teaming and safeguards. The partners plan to build deeper evaluation capacity, while operational details are still developing; this is not a completed comprehensive external audit.
- IT之家
WorkBuddy 5.5.6 generates web apps with databases and sign-in
WorkBuddy 5.5.6 can generate full-stack web apps with optional databases, file storage, authentication and AI calls. Connecting cloud resources requires user confirmation. Running an app consumes cloud-service resource points subject to plan quotas, so deployment-free setup does not mean unlimited free hosting.
- Huawei
Huawei Cloud Rolls Out Enterprise AI Products Across the Board, Building an Open Agentic Cloud
At Huawei Connect, Huawei outlines AI cluster services, model services and AgentArts for enterprise agents. The announcement includes staged availability across markets. A global product announcement should not be interpreted as every service being commercially available in every region that day.
- Huawei
Huawei Unveils New UnifiedBus Computing Architecture for SuperPoDs and Clusters
Huawei introduces UnifiedBus for interconnecting SuperPoDs and clusters, describing the coordination of compute, memory and storage for AI workloads. This is an infrastructure announcement; performance targets and subsequent product plans depend on their stated implementation and delivery conditions.
- Z.ai / 智谱
Toward Recursive Self-Improvement: How GLM Built Its Own Inference Infrastructure
Z.ai describes an Infra Agent powered by GLM-5.3 assisting with model adaptation, correctness checks and inference optimization on Chinese accelerators. The account emphasizes detailed feedback and engineering collaboration. It is an early infrastructure case, not proof of fully autonomous model self-improvement.
- Mistral AI
Mistral and Mozilla are bringing open, private and multilingual AI to your web browser
Mistral announces models for Firefox Smart Window beta, supporting assistance grounded in browsing context. The release distinguishes initial coverage in France and North America from later planned markets and discusses user control and retention arrangements. It is not a globally enabled browser default.
- Google
Introducing Gemini 3.8 Live and 3.8 Live Extended Thinking
Google introduces Gemini 3.8 Live and its Extended Thinking variant, emphasizing spoken interaction, live visual context and parallel task handling. The aim is to continue a conversation while more involved work runs, with access varying by product and API.
- Google
Sharpen your study routine with new Gemini Notebook tools
Gemini Notebook adds live conversations around notebooks, recording and interactive learning overviews, including quizzes, flashcards and video summaries. These features help organize study from existing materials, while generated explanations still need checking against their sources.
- Cohere
Introducing North Small Translate: A leading sovereign open-weight machine translation model
Cohere introduces North Small Translate for more than fifty languages, developed with RWS. The research weights use a non-commercial license, while commercial translation is offered through separate product channels. Open weights do not imply unrestricted commercial use.
- Cursor
Introducing Projects
Cursor Projects maintains context for larger work such as features and migrations. A coordinator delegates to subagents and handles recurring work, while developers set goals, steer progress and review results rather than manually organizing every session.
- DeepSeek
DeepSeek V4.1 Flash introduces native visual understanding
DeepSeek introduces V4.1 Flash with native visual understanding and a redesigned architecture for agent workloads. This entry records the release, not a blanket retirement of V4 Pro: the later API changelog says that Pro service continues.
- Google
The Gemini app is now available for Windows
Google introduces a Windows Gemini app with a desktop shortcut for assistance while working.
- OpenAI
Build more natural voice experiences with GPT-Live-1 in the API
OpenAI introduces full-duplex voice capabilities for developers, including tool delegation and voice-agent workflows.
- Microsoft
What’s new in Microsoft Foundry: July and August 2026
Microsoft Foundry reviews July and August changes to hosted agents, Voice Live, Toolboxes, model routing and SDKs. The roundup mixes generally available features with preview local-development capabilities, so status and migration guidance must be checked for each component.
- Adobe
Generate and create directly in your timeline with new AI-powered innovations in Premiere and After Effects
Adobe introduces a Generative Media Tool for producing video and sound effects within Premiere timelines, and brings AI Assistant to After Effects in public beta. Editors retain responsibility for reviewing results and applicable media-use terms.
- Cohere
Inside the megakernel serving engine for North Mini Code
Cohere describes a serving system organized around a decode megakernel for North Mini Code, with continuous batching, paged attention and tool-call support. Its implementation and benchmarks are workload- and hardware-specific, not a universal speedup for every language model.
- IT之家
Hunyuan and WorkBuddy tune Hy4 preview to reduce excessive reasoning
Tencent’s Hunyuan and WorkBuddy teams update Hy4 preview to address overly long reasoning and repeated self-verification on complex tasks. They report fewer task rounds and lower input/output token use under benchmark and human evaluation, without a numerical reduction. This is a preview update, not a formal Hy4 release.
- IT之家
WorkBuddy brings task synchronization and audio results to HarmonyOS watches
WorkBuddy adds a HarmonyOS watch app for several HUAWEI WATCH families. It synchronizes tasks, notes and conversations and lets users review computer-generated results or listen to them as audio. Longer content can move back to a phone, while HarmonyOS sharing can pass files into desktop conversations.
- OpenBMB / 面壁智能
MiniCPM5-2B opens weights and training data for local agents
MiniCPM5-2B releases weights, intermediate checkpoints and training datasets alongside GGUF and MLX deployment options for local assistants and tool use. Its full parameter count is about 2.52 billion, including roughly 1.98 billion non-embedding parameters. The 2B-class name should not be confused with the complete parameter total.
- OpenAI
Research acceleration: The view inside OpenAI
OpenAI reports observations about coding-agent use by its researchers, including experiment velocity and changing workflows. This is an internal usage and research account; the correlations should not be presented as independently established causal effects.
- Google
Create your best tracks yet with Lyria 3.5 in Gemini
Google describes updated music generation with genre prompts, vocal or instrumental options and creative templates.
- IT之家
WorkBuddy adds Kylin, UOS and other Linux distribution support
WorkBuddy is listed in the app stores for Kylin, UOS, openKylin and deepin. Tencent says the adaptation covers system sign-in, device connections, skill execution and output previews, extending its desktop workspace beyond Windows, Mac and HarmonyOS to these Linux environments.
- OpenAI
GPT-6 Astra: A new generation of intelligence
The announcement describes a new model for computer use, coding and knowledge work, alongside safeguards.
- Hugging Face
Give Your Coding Agents a Memory You Own
Funes builds searchable memory from local coding-agent sessions with provenance, helping different tools retain project context. It supports local processing and optional user-directed dataset synchronization rather than automatically uploading all conversations to a public service.
- Cursor
Run cloud agents on machines you manage
Cursor cloud agents can execute on machine pools managed by a team, closer to internal services, custom hardware and specialized build environments. Infrastructure is under the team's control, while Cursor still orchestrates agents; this is not simply an offline mode.
- IT之家
WorkBuddy opens a developer platform for skills, experts and connectors
The WorkBuddy platform exposes agent capabilities to hardware partners, industry applications and developers. It manages skills, experts, connectors and external integrations in one place. Individual and business developers must complete onboarding and verification. This expands the application ecosystem rather than introducing a foundation model.
- Microsoft
New and improved: GitHub Copilot harness, agent skills, and richer context
Microsoft summarizes Copilot Studio changes including general availability of the GitHub Copilot harness, skills, memory, MCP connections and enterprise context. Although it covers August updates, the article belongs to September in the archive because that is when it was published.
- Anthropic
Introducing Claude Fable 5.1 and Claude Mythos 5.1
Anthropic introduces Fable 5.1 and Mythos 5.1 as the same model with different safeguards. Fable 5.1 is generally available, while Mythos 5.1 is limited to trusted-access programs rather than being an ordinary release for all users.
- Google
The latest AI news we announced in August 2026
Google’s monthly roundup covers Gemini models, transcription, creative tools and device-related announcements.