Qwen-Image-2.1 unifies editing and transparent image generation
Qwen announced an open-source image model combining generation and editing. Its visual generation component has seven billion parameters.
From model advances to new AI products, agents and tools. Follow the updates, check the sources and understand what changes for you.
Explore the latest
Qwen announced an open-source image model combining generation and editing. Its visual generation component has seven billion parameters.
Qwen's interpretation update adds speaker separation and synchronized source-and-translation output. Developer testing reports average lag falling from 2.8 to 2.3 seconds.
Grok's new transcription model targets multilingual and difficult audio. Batch and streaming access retain previous rates.
OpenAI is testing the feature with selected US advertisers. Users can opt into labeled agent conversations after clicking ads.
Six launch and cookbook figures explain typed decisions, cost comparisons, re-ranking and a practical Python integration for TypeSafe’s Jev.
Notion added reusable team skills. Members can share instructions and invoke them in Agent conversations.
ElevenLabs expanded its MCP with creative tools. Connected assistants can generate speech, music, images and video.
Kimi completed the K2.8 Preview rollout in Kimi Code. Existing clients retain the kimi-for-coding identifier.
Google released its desktop app for Windows 10 and 11. The shortcut opens assistance above the current workspace.
OpenAI's full-duplex model can listen while speaking. It delegates deeper reasoning and tool use to backend models.
Cursor launched Projects in beta. A coordinator delegates work while retaining shared project context.
DeepSeek released a native multimodal model under deepseek-flash. Its architecture separates input and output computation.
OpenAI introduced a Data agent to investigate metrics and build interactive dashboards from approved sources while retaining connected-account permissions.
The API exposes Codex's agent harness, supporting extended sessions, tools, and subagents with hosted or developer-selected compute environments.
Runway added Premiere Pro and After Effects panels. Generated images and video return directly to sequences or compositions.
A practical look at Astra’s context, async tools, mid-turn steering and Ultrafast, with concrete workflows, API examples and a method for evaluating completed work.
Google introduced Flash for agent workflows and Flash Cyber for defense. Flash retains 3.7 Flash's introductory rates.
Anthropic says the releases share a model but use different safeguards. Fable is generally available; Mythos requires trusted access.
Figma updated agent-generated plugins and shaders. New controls support sharing, animation, interaction and code downloads.
Z.ai released GLM-5.3 using the GLM-5.2 base with expanded post-training. Disabling thinking is no longer supported.