Skip to content
B Bloger.fm

Organisation

Alibaba

The Chinese cloud and commerce group whose Qwen team publishes the Qwen model family. Qwen models are both served through Alibaba Cloud and released openly, which has made the family a common base for work done elsewhere.

Alibaba was recorded in 14 items across 8 of the 8 briefings in the current window.

Its share of coverage was steady: 5 items in the first half of the window and 9 in the second, tracking the feed as a whole, which grew about 2.2×.

It appeared most often alongside Qwen, Claude and Anthropic.

Tracking the feed
items
14
briefings
8
mentions
27
last seen
2026-09-19
Official channel
alibabagroup.com

Coverage timeline

Sat 12 Sept – Sat 19 Sept / 8 briefings

Products and models

Everything recorded

19 September 2026 4 items

Alibaba kept this AI code reviewer inside for 2 years. Now it’s free. OpenCodeReview reportedly found millions of defects across tens of thousands of developers before being released publicly. Repo:

A great example of practical AI solving real-world problems! Faster than ever. Thanks for building with Qwen. @cerebras Try it out for yourself! 🏠

@cerebras

We built Money Agent, a personal finance assistant powered by @Alibaba_Qwen 3.8 27B on Cerebras. > Now you can turn a home-buying question into a real-time conversation, then into a financial goal, faster.

truncated at source

qwen's new live translation model supports 60 languages and cuts average lag from 2.8 to 2.3 seconds. it can separate speakers in a group conversation, preserve their voices, show both languages at once and use earlier context to keep names and terms consistent.

@Alibaba_Qwen

Meet Qwen3.8-LiveTranslate, Qwen's next-generation real-time simultaneous interpretation model! 📢 > Built on an Interleave architecture, it improves faithfulness, fluency, and conciseness while reducing average lagging (LAAL) from 2.8s to 2.3s across 60 languages. > New capabilities: 🙌 - Real-time speaker diarization — distinguishes speakers in multi-party speech and preserves each speaker's voice through more stable voice cloning. - Synchronized bilingual display — source and translation on screen together. - Long-context disambiguation — leverages conversation history to clarify names and terminology for consistent translations. > Let's try Qwen3.8-LiveTranslate! 🥳 - Blog: ht

truncated at source

UPDATE: Qwen3.8-Flash for a single DGX Spark 🔥 - 117 tok/s prose & 180 tok/s code at 8 streams. - Optional official Nvidia NVFP4. - 24/7 auto-restart supervisor. - Cached-token reporting in every response. - Peak memory down from 101 to 91 GiB. - LOTS of bugs were fixed. This is still the BEST model to run on a single spark. Full details below 👇 Get it here:

@jvr0x

Big update to the @Alibaba_Qwen Qwen3.8-Flash-Next single DGX Spark recipe! > 𝗪𝗵𝗮𝘁'𝘀 𝗻𝗲𝘄 🎁 > • Measured on one DGX Spark, 262K context, MTP k=3, aggregate tok/s at 1 / 2 / 4 / 8 streams: > Prose: 38.0 / 61.1 / 89.2 / 117.4 Code: 53.8 / 87.5 / 131.8 / 180.2 > • Long context holds: MTP keeps working at a 185K-token prompt (35.4 tok/s decode), prefill ~2,000 tok/s from 4K to 185K > • ~1M-token KV pool at the full 262K context (FP8 KV) > • NVIDIA's official NVFP4 checkpoint now runs on one Spark, with chat, tool calls and vision w

18 September 2026 2 items

Open Research is really unbeatable!

@gajesh

Together, the MLX(.)fast community has made Qwen 3.8 Flash nearly 2x faster on Apple Silicon! > We're ready to bring it to @DarkbloomAI: an open network of local Mac machines providing inference to the world. One thing remains: the community flagged that its license requires a separate agreement for commercial model serving, so we're holding the launch until that's in place. > We believe this is a great opportunity for the local community: one where we make Qwen models faster and more accessible, and the people running them share in the value they create. > .@Alibaba_Qwen @QwenDevs, we'd love to work together on this. > If anyone else knows someone we can talk to, we'd love to have that conversation. Let's make Qwen 3.8 Flash on Darkbloom a reality!

truncated at source

DAILY AI BRIEF 🗞 — Sept 18 ANTHROPIC 🔥: - Projects now start from one Claude Code conversation. Claude spins parallel cloud threads, keeps shared memory, and surfaces an Overview panel. XAI 🔥: - Grok Bot voice is live. Desktop and mobile, rolling out over the next couple of days. META 🔥: - Muse for Mac is out, US only. Computer use across apps, files, calendar, notes, and messages. You pick what it can access. PERPLEXITY 🔥: - Effort selector is live in Computer on web. Presets pair the orchestrator model with reasoning depth. Mobile and desktop next. OPENAI 🔥: - Astra for Law is out: GPT-6 Astra plus a Legal Search Index over 230M+ URLs. Trusted Access first, API soon. - ChatGPT in Word hits all plans including Free, with usage limits. Business and Enterprise get a two-week GPT-5.6 Sol preview. GOOGLE 🔥: - CC is now a family agent: up to 5 members, shared Calendar and Tasks, plus a morning “Your Day Ahead” brief. Waitlist, US 18+. ALIBABA 🔥: - Qwen3.8-Omni-Flash is out. First omni-modal agent model, 1M

17 September 2026 1 item

This is where Wan3.0 starts fitting into real production. A year ago: 5-second clips. Then: 15 seconds. Now: a single 30-second shot, straight from the model. Add director-level control and omni-reference that takes up to five videos. Not five images, five videos. The filmmaker behind Soulscape and Johnny Mai from Alibaba Cloud put Wan3.0 into an actual production workflow. No more stitching together endless short clips. Generate long takes, then cut them to the script and the story. Want to bring Wan3.0 to your team? Learn More → #Wan3 #VideoProduction #AIVideo #Filmmaking

16 September 2026 2 items

Wan3.0 isn't an incremental update. Look at the last year of releases: Wan 2.1 → 2.2 → 3.0 That's more than a generational leap. Motion reference. Consistency. Creative control. Resolution. Every step makes the model more friendly to the people actually using it — creators, filmmakers, storytellers. The filmmaker behind Soulscape and Johnny Mai from Alibaba Cloud on what changed with Wan3.0, and what surprised them most. The tools are finally getting out of the storyteller's way. Want to bring Wan3.0 to your team? Learn More →

Motion transfer without the usual motion capture setup. @alibaba_cloud's Wan Animate 2 (14B, open-source, Apache 2.0) lets you take a character image and a reference video, then transfer the movement onto the character while keeping their face, outfit, and identity consistent. No pose rig. No skeleton extraction. Record the movement on your phone, pair it with a character image, and run it. It works especially well for full-body movements with a steady camera. Fine finger details and longer clips can drift, so shorter clips around 81 frames tend to give cleaner results. Try it on Floyo:

15 September 2026 1 item

I've added 3 new models to the H3 Acceleration Arena for evaluation VDN-H3 (8 steps), Lightx2v 1.2 (8 steps), Alibaba TaoMate H3 (3 steps)

14 September 2026 1 item

truncated at source

Open weights. Shared progress. MiniMax H3 is moving fast. We built MiniMax H3 for video generation with native stereo audio and multimodal reference control. The open-source community is making that capability faster, more accessible, and easier to build on. Recent highlights: • FastH3 — FastVideo, Nuva Lab and NVIDIA: 4-step distillation, now running on DGX Spark and Apple Silicon. • Sol-H3 — NVIDIA’s SANA team: now on DGX Spark with a two-stage H3 + LTX-2.5 pipeline. On 8×B300, the team reports 15 seconds of 768p video + audio in 6.6 seconds of warm inference.* • VDN — Haocheng Xi and the OpenVDN team: rethinking attention for faster H3 inference, with weights, training and inference code released. • PDD — NVIDIA’s distillation method, brought to H3 by Alibaba PAI as 8-step Acc-LoRAs, now supported in ComfyUI. • LightX2V — 4- and 8-step Turbo LoRAs, with workflows for text, image and reference-conditioned video + audio. Behind every release are people training, optimizing, quantizing, testing and sh

13 September 2026 1 item

Taomate H3 by Alibaba, a new acceleration method for Minimax H3. This streams the generation in small chunks and supports continuous, long videos.

12 September 2026 2 items

Fast meets open. 🚀 Qwen3.8-27B is now running on @cerebras with rapid inference. Try it now!

@cerebras

Qwen3.8-27B is now live at Cerebras speed. > The dense, open-weight model from @Alibaba_Qwen scores 34 on the Artificial Analysis Intelligence Index—making it comparable to models such as GPT-5.6 Luna, DeepseekV4 Pro, and Claude Sonnet 4.6.

truncated at source

DAILY AI BRIEF 🗞 — Sept 12 OPENAI 🔥: - GPT-Rosalind is out of research preview for eligible orgs worldwide. API, Codex, and ChatGPT Enterprise, with new Rosalind models as they ship. - Codex adds Life Sciences plugins for genomes, protein structure, QC reports, and notebooks. - ChatGPT Sites hit 5M apps. New: collab editing, private invites, custom domains, and DB inspect. - Desktop pets can start a new chat. Mini is the compact no-pet option. XAI 🔥: - Elon: Grok 4.7 needs a few more days. RL still quits hard tasks too early and undershoots self-checks. - Grok Bot is rolling out on Grok web for Heavy users. Create and chat with bots in the UI. No official post yet. ANTHROPIC 🔥: - Claude Code ships `claude plugin eval`. Score a plugin on test cases, then rerun without it. MICROSOFT 🔥: - MAI-Transcribe-2 hit 1M OpenRouter requests in 5 days. ALIBABA 🔥: - Qwen3.8-27B is live on Cerebras. * Used Grok to compose this brief, cherry-picking the news and doing some post-editing. > [@testingcatalog](