Next week in AI: We’re probably getting many models, it will be a huge week, especially for those without unlimited money to spend. We’re getting Opus 5.2 (and possibly Fable 5.2/Sonnet 5.2), GPT-6 Sol and Luna, and we also might get Kimi K3.1. I’m so excited, especially about GPT-6 Sol
Product / Moonshot AI
Kimi
Moonshot AI's assistant and long-context model family.
Kimi was recorded in 21 items across 6 of the 8 briefings in the current window.
Its share of coverage was steady: 7 items in the first half of the window and 14 in the second, tracking the feed as a whole, which grew about 2.2×.
It appeared most often alongside GLM, DeepSeek and GPT-5.6 Sol.
- items
- 21
- briefings
- 6
- mentions
- 84
- last seen
- 2026-09-19
- Background
- en.wikipedia.org
Coverage timeline
Sat 12 Sept – Sat 19 Sept / 8 briefings
Appears alongside
GLM
ProductZhipu AI's model family, published openly and served through its Z.ai platform.
32 items / 7 briefings
RisingDeepSeek
OrganisationA Chinese AI lab known for releasing open-weight reasoning and coding models. Its releases are watched for what they show about training costs as much as for what the models can do.
37 items / 8 briefings
FallingGPT-5.6 Sol
ProductNo definition written; coverage recorded from the feed.
54 items / 8 briefings
Anthropic
OrganisationAn AI safety and research company, and the maker of the Claude model family and Claude Code. It serves its models as a hosted product and publishes research on model behaviour alongside them.
44 items / 7 briefings
DeepSeek V4.1 Flash
ProductNo definition written; coverage recorded from the feed.
31 items / 8 briefings
FallingGemini
ProductGoogle's multimodal model family and the assistant built on it. It is embedded across Google's products, which gives it a distribution no standalone assistant matches.
72 items / 8 briefings
Everything recorded
19 September 2026 3 items
We’re bringing Databricks Unity Gateway to developers on Neon (@neondatabase), and it’s the fastest AI Gateway for Kimi K3! Databricks inference is moving super fast!
Kimi 新模型,万众期待的 K3.1 即将发布。
18 September 2026 5 items
月之暗面上线新版 Kimi 会员。四档套餐改名为 Go、Plus、Pro 和 Max,年付分别为 468、948、1908 和 6708 元,折合每月 39、79、159 和 559 元。 Kimi Code 没有按照最早公布的方案拆出去单独收费。新版 Plus、Pro 和 Max 仍然包含 Kimi Code,只有最低档 Go 不支持。 Kimi 7 月因 K3 上线后算力紧张暂停新用户订阅,并宣布后续会将 Kimi 主会员和 Kimi Code 分开收费。随后在用户压力下,拆分收费方案基本撤回。 新四档年费与旧 Andante、Moderato、Allegretto、Allegro 完全一致。不过旧套餐的最低档还能用 Kimi Code,现在同价位的 Go 已经取消了这项权益;要用 Code,年付门槛从 468 元升到了 948 元。
@0xLogicrwKimi 重做个人会员体系。原本同时包含网页、App 和 Kimi Code 的会员,被拆成通用会员与 Coding Plan。两类功能都要用,今后需要分别付费。 > 按年付计算,旧 Andante 为 468 元,同时包含网页端权益和基础 Code 额度。新体系最低组合是 Go 加 Starter,共 1536 元。入门门槛变为原来的 3.28 倍,上涨 228%。 > 纯编程用户的变化更复杂。新 Starter 与旧 Moderato 都是每月 99 元,但不再包含网页和 App 权益。需要高速版和 100 万上下文时,月付门槛则从旧 Allegretto 的 199 元升至新 Explorer 的 299 元,上涨约 50%。 > Kimi 称,新 Code 套餐取消月总额度上限,同价位可用 Token 更多。但周额度和 5 小时频控仍然存在,官方也没有公开可直接比较的新旧绝对额度。目前无法判断增加的额度,能否抵消被移除的通用权益和部分档位涨价。 > 老套餐不会被强制迁移。现有用户可以继续使用和自动续订,暂时没有明显理由主动切换。
月之暗面正式上线 Kimi Code Desktop,把原本偏终端的 Kimi Code 做成了独立桌面应用。用户可以直接打开本地项目,让 Agent 读写代码、跑命令、查看 Git 改动,目前支持 macOS 和 Windows。 桌面版直接内置了浏览器和终端。做网页时,Agent 可以读取右侧网页的页面信息,用户也能直接点选元素、框出区域或者截图批注,告诉它具体哪里要改。改代码、看网页效果、跑测试都不用再切窗口。 它还支持计划、目标和 Swarm 三种任务模式。Swarm 可以把任务拆给多个子 Agent 并行处理,长任务也能放到后台继续跑。除了 Kimi 自家的模型,用户还可以自己填 API Key 和 Base URL,接入其他模型供应商。
@real_kai42Kimi Code Desktop 发布啦 > 高情商: 测 broswer-use 低情商:带薪做旅游计划()
Kimi K3 is now free in Cline Desktop to celebrate all the new users and give them more tokens to continue exploring the app ❤️
@clineIntroducing Cline Desktop - a native interface for working with open weights models. > Use with ClinePass and all our free models like DeepSeek-V4.1-Flash, Musespark-1.3, or BYOK with any provider!
I HAVE ALWAYS SAID THE NARRATIVE AND TIMELINE IS MOVING FROM USA LABS TO CHINESE LABS this is another proof of that look at the numbers Bolt shared from the first few days of open models on Forge - 1. GLM 5.3 Flash 54% 2. DeepSeek V4 Pro 17% 3. GLM 5.3 15% 4. Kimi K3 14% Chinese labs are moving fucking fast rn
@boltdotnew11M+ people build on Bolt. On Monday, we gave them open models with up to 50x usage. > Which is #1 so far? The fastest, cheapest one, with prompts up to 2x the size. > 🥇 GLM 5.3 Flash 54% 🥈 DeepSeek V4 Pro 17% 🥉 GLM 5.3 15% 👏 Kimi K3 14% >
Open models are proving they can lead, not just fill the gaps. Real builders are putting them to work, and the numbers show it. GLM 5.3 Flash captured 54% of Forge prompts through Wednesday, with Flash getting up to 50x usage.
@boltdotnew11M+ people build on Bolt. On Monday, we gave them open models with up to 50x usage. > Which is #1 so far? The fastest, cheapest one, with prompts up to 2x the size. > 🥇 GLM 5.3 Flash 54% 🥈 DeepSeek V4 Pro 17% 🥉 GLM 5.3 15% 👏 Kimi K3 14% >
17 September 2026 3 items
You can run Uncensored Kimi K3 locally without refusal. - Frontier MoE. - Native vision. - 1M context. - Refusals mostly gone. - EN/JA calibration. - Parent card claims 98% of several safeguard directions removed. If you already run Unsloth K3 quants, this is the abliterated twin. Not for laptops, For people who already knew that. -
@0x0SojalSec35B MoE model Run locally on your iPhone. Edge0-35B-A3B (Qwen3.6-based, 4-bit + Recover-LoRA) > - streams unused experts from SSD instead of loading the whole model. - 35B-class Qwen MoE. - Under 3GB active RAM. - 15-18 tok/s decode on macbook > -
union alpha vs kimi k3 tested both models with same prompt at highest reasoning available > union alpha took 30 minutes to make this > k3 took 10 minutes to make this stealth model looks on par with Kimi, it could actually be the kimi next mode can't do more tests now as it's almost unusable right now so gonna try again in morning for now here's output, which one did better?
@notjaziiis union alpha working for anyone? > been trying to make it work for the fast few hours and it just keep giving me error > tried it in opencode and via openrouter too but still same > worst stealth model launch ever > all i wanna do is test few of my prompts, is that too much to ask?
月之暗面推出 Kimi 金融行业解决方案,把金融数据、Skill 和 Agent 能力打包到一起。 方案接入 Wind、东方财富、标普全球、财联社、财新数据等 10+ 数据源,并提供财务建模、机构研报、财报点评、组合复盘、持仓早报等 9 项金融 Skill。 Kimi 称,合作案例中,财务建模的人力投入从 5–7 人天降至 0.5–1 人天,深度研究从 10–20 天缩短至约 2 天。 Kimi 还与中信建投共建「风险评估网关」,处理数据分级、个人信息保护、工具授权、内容核验和审计追溯。试点中,单份临时受托报告的人工制作时间从约 30 分钟降至 10 分钟。 短短一周,OpenAI、Anthropic 和 Kimi 相继推出金融行业产品。金融 AI 的竞争也从聊天和信息检索,开始深入数据、建模、报告和合规等完整工作流。
16 September 2026 3 items
🚀 vLLM's Humming backend can run Chord, @novita_labs' open-source W4A16 MoE kernels for Kimi K2.x. Kernel gains reach 1.33x on H200 TP8 vs tuned public Humming and 2.15x on B300 EP8 decode vs its untuned default. The indexed path works on compatible vLLM revisions; grouped integration is WIP. Great to see the kernels and benchmarks open-sourced! Details:
The Forge gates are open. You all came running 🏃 We’re opening up access as fast as we can, with another wave coming soon. Join the Bolt Lite waitlist, our $9/month plan with Forge access → Or skip the wait entirely. Forge is live on all Pro plans.
@boltdotnewIntroducing Bolt Forge. Free until Oct 14th: > - Up to 50x more usage - The new frontier: GLM, DeepSeek, Kimi - Zero usage charges > Live now in your model picker on > And one more thing... 👇
THIS FREE OPEN-SOURCE REPO LETS YOU RUN GLM-5.3 FLASH, DEEPSEEK V4 FLASH AND KIMI K3 LOCALLY WITHOUT A GPU OR HOSTED TOKEN QUOTAS. THE TRADEOFF: YOU’LL NEED A LOT OF STORAGE.
15 September 2026 6 items
i cannot in good conscience recommend @nahcrof anymore, i used them a lot in the past, they were an absolute steal, exceptionally reliable (for a small inference provider), atleast to my knowledge up until a few months ago they were entirely fine post Kimi K3 launch things seem to have shifted and this feels like a rugpull, i've independently confirmed the tokenizers don't match the model ids, there is an extremely detailed blogpost by @KTibow that goes further into this linked below, where all the evidence seems to check out
Kimi Code New Features 💡 Remote Control is now always on; the experimental KIMI_CODE_EXPERIMENTAL_REMOTE_CONTROL flag has been removed. 💡 Add native clipboard support on Linux X11, so copying from the TUI no longer depends on the terminal's OSC 52 support. 💡 Kimi web: Add a Plugins panel to Settings for browsing the plugin marketplace and installing, enabling, disabling, and removing plugins. See the Changelog for more technical entries.
Anyone who’s built with AI knows the annoying part: You’re finally in flow, then you start rationing prompts because every small change eats into your usage. @boltdotnew’s new Forge mode is built for exactly this. It’s an experimental mode in AI app builder with up to 50x more usage for building full-stack web apps. I took one from a prompt to a live app, then kept iterating without staring at the meter. You can also opt into Forge to help train open models. The research preview runs from Sep 14 to Oct 14, 2026. It’s free on Pro plans until Oct 14, or $9/month through the Bolt Lite early-access plan.
@boltdotnewIntroducing Bolt Forge. Free until Oct 14th: > - Up to 50x more usage - The new frontier: GLM, DeepSeek, Kimi - Zero usage charges > Live now in your model picker on > And one more thing... 👇
@Kimi_Moonshot on the verge of releasing K2.8🤩 K2.8-Preview available now in the Kimi Code CLI!
UCSD 助理教授黄碧薇创办的因果世界模型公司 Aether AI 开源 RSIAgent。 这是一套不训练模型的递归自我改进框架。底层模型参数全程固定,Agent 会自己寻找值得练习的任务、实际操作、检查结果,再把成功方法和失败教训写进长期 Memory。下一轮继续利用这些经验,逐步补上能力短板。 系统由三个 Agent 配合。Curriculum Agent 决定接下来练什么,Actor Agent 真正操作软件,Verifier Agent 独立检查结果。探索分成两步:先广泛尝试不同任务,再针对失败、隐藏限制和边界情况继续深挖。最后 Memory 会被冻结,直接拿去执行正式任务。 在 OSWorld 2.0 上,加入这套 RSI 后,平均部分得分从 71.97% 提升到 78.98%;Agents’ Last Exam 从 83.75% 提升到 84.82%。不过这不是整套测试的严格 A/B 对比。OSWorld 只有一半任务实际用了 RSI 后的新结果,其余任务继续沿用原成绩。 它和 Prime Agent 这类 Harness 自我改进思路属于同一个大方向:模型权重不变,持续更新模型外面的东西。 RSIAgent 的特点是把改进重点放在 Memory,再用自主出题和独立验证,让这份外部经验库不断积累。
@huang_biweiCan an agent explore a new environment, learn its causal structure, and keep improving without updating its model weights? > We introduce RSIAgent, a framework for recursive self-improvement through autonomous exploration. Using Kimi-K3 and GLM-5.3 as base models, RSIAgent outperforms GPT-6 Astra on both OSWorld 2.0 and Agents’ Last Exam. > RSIAgent decides what to explore, executes tasks,
next few days could get VERY interesting for AI 👀 Grok 4.7 Opus 5.1 Gemini 4 Kimi K3.1 GPT-6 Sol some of these are rumors, some have stronger signals than others but if even a few actually drop, the AI leaderboard is about to get chaotic who are you betting on?
14 September 2026 1 item
Kimi-K3 vs new Kimi-K2.8 Preview build the exact same game using both models with the same prompts i believe the whole point of releasing kimi-k2.8 after kimi-k3 was to fix the issues kimi-k3 has which is speed and cost efficiency keeping kimi-k3 as a premium option and kimi-k2.8 as mid tier option would you switch kimi-k3 for kimi-k2.8?
@TimJayaswe asked Moonshot to drop Kimi-k3.1 but they ended up dropping Kimi-k2.8 preview instead? if the whole point was to fix speed then Kimi-k3-Flash would've been a better name