Skip to content
B Bloger.fm

Organisation

OpenRouter

A routing service that exposes many models from many providers behind one API, which makes it a common way to compare them.

OpenRouter was recorded in 14 items across 7 of the 8 briefings in the current window.

Its share of coverage was steady: 5 items in the first half of the window and 9 in the second, tracking the feed as a whole, which grew about 2.2×.

It appeared most often alongside OpenCode, Claude Code and Codex.

Tracking the feed
items
14
briefings
7
mentions
27
last seen
2026-09-19
Official channel
openrouter.ai

Coverage timeline

Sat 12 Sept – Sat 19 Sept / 8 briefings

Everything recorded

19 September 2026 1 item

truncated at source

A significant benefit is the ability to compare different agents and models when performing the same task. While performance is important, cost is also a crucial factor. Identifying a model that achieves the desired outcome at a considerably lower price point provides the kind of adaptability required by users of artificial intelligence.

@quxiaoyin

We just launched Agentsky @agentsky_dev, world’s 1st Agent Market! OpenRouter is for models. AgentSky is for agents. > Use 40+ agents—Claude Code, Codex, OpenCode, Hermes, Pi in your browser(even your phone!) or via one API. All without installing or setting up anything. > Hit Codex Astra’s weekly limit? Hand off to another agent such as OpenCode + DeepSeek V4.1 in browser without losing any context. > You can compare any agent + model directly in browser and that's how I found Astra costs $5.3 while deepseek v4.1 cost $0.12 on the same dashboard task. (I actually preferred deepseek) > Try it at

18 September 2026 1 item

17 September 2026 6 items

She’s right. This cheapens a thing that is actually really cool (stealth drops) The fact that they made a twitter account for the stealth model is super “fellow kids” core

@maria_rcks

very disappointed all of the model providers (opencode, openrouter, whatever) > released 'union alpha' without telling us it was a crappy model router > kudos to the cloudflare guys for telling people what they're getting > AND FUCK union alpha

Significantly more capacity for Union Alpha is coming online tomorrow morning Thanks for bearing with the hiccups in the meantime!

@OpenRouter

🥷 New stealth model: Union Alpha (@unionalphaai) > A multimodal model for research, coding, and agentic workflows. > - Free to use - 256K context - Tool calling - Frontier-level general-purpose performance > Try it now and share your feedback:

1/ Any model on OpenRouter can now run code in a hosted Linux container. Add one tool to a Responses API request. The model writes the script, runs it, and reads back the result. Nothing to install, nothing to host. Meet openrouter:shell. To get started: "tools": [{ "type": "openrouter:shell", "parameters": { "engine": "openrouter" } }]

union alpha vs kimi k3 tested both models with same prompt at highest reasoning available > union alpha took 30 minutes to make this > k3 took 10 minutes to make this stealth model looks on par with Kimi, it could actually be the kimi next mode can't do more tests now as it's almost unusable right now so gonna try again in morning for now here's output, which one did better?

@notjazii

is union alpha working for anyone? > been trying to make it work for the fast few hours and it just keep giving me error > tried it in opencode and via openrouter too but still same > worst stealth model launch ever > all i wanna do is test few of my prompts, is that too much to ask?

truncated at source

What happens when your AI model becomes good enough to build its own infrastructure? Zhipu just found out. Their AI model GLM-5.3 helped build and optimize the inference system that serves GLM-5.3-Flash to users. The model improving the system that runs the model. 100,000+ Chinese-made AI accelerators. Nobody had deployed at this scale on that hardware before. Limited memory. Incomplete ecosystem. Most of it undocumented. The Infra Agent powered by GLM-5.3 did the engineering work. Found bugs in kernels. Fixed concurrency bottlenecks. Studied optimization patterns from other codebases and applied them to its own inference. First successful run to production ready in two weeks. Throughput tripled. Then it went live anonymously as "Ox-Alpha" on OpenCode and OpenRouter. Became the most-used model on both platforms in six days. 62 trillion tokens processed. Zhipu's own words: "The model optimizes the system. The system runs the model." They added: "We have not yet reached full recursive self-improvement. B

Stealth models in 2026: Hunter Alpha → Xiaomi MiMo-V2 Owl Alpha → Meituan LongCat Pony Alpha → GLM-5 Ox Alpha → GLM-5.3-Flash Now Union Alpha is free for a week on OpenCode and OpenRouter.

16 September 2026 1 item

Ox Alpha: The Most Brilliant Marketing Move In AI This Year 👀 >No logo, no branding, no PR just showed up on OpenRouter on Aug 20 as an anonymous "stealth model" >1M-token context, multimodal (text, images, video), zero data retention >Free for a week, with a claimed 100 trillion tokens/day of capacity behind it >Beat GPT-5.6-Sol and Claude Fable 5 on the DeepSWE coding benchmark >Usage blew past DeepSeek by more than 2x in days Stripe's CEO even called it "very impressive"

15 September 2026 1 item

太魔幻了,Astra 还在让大家惊叹:AI 终于会用 Blender 了。 结果Nex-AGI 已经直接出现在现实世界里了。 我只给了一个 Prompt: 做一根 6cm 的粉色活动香蕉,省点料。 然后我基本没碰 Blender。 它自己用 MCP 建模,Computer Use 看结果、自己修,最后直接吐给我一个 STL。 我顺手扔进拓竹。 👇 视频里正在一层一层长出来的,就是它自己设计的香蕉。 这一刻我是真有点惊到了。 Nex-AGI + Computer Use,能力完全超出我预期。 一句话 → Blender → STL → 3D 打印 → 实物。 🔥 这真的有点科幻了。 直接免费体验nex-agi :

13 September 2026 2 items

truncated at source

Nex-N2.5 feels less like “another model launch” and more like a push toward agents that can actually finish real work. The part that stands out to me is Nex-N2.5 Pro: a 397B multimodal model built for Vision, Computer Use, and long-horizon interaction. Instead of only understanding a screen, it can keep reading changing UI states and continue operating software step by step. That means workflows like: → moving from an idea to a Blender scene → operating CAD and other professional tools → navigating browsers and software end to end → writing code, running it, checking the interface, fixing issues, and verifying again The brief even shows examples spanning games, Blender, FreeCAD, LabPlot, GeoGebra, and other real desktop workflows. Nex-N2.5 comes in Mini, Pro, and Max, with Mini and Pro focused on multimodal agent capabilities and Max using a 1.6T text-only MoE base for more complex agentic reasoning and coding. The direction is clear: AI agents are moving from “tell me what to do” toward “open the tool a

Amp 推出免费 Hobby 档。 Amp 是从 Sourcegraph 拆出来的 Coding Agent 公司,产品和 Claude Code、Codex 属于一类,可以让 AI 自己读代码、改代码、跑命令和测试。现在只要用自己的电脑,再接自己的 ChatGPT 订阅或模型 API Key,就能免费使用 Amp,不再额外交月费和 BYOK Token 费。 今年 7 月 Amp 推出订阅制后,想把自己的 ChatGPT 订阅接进 Amp,至少要开每月 20 美元的 Megawatt。BYOK 即便用的是自己的模型 Key,Amp 也会收额外 Token 费用并设限制。现在这两层都取消了。ChatGPT 订阅费或 API 本身的模型费用,还是用户自己承担。 Amp 还有一个叫 Orb 的远程云电脑产品,可以给每个任务开一套独立环境。代码和开发工具都放在里面,关掉自己的电脑后 Agent 还能继续干活,也方便同时跑多个任务。Orb 依然按使用量收费;不想付这笔钱,也可以让 Amp 直接跑在自己的电脑上。 另外,Amp 还在扩大 BYOK 范围。付费用户已经可以接 OpenRouter、Amazon Bedrock、Azure Foundry、Ollama Cloud 和自定义模型接口等,之后会开放给所有用户。

@sqs

Amp is now free to use when you bring your own compute and model subscriptions/keys. > No more limits or fees for BYOK. >

12 September 2026 2 items

truncated at source

This is happening “the boon in using open models for cost savings” is real. Just four months ago we launched @CommandCodeAI coding agent built specifically for open models. 60T tokens scale and 43K paying active customers later, more than 90% of our usage is open models. In the first 24hrs of DeepSeek V4.1 launch, Command Code processed 3.1T paid tokens that’s 3x more than entire market on OpenRouter. 1T tokens on SOTA Claude/GPT cost you ~$5M 1T tokens on SOTA open models cost you ~$50K I’m not joking. I’m literally looking at our data. These numbers are real. That’s 100 times cheaper. While nearly as good. Run twice as many. I believe top ten open models collectively beat single SOTA Claude/GPT model. And when you discover a workflow that works for you, there’s literally nothing stopping you. You are not bound by subscription subsidies. Regular API prices are ten times or more manageable. In near future, enterprise will wake up to this. Better, faster, cheaper, private, and available now

truncated at source

DAILY AI BRIEF 🗞 — Sept 12 OPENAI 🔥: - GPT-Rosalind is out of research preview for eligible orgs worldwide. API, Codex, and ChatGPT Enterprise, with new Rosalind models as they ship. - Codex adds Life Sciences plugins for genomes, protein structure, QC reports, and notebooks. - ChatGPT Sites hit 5M apps. New: collab editing, private invites, custom domains, and DB inspect. - Desktop pets can start a new chat. Mini is the compact no-pet option. XAI 🔥: - Elon: Grok 4.7 needs a few more days. RL still quits hard tasks too early and undershoots self-checks. - Grok Bot is rolling out on Grok web for Heavy users. Create and chat with bots in the UI. No official post yet. ANTHROPIC 🔥: - Claude Code ships `claude plugin eval`. Score a plugin on test cases, then rerun without it. MICROSOFT 🔥: - MAI-Transcribe-2 hit 1M OpenRouter requests in 5 days. ALIBABA 🔥: - Qwen3.8-27B is live on Cerebras. * Used Grok to compose this brief, cherry-picking the news and doing some post-editing. > [@testingcatalog](