NEAR AI Cloud becomes provider on OpenRouter, serving GLM 5.3 Flash with 1M token context

1 week ago 6



NEAR AI Cloud is now live as a provider on OpenRouter, giving developers access to the GLM 5.3 Flash model through the provider slug near-ai. The model comes with a 1 million token context window, and NEAR’s implementation wraps it in something most AI providers don’t offer: cryptographically verified confidential inference. What GLM 5.3 Flash actually is The model being served isn’t homegrown. GLM 5.3 Flash was developed by Z.ai and publicly launched on August 26 after an anonymous preview period called “Ox Alpha.” It’s a multimodal mixture-of-experts architecture, which means it routes different parts of a query to specialized sub-networks rather than firing up the entire model for every request. The numbers are substantial: 320 billion total parameters, with 18 billion active at any given time. It handles text, images, and video as inputs, producing text outputs. The model’s open weights are released under the MIT license. On NEAR AI Cloud, the pricing sits at $0.15 per million input tokens and $0.50 per million output tokens. The privacy angle isn’t just marketing NEAR AI Cloud runs inference inside hardware-based Trusted Execution Environments, specifically Intel TDX and NVIDI...

Read Entire Article