Moonshot AI in Revenue-Sharing Talks with Microsoft, Amazon, and Google to Host Kimi K3

Beijing-based artificial intelligence developer Moonshot AI is negotiating revenue-sharing partnerships with Microsoft, Amazon, and Alphabet's Google to host its flagship open-weight model, Kimi K3, across major cloud platforms. According to reporting from Reuters, the startup is seeking up to a 30 percent share of revenue generated from Kimi K3 inference services hosted on Microsoft Azure, Amazon Web Services (AWS), and Google Cloud. Moonshot released Kimi K3 in July 2026 as a 2.8-trillion par

2 min
Moonshot AI in Revenue-Sharing Talks with Microsoft, Amazon, and Google to Host Kimi K3

Beijing-based artificial intelligence developer Moonshot AI is negotiating revenue-sharing partnerships with Microsoft, Amazon, and Alphabet's Google to host its flagship open-weight model, Kimi K3, across major cloud platforms. According to reporting from Reuters, the startup is seeking up to a 30 percent share of revenue generated from Kimi K3 inference services hosted on Microsoft Azure, Amazon Web Services (AWS), and Google Cloud.

Moonshot released Kimi K3 in July 2026 as a 2.8-trillion parameter open-weight model featuring a 1-million-token context window. While the model weights are publicly downloadable for research and self-hosting, the sheer scale of a multi-trillion parameter architecture makes private infrastructure hosting cost-prohibitive for most enterprise organizations, establishing hyperscaler cloud catalogs as the principal distribution channel for enterprise adoption.

Enterprise Cloud Inference and Telemetry Architecture

Open-Weight Licensing and Commercial Monetization

Moonshot's commercial strategy relies on provisions embedded in Kimi K3's licensing terms. Under the model's license agreement, any commercial provider offering Kimi K3 as a managed service and generating over $20 million in annual gross revenue must establish a formal commercial agreement with Moonshot.

The proposed 30 percent revenue split aligns with the licensing tiers Moonshot has presented to enterprise software vendors and regional infrastructure operators. Chinese IT services provider Chinasoft International previously disclosed a commercial revenue-sharing partnership with Moonshot in a regulatory filing, though the precise split was not publicly detailed.

Key points currently under discussion between Moonshot and US cloud providers include:

  • Revenue distribution percentages across managed API calls and dedicated instance hosting.
  • Verification mechanisms and data audit protocols for tracking token consumption across third-party cloud billing systems.
  • Compliance and governance boundaries regarding enterprise data security and regional sovereign hosting requirements.

Hyperscaler Economics for Trillion-Parameter Models

The negotiations reflect a shifting dynamic between foundational model developers and cloud infrastructure giants. As frontier open-weight models reach trillions of parameters, model creators are using tiered commercial licenses to capture direct monetization from the massive inference volumes generated on public cloud platforms.

Moonshot, which has secured backing from Alibaba and Tencent and is preparing for a potential Hong Kong initial public offering, has also explored similar revenue-sharing arrangements with regional cloud operators. The discussions with US cloud providers remain in early stages, with formal commercial terms still subject to negotiation.

Sources

Written by

More to read

  • LLM Observability and Tracing in Production: Comparing Langfuse, Arize Phoenix, OpenLLMetry, and Helicone Architecture, OpenTelemetry Ingestion, Eval Pipelines, and Serving Economics

    Tracing multi-step LLM pipelines, autonomous agent graphs, and retrieval-augmented generation (RAG) systems in production introduces telemetry challenges that traditional Application Performance Monitoring (APM) tools cannot address out of the box. While standard microservices rely on CPU utilization, HTTP status codes, and network latency percentiles, LLM workflows require deep inspection into non-deterministic text generation, nested execution graphs, prompt token counts, retrieved context rel

    1 min
  • Huawei Proposes 2,000 Ascend 950 AI Chips for Egyptian Government Cloud in Key Export Test

    Huawei Technologies has submitted a proposal to build sovereign artificial intelligence infrastructure for the Egyptian government, offering to export more than 2,000 of its proprietary Ascend AI accelerators. The tender represents China's most significant known push to export its highest-end AI silicon to international public sector clients. The proposal has drawn immediate attention in Washington, prompting the U.S. State Department to contact American semiconductor and cloud providers to ass

    1 min
  • Meta Explored Slashing Teams by Up to 60% in AI-Native Shift Before Agent Failures Forced Retreat

    Internal planning documents and reporting revealed that Meta explored cutting team headcounts by up to 60% as part of an initiative code-named Project OT (Organization Transformation), designed to shift the company into an "AI-native" operating structure where small pods of engineers would oversee autonomous AI agents. The initiative unraveled following internal workforce pushback and operational data demonstrating that generative AI agents caused severe reliability problems while failing to de

    1 min