Cloud11 articles

Cloud

Articles

  • Anthropic Agrees to 5 Billion Cloud Deal with Nscale for 460MW of Vera Rubin Compute

    Anthropic has finalized a six-year, $45 billion cloud computing agreement with AI infrastructure provider Nscale. Under the terms of the deal, Anthropic will secure approximately 460 megawatts of dedicated computing capacity at Nscale's Monarch data center development in West Virginia, scheduled to come online in late 2027. The deployment will be powered by Nvidia's upcoming Vera Rubin architecture, providing compute bandwidth for next-generation foundation model training and enterprise inferen

    1 min
  • Moonshot AI Seeks Up to 30% Revenue Share from Microsoft, Amazon, and Google to Host Kimi K3

    China-based artificial intelligence startup Moonshot AI is in early negotiations with Microsoft, Amazon Web Services (AWS), and Google Cloud regarding revenue-sharing agreements to host its open-weight Kimi K3 model on their respective cloud platforms, according to a report from Reuters. According to people familiar with the matter, Moonshot is seeking up to a 30% cut of all revenue generated from hosting and serving Kimi K3 on the US hyperscaler platforms. Commercial Licensing Clauses on Ope

    1 min
  • AWS and NVIDIA Expand AI Partnership to Deploy 2 Million Additional Blackwell Ultra and Rubin GPUs

    Amazon Web Services (AWS) and NVIDIA have announced a major expansion of their cloud infrastructure partnership, committing to deploy two million additional high-end NVIDIA GPUs across AWS global data centers in 2027 and 2028. The deployment expands on AWS's previous commitment from GTC 2026 to add one million GPUs starting in 2026, bringing total forward allocations across the multi-year cycle to three million units. The upcoming capacity will comprise NVIDIA Blackwell Ultra, Rubin, and Rubin

    1 min
  • Amazon Acquires DuckLabs to Integrate DuckDB into AWS Analytics and AI Agent Workflows

    Amazon has entered into a definitive agreement to acquire DuckLabs, the Amsterdam-based company behind the open-source columnar database DuckDB. The acquisition brings the DuckLabs development team into Amazon Web Services (AWS), where they will operate as a wholly owned subsidiary starting in early September. Financial terms of the transaction were not disclosed. DuckDB creators and DuckLabs co-founders Hannes Mühleisen and Mark Raasveldt will continue leading the team from Amsterdam, maintain

    1 min
  • Moonshot AI in Revenue-Sharing Talks with Microsoft, Amazon, and Google to Host Kimi K3

    Beijing-based artificial intelligence developer Moonshot AI is negotiating revenue-sharing partnerships with Microsoft, Amazon, and Alphabet's Google to host its flagship open-weight model, Kimi K3, across major cloud platforms. According to reporting from Reuters, the startup is seeking up to a 30 percent share of revenue generated from Kimi K3 inference services hosted on Microsoft Azure, Amazon Web Services (AWS), and Google Cloud. Moonshot released Kimi K3 in July 2026 as a 2.8-trillion par

    1 min
  • AWS Backs Open Agentic Resource Discovery Specification for Agent Registry Federation

    Amazon Web Services announced support for the Agentic Resource Discovery (ARD) open specification, detailing how the federation standard will integrate with AWS Agent Registry, its managed catalog for AI agents, tools, and skills currently in preview within Amazon Bedrock AgentCore. The alignment targets cross-platform discovery across heterogeneous enterprise stacks. While AWS Agent Registry provides centralized indexing and access controls within an AWS environment, real-world deployments fre

    1 min
  • Groq Secures 50M at .5B Valuation to Expand Nvidia-Powered AI Neocloud

    AI infrastructure provider Groq has raised $350 million in a Series A funding round at a $3.5 billion valuation, led by investment firm Disruptive with expected participation from Nvidia subject to customary closing conditions. The financing accelerates the company's structural pivot from developing custom inference silicon toward operating an enterprise-grade inference cloud powered by Nvidia accelerated computing systems. The round follows a $650 million capital raise completed in June 2026 a

    1 min
  • Meta Emerges as Major Microsoft Azure AI Customer with Multi-Hundred-Million-Dollar Spend

    Meta Platforms has emerged as one of Microsoft Azure's largest artificial intelligence customers, spending hundreds of millions of dollars annually to access hosted AI models and inference compute, according to reporting by Bloomberg. The multi-hundred-million-dollar commitment underscores how current commercial demand for large-scale AI infrastructure remains intensely concentrated among frontier technology companies themselves. Bridging Internal Compute Gaps with Third-Party Infrastructure

    1 min
  • Anthropic Modifies Enterprise Data Retention to Allow Customer Cloud Logging for Frontier Models

    Anthropic is preparing to revise the mandatory 30-day data retention requirement on its frontier models, allowing enterprise customers to retain logs on their own cloud infrastructure rather than storing conversation records on Anthropic servers. According to reporting from Bloomberg and Reuters, the upcoming safety architecture preserves the 30-day logging mandate for safety audits and abuse monitoring while shifting physical custody of the stored data into customer virtual private clouds. E

    1 min
  • LLM Autoscaling and Cold Starts in Kubernetes: Architecture, KEDA Metrics, Model Weight Caching, and Ephemeral GPU Provisioning

    Autoscaling large language model workloads on Kubernetes presents a fundamentally different engineering problem than traditional stateless microservices. While web APIs scale on CPU utilization or request rate within seconds, LLM inference instances require specialized GPU accelerators, massive container images, multi-gigabyte weight tensors, and intensive runtime compilation before serving a single token. Without proactive architectural design, a cold-starting LLM pod on Kubernetes often requi

    1 min
  • Grok 4.6 Launches on Amazon Bedrock with 500K Context and Cross-Region Routing

    xAI's flagship reasoning model, Grok 4.6, is now generally available across Amazon Web Services through Amazon Bedrock. Released on August 19, 2026 under the model ID xai.grok-4.6, the deployment gives enterprise AWS customers managed API access to xAI's frontier model alongside existing foundational offerings from Anthropic, Meta, and Mistral. The integration comes one week after xAI initially launched Grok 4.6 on August 12, marking a significantly faster enterprise cloud deployment than its p

    1 min