Observed Signal · Sep 23, 2026 · Product Launch · Source: SemiAnalysis · Impact: 3/5 · Sentiment: Positive
SemiAnalysis Releases ClusterMAX 3.0 GPU Cloud Ratings
SemiAnalysis published ClusterMAX 3.0, the third edition of its industry-standard GPU cloud rating system for neoclouds. The report covers 77 providers, expanding its market view to 323 providers. Nebius joined CoreWeave in the Platinum tier, while Google Cloud moved up to Gold. The testing methodology includes audits of compute, networking, storage, orchestration, reliability, and security through hands-on benchmarks and failure injections. Key trends include financing structures, Blackwell/GB300 deployments, the transition to Vera Rubin, and the impact of agentic coding on cluster management. The report also introduces standardized SLAs and highlights security concerns across neoclouds.
This is a significant industry report card for the neocloud sector, covering major providers and trends. It provides detailed evaluations and could influence customer decisions.
Marktsignale zu SemiAnalysis in Echtzeit verfolgen
Polaris7 erfasst behördliche Registrierungen, Primärquellen, Führungswechsel und Deal-Aktivitäten rund um die Uhr. Erstellen Sie Ihren kostenlosen Explorer-Workspace, um automatisierte Executive Briefings zu erhalten.
Wichtigste Kernpunkte & Evidenz
- ClusterMAX 3.0 was released covering 77 providers, with the market view expanded to 323 providers from 209 in ClusterMAX 2.0.
- Nebius was upgraded to Platinum tier, joining CoreWeave, while Google Cloud moved to Gold.
- Only 19 neoclouds globally achieved a Medallion rating in ClusterMAX 3.0.
- SemiAnalysis introduced a new 'Participation Ribbon' tier for 15 providers that do the bare minimum.
- CoreWeave announced a VR200 NVL72 system passing L11 diagnostics, and a multi-year agreement with Anthropic.
Verknüpfte Unternehmen
31 verknüpfte Unternehmen“SemiAnalysis published ClusterMAX 3.0, the industry standard GPU cloud rating system....”
“Nebius joins CoreWeave in the Platinum tier after strong business decisions....”
“Google Cloud joins Oracle in the Gold tier....”
“Azure moves to Silver tier in the ClusterMAX 3.0 rankings....”
“Microsoft is mentioned as a customer of Nebius and partner with CoreWeave....”
“CoreWeave remains in the Platinum tier and sets the technical bar....”
“Meta is mentioned as a customer of Nebius and CoreWeave....”
“Anthropic is mentioned as a frontier lab and a customer of neoclouds....”
“Google completed the acquisition of Wiz for $32B....”
“Jane Street signed a $6B AI cloud agreement with CoreWeave....”
“Palantir partnered with Nebius....”
“Oracle was tested and remains in the Gold tier....”
“Together AI's clusters were plagued with reliability issues....”
“OVHcloud has substantial infrastructure globally....”
“Nscale announced a $2B Series C and acquisition of AnyScale....”
“Andromeda takes capacity from other providers and puts a managed cluster service on top....”
“Akamai/Linode has no managed Slurm or Kubernetes available....”
“DigitalOcean markets itself as 'the first cloud built end-to-end for the inference and agentic era'....”
“Blackstone announced a $5B joint venture with Google for TPU cloud....”
“NVIDIA is mentioned as a major technology provider and investor....”
“Google Cloud signed a multi-billion-dollar deal with Thinking Machines Lab....”
“AMD is mentioned as a chip provider for some neoclouds....”
“Alibaba has an ACK Slurm operator....”
“Groq pivoted to renting Nvidia GPUs....”
“Google Cloud landed a $10B deal with Palo Alto Networks....”
“Prime Intellect closed a $130M Series A....”
“Naver operates hyperscaler-class datacenters....”
“Mithril is a marketplace wrapping 3 Nebius availability zones....”
“OpenAI is mentioned as a frontier lab that manages its own infrastructure....”
“Mistral raised a €3B Series D....”
“IBM Cloud is mentioned as a provider with a horrendous experience....”
Ontology Mapping & Concepts
Verwandte Marktsignale & Trends
Aktuelle verifizierte Unternehmensentwicklungen und Deal-Aktivitäten in diesem Marktsegment.
Samsung investiert 1 Mrd. Dollar in KI-Infrastruktur-Firma Helix
Samsung Electronics und fünf Tochtergesellschaften investieren gemeinsam 1 Milliarde US-Dollar in Helix Digital Infrastructure, ein von KKR gegründetes und von Nvidia unterstütztes KI-Infrastrukturunternehmen. Samsung Electronics steuert 500 Millionen US-Dollar bei, den Rest übernehmen Samsung C&T, Samsung SDS, Samsung SDI, Samsung Life Insurance und Samsung Fire & Marine Insurance. Helix, geführt vom ehemaligen AWS-CEO Adam Selipsky, konzentriert sich auf Hyperscale-Rechenzentren, Stromerzeugung, Übertragung und Glasfasernetze. Die Investition ergänzt die bereits zugesagten über 10 Milliarden US-Dollar anderer Investoren wie KKR, Kuwait Investment Authority, Nvidia und Vistra. Der Schritt ermöglicht es Samsung, seine Halbleiter-, Kühlungs-, Rechenzentrumsbau- und Batteriekompetenzen zu nutzen, um im KI-Infrastrukturmarkt zu expandieren.
Sparse Attention in GLM-5.3 und ihre Auswirkungen auf DRAM-Speicher
Dieser Artikel analysiert die Auswirkungen von Sparse-Attention-Mechanismen, insbesondere DeepSeek Sparse Attention (DSA), die im GLM-5.3-Modell von Z.ai verwendet werden, auf den adressierbaren Gesamtmarkt (TAM) für DRAM-Speicher, einschließlich HBM und NAND. Es wird erklärt, dass Sparse Attention zwar KV-Cache-Speicher und Bandbreite während der Attention-Operation reduziert, jedoch nicht den Gesamtspeicherbedarf, da die Top-k-Auswahl den vollständigen Kontext im HBM erfordert. Der Artikel diskutiert Systemoptimierungen wie HiSparse, das KV-Cache auf Host-DRAM auslagert, um Kapazitätsengpässe zu überwinden. Er enthält auch detaillierte Leistungs- und Kostenvergleiche für das Serving von GLM-5.3 auf verschiedenen Hardware-Plattformen (GB200, GB300, MI355X) mit Inferenz-Engines wie Dynamo-SGLang, Dynamo-TRT-LLM und ATOM, wobei Kosteneffizienz und Interaktivitäts-Kompromisse hervorgehoben werden. Die Analyse umfasst einen tiefen Einblick in die Architektur von GLM-5, einschließlich Lightning Indexer, MLA-Konfiguration und Post-Training-Pipeline.
Alibabas T-Head-Chips: Cloud-Kunden oder Qwen-Training?
Alibaba kündigte auf seiner Apsara-Konferenz an, dass sein neuer KI-Chip Zhenwu V900 im ersten Quartal 2027 in Massenproduktion gehen und verkauft wird, zwei Quartale früher als geplant. Dies folgt auf Huaweis Ankündigung, dass sein Ascend-960DT-Chip im ersten Quartal 2027 fertig sein wird. Beide Unternehmen sehen sich einer hohen Nachfrage und begrenzten Angebot für ihre Chips gegenüber. IDC-Daten zeigen, dass Nvidia 55 % der chinesischen Server-KI-Beschleuniger-Lieferungen hält, Huawei 20 % und T-Head 7 %. Alibaba plant, Qwen-Modelle mit 5-10 Billionen Parametern zu trainieren, hat aber nicht offengelegt, welche Chips verwendet werden, was Bedenken über Konkurrenz zwischen internem Modelltraining und zahlenden Cloud-Kunden um knappe Chip-Kapazitäten aufwirft.
Marktsignale & Strategische Shifts in Echtzeit verfolgen
Erstellen Sie benutzerdefinierte Watchlists, um automatisierte, evidenzbasierte Executive Briefings zu erhalten, sobald wesentliche Signale oder Marktverschiebungen auftreten.
