Observed Signal · Aug 25, 2026 · Partnership · Source: https://martechseries.com/feed/ · Impact: 4/5 · Sentiment: Positive

Cisco Expands Secure AI Factory with NVIDIA for Rack-Scale

Executive Signal Summary

Cisco announced an expansion of its Secure AI Factory through a partnership with NVIDIA and Supermicro to deliver rack-scale, high-density AI computing solutions for enterprises, neoclouds, and sovereign clouds. The offering adds Supermicro liquid- and air-cooled servers validated as part of Cisco’s full-stack Secure AI Factory architecture, supports trillion-parameter training and high-throughput inference, and is NVIDIA Cloud Partner (NCP) compliant. Cisco says the solution integrates Cisco Silicon One and NVIDIA Spectrum‑X networking, Cisco Nexus One orchestration, and new Cisco Validated Infrastructure Services aligned with NVIDIA Infrastructure Services to simplify deployment, improve observability, and reduce deployment risk. Executives from Cisco and NVIDIA framed the expansion as enabling faster time-to-value, stronger security, and operational simplicity for large-scale AI deployments.

Polaris7 AgentPolaris7 Strategic Assessment
High Confidence

Major infrastructure partnership (Cisco + NVIDIA + Supermicro) expands validated, rack-scale AI architectures for enterprise and sovereign cloud deployments, accelerating large-model training/inference adoption and reducing deployment risk; important for enterprise AI infrastructure planning.

SIGNAL RADAR

Track Cisco Signals & Market Shifts in Real-Time

Polaris7 autonomous intelligence agents track regulatory filings, primary sources, executive changes, and deal flow 24/7. Create your free Explorer workspace to monitor these entities.

Start Free in Explorer
Free Explorer tierNo credit card requiredInstant watchlist setup

Key Takeaways & Evidence Grounding

  • Cisco announced an expansion of its Secure AI Factory through a partnership with NVIDIA and Supermicro.
  • Cisco will offer Supermicro liquid- and air-cooled rack-scale GPU systems validated and sold as part of Cisco’s AI infrastructure portfolio.
  • The solution is NVIDIA Cloud Partner (NCP) compliant and targets enterprise, neocloud, and sovereign cloud customers for trillion-parameter training and high-throughput inference.
  • Cisco integrates Cisco Silicon One front-end switches and NVIDIA Spectrum-X based back-end switches unified by Cisco Nexus One in the reference architecture.
  • Cisco introduced Cisco Validated Infrastructure Services (CVIS) aligned with NVIDIA Infrastructure Services (NVIS) and is investing in a large-scale AI Lab to certify and test AI infrastructure.

Connected Companies & Entities

4 Entities mapped

“Today, Cisco is expanding the Secure AI Factory with NVIDIA, through a partnership with Supermicro, bringing its leadership in high-density,...”

““AI factories are revenue-generating infrastructure, where compute produces intelligence, and intelligence drives revenue,” said Justin Boit...”

“PR Newswire, a Cision company, is the premier global provider of multimedia platforms and distribution that marketers, corporate communicato...”

“MarTech Series is a leading publishing platform that provides daily updates on marketing technology news, in-depth interviews with industry ...”

Primary Source Grounding & Direct Attribution
Direct Origin Attribution
Primary Reporting: https://martechseries.com/feed/•Published: Aug 25, 2026
Original Coverage Title: “Cisco Expands Secure AI Factory with NVIDIA for the Rack-Scale Era”

Related Market Signals & Shifts

Recent verified developments and strategic activity across this market segment.

InfrastructureSep 29, 2026

Samsung to invest $1B in AI infrastructure firm Helix

Samsung Electronics and five affiliates will invest a combined $1 billion in Helix Digital Infrastructure, an AI infrastructure company launched by KKR and backed by Nvidia. Samsung Electronics contributes $500 million, with the rest from Samsung C&T, Samsung SDS, Samsung SDI, Samsung Life Insurance, and Samsung Fire & Marine Insurance. Helix, led by former AWS CEO Adam Selipsky, focuses on hyperscale data centers, power generation, transmission, and fiber-optic networks. The investment adds to over $10 billion already committed by other investors including KKR, Kuwait Investment Authority, Nvidia, and Vistra. The move allows Samsung to leverage its semiconductor, cooling, data center construction, and battery capabilities to expand in the AI infrastructure market.

Read assessment
InfrastructureSep 28, 2026

Modal Labs closing in on $750M round at $15.75B valuation

AI inference infrastructure provider Modal Labs is nearing a $750 million funding round led by Accel at a $15.75 billion valuation, according to a source. This would more than triple its valuation from $4.65 billion in May. The round comes amid surging demand for inference services, with other startups like Baseten, Fireworks, and Fal also raising at higher valuations. Modal Labs, founded in 2021 by Erik Bernhardsson and Akshat Bubna, provides infrastructure for training and running AI models without managing servers. The company has surpassed $300 million in annualized revenue as of May. The funding talks follow a security incident in July where a customer's data was compromised, but Modal's platform was not breached.

Read assessment
AI InfrastructureSep 28, 2026

GLM-5.3 Sparse Attention Impact on DRAM Memory TAM

This article analyzes the impact of sparse attention mechanisms, specifically DeepSeek Sparse Attention (DSA) used in Z.ai's GLM-5.3 model, on the total addressable market (TAM) for DRAM memory, including HBM and NAND. It explains that while sparse attention reduces KV cache memory and bandwidth during the attention operation, it does not reduce overall memory capacity requirements because the top-k selection still requires full context in HBM. The article discusses system optimizations like HiSparse, which offloads KV cache to host DRAM to overcome capacity bottlenecks. It also provides detailed performance and cost comparisons for serving GLM-5.3 on different hardware (GB200, GB300, MI355X) using inference engines like Dynamo-SGLang, Dynamo-TRT-LLM, and ATOM, highlighting cost-efficiency and interactivity trade-offs. The analysis includes a deep dive into GLM-5's architecture, including the lightning indexer, MLA configuration, and post-training pipeline.

Read assessment

Track Real-Time Market Signals & Shifts

Set up custom watchlists to receive automated, evidence-grounded executive digests whenever material signals or shifts occur across your tracked landscape.