Observed Signal · May 28, 2026 · Funding · Source: techcrunch · Impact: 3/5 · Sentiment: Positive
General Compute raises $15M to build inference neocloud
General Compute, an inference-focused neocloud that rents AI processing for model inference (not training), raised a $15 million seed round at a $60 million post-money valuation led by FUSE VC with participation from Carya Venture Partners and Village Global Ventures. The startup plans to deploy SambaNova’s upcoming SN50 inference chips — it has $300 million of SN50s on order and says it will be the first neocloud to deploy them — arguing SambaNova’s architecture offers higher inference throughput (company claims 600–700 tokens/sec vs ~250 for GPUs). The chips are air-cooled and lower-power, enabling installation in existing data centers; General Compute is pursuing colocation deals including partnerships with crypto miners. The company launched its cloud offering and claims top performance on the open-source MiniMax 2.7 LLM.
A startup funding and major chip order tied to a new inference-focused cloud highlights shifts in AI inference infrastructure and supply choices beyond GPUs; could affect availability, cost and competitive dynamics in the inference cloud market.
Track CoreWeave Signals & Market Shifts in Real-Time
Polaris7 autonomous intelligence agents track regulatory filings, primary sources, executive changes, and deal flow 24/7. Create your free Explorer workspace to monitor these entities.
Key Takeaways & Evidence Grounding
- General Compute raised a $15 million seed round at a $60 million post-money valuation.
- The seed round was led by FUSE VC with participation from Carya Venture Partners and Village Global Ventures.
- General Compute has $300 million of SambaNova SN50 inference chips on order and says it will be the first neocloud deploying them.
- SambaNova claims its upcoming chips can deliver 600–700 tokens per second for inference versus ~250 tokens per second for GPUs.
- General Compute launched its cloud offering and claims it is the fastest at running the open-source MiniMax 2.7 LLM.
Connected Companies & Entities
4 Entities mappedOntology Mapping & Concepts
Related Market Signals & Shifts
Recent verified developments and strategic activity across this market segment.
Startup Secures $400M Loan Using Inference Chips
General Compute, an AI inference cloud startup founded by CEO Finn Puklowski, has received a $400 million loan from tech investor Upper90. The deal may be among the first to use inference-specific chips as collateral — chips optimized to run trained AI models efficiently rather than to train them. General Compute plans an inference neocloud built around SambaNova silicon and says its SN50 chips will deliver faster, more power-efficient inference than GPU-based clouds. Upper90, led by co-founder and CEO Billy Libby, has prior experience financing advanced chips and is applying that playbook to inference-focused infrastructure as demand grows for lower-cost open-source model deployments.
Gimlet Labs raises $80M for multi-silicon inference cloud
Gimlet Labs, founded by Stanford adjunct professor and founder Zain Asgar with cofounders Michelle Nguyen, Omid Azizi and Natalie Serrino, raised an $80 million Series A led by Menlo Ventures to commercialize what it calls a "multi-silicon inference cloud." The software orchestrates AI workloads across diverse hardware (CPUs, AI GPUs, high-memory systems), claims 3x–10x inference speedups for the same cost and power, and can slice models to run across different architectures. Gimlet has partnerships with chip makers NVIDIA, AMD, Intel, ARM, Cerebras and d‑Matrix, offers its product as software or via an API/Gimlet Cloud, and targets large model labs and data centers. The company reported eight-figure revenues at launch, has roughly 30 employees, and has now raised $92 million in total including prior seed and angel investments.
SpaceX Neocloud Estimated at $28B/Year
A Latent Space AI News roundup highlights large GPU compute contracts and multiple AI infrastructure developments. Social posts indicate SpaceX signed a third large GPU rental deal — reportedly a $6.3B compute agreement with Reflection that will pay $150M per month from July 1, 2026 through 2029 — and, when combined with other publicized deals (Anthropic, Google), yields a tally of roughly $2.32B per month (~$28B/year). The piece contrasts that scale with CoreWeave’s current revenue/valuation and also summarizes major AI product and infra news: OpenAI’s Daybreak expansion and GPT-5.5-Cyber for security remediation, Sakana’s Fugu orchestration release and accompanying critical debate, GLM-5.2’s traction as an open-weight frontier-adjacent model, and Google promoting the Gemini Interactions API to GA. Baseten’s announced $13B Series F and the broader move toward post-training and compute leasing are noted as market signals for inference and owned intelligence trends.
Track Real-Time Market Signals & Shifts
Set up custom watchlists to receive automated, evidence-grounded executive digests whenever material signals or shifts occur across your tracked landscape.
