V

Vector Core Compute

by Vector Core Compute (VC2)

Infrastructure & CloudEnterprise Platform

Disaggregated enterprise inference cloud running CPUs, GPUs and RDUs in one pipeline

Contact for pricing·Added Jul 31, 2026·Updated Jul 31, 2026
Share:
THE DAILY BRIEF
Vector Core Compute

by Vector Core Compute (VC2)

Infrastructure & CloudEnterprise Platform

Disaggregated enterprise inference cloud running CPUs, GPUs and RDUs in one pipeline

Contact for pricing

Vector Core Compute (VC2) is an enterprise inference cloud, launched by Vista Equity Partners and Cambium Capital, that splits AI inference across Intel Xeon CPUs, NVIDIA Blackwell GPUs and SambaNova RDUs instead of running everything on GPUs. It targets enterprises and AI platform teams running agentic workloads that need low-latency, US-metro-local inference capacity.

At a Glance

Category
Infrastructure & Cloud
Pricing
Contact for pricing
Target Market
CTOs, CIOs, VPs of Engineering, AI Infrastructure Leaders, Enterprise Architects
Founded
2026

Key Features

  • Fully disaggregated inference
  • Tri-silicon architecture
  • Metro-distributed capacity
  • Agentic-workload orchestration
  • Independently benchmarked speed

Use Cases

  • Serve production agentic AI at scale
  • Low-latency regional inference
  • Diversify away from GPU-only capacity

Ideal For

Best For

  • High-volume agentic AI inference that needs low latency close to end users
  • Enterprises seeking an alternative to GPU-only inference capacity
  • AI platform teams benchmarking cost-per-token against hyperscaler inference

Market Analysis

Enterprise-gradeInfrastructure providerAgentic-first

Pros

  • Independently benchmarked at two to three times faster than a GPU-only stack
  • Metro-distributed footprint reduces inference latency for US enterprises
  • Anchor customer Together AI and Vista's portfolio give it immediate demand

Cons

  • Only the Los Angeles site is live; Chicago, Seattle and Phoenix are still in development
  • No public pricing, published API documentation or self-service onboarding yet
  • Very new company with a short operating track record

Pricing

Enterprise inference capacity

Contact for pricing

  • Disaggregated CPU/GPU/RDU inference
  • Metro-local endpoints
  • Agentic workload orchestration

No public pricing has been published; capacity is sold to enterprise customers directly.

THE DAILY BRIEF

Enterprise AI insights for technology and business leaders, twice weekly.

beri.net

Subscribe at beri.net/subscribe for twice-weekly AI insights delivered to your inbox.

LinkedIn: linkedin.com/in/rberi  |  X: x.com/rajeshberi

© 2026 Rajesh Beri. All rights reserved.

Vector Core Compute (VC2) is an enterprise inference cloud, launched by Vista Equity Partners and Cambium Capital, that splits AI inference across Intel Xeon CPUs, NVIDIA Blackwell GPUs and SambaNova RDUs instead of running everything on GPUs. It targets enterprises and AI platform teams running agentic workloads that need low-latency, US-metro-local inference capacity.

Vector Core Compute, known as VC2, launched on June 3, 2026 as what its backers describe as the world's first commercially available enterprise inference cloud built on a fully disaggregated inference architecture. Instead of running every stage of inference on GPUs, VC2 assigns each stage to the processor best suited to it: Intel Xeon 6 CPUs handle orchestration and tool execution for agentic workloads, NVIDIA Blackwell GPUs handle prefill and prompt caching, and SambaNova SN40 RDUs handle decode. Independent measurement by Artificial Analysis found the architecture to be at least two to three times faster than a GPU-only stack. The company was formed by Vista Equity Partners and Cambium Capital and is backed by a $3.5 billion compute commitment to SambaNova with support from Intel; SambaNova CEO Rodrigo Liang has called VC2 the largest commercial deployment of SambaNova. VC2 went live from a Los Angeles facility, with additional sites in development in Chicago, Seattle and Phoenix and a stated plan to distribute inference endpoints across more than 50 US metropolitan markets rather than concentrating capacity in a handful of remote mega-sites. Together AI, which serves hundreds of trillions of inference tokens a month, is VC2's first commercial customer, and Vista Equity Partners has secured early access for its 90-plus portfolio companies. The company positions itself publicly as 'building North America's agentic-first enterprise cloud.'

At a Glance

Category
Infrastructure & Cloud
Pricing
Contact for pricing
Target Market
CTOs, CIOs, VPs of Engineering, AI Infrastructure Leaders, Enterprise Architects
Founded
2026

Key Features

  • Fully disaggregated inference

    Each inference stage is assigned to a different processor type rather than running the whole pipeline on GPUs.

  • Tri-silicon architecture

    Intel Xeon 6 CPUs for orchestration and tool execution, NVIDIA Blackwell GPUs for prefill and prompt caching, and SambaNova SN40 RDUs for decode.

  • Metro-distributed capacity

    Inference endpoints are being distributed across 50-plus US metropolitan markets instead of concentrated in remote mega-sites, putting compute near the enterprises using it.

  • Agentic-workload orchestration

    CPU-side orchestration and tool execution is designed specifically for multi-step agentic AI rather than single-shot prompts.

  • Independently benchmarked speed

    Artificial Analysis measured the disaggregated architecture at least two to three times faster than a GPU-only stack.

Use Cases

  • Serve production agentic AI at scale

    Run multi-step agent workloads where CPU-based tool execution and orchestration sit alongside GPU prefill and RDU decode in a single pipeline.

  • Low-latency regional inference

    Place inference endpoints in the same metro as the applications and customers consuming them to cut round-trip latency.

  • Diversify away from GPU-only capacity

    Enterprises constrained by GPU supply or GPU-only economics can buy inference capacity built on a mixed CPU/GPU/RDU fleet.

Ideal For

Best For

  • High-volume agentic AI inference that needs low latency close to end users
  • Enterprises seeking an alternative to GPU-only inference capacity
  • AI platform teams benchmarking cost-per-token against hyperscaler inference

Market Analysis

Enterprise-gradeInfrastructure providerAgentic-first

Pros

  • Independently benchmarked at two to three times faster than a GPU-only stack
  • Metro-distributed footprint reduces inference latency for US enterprises
  • Anchor customer Together AI and Vista's portfolio give it immediate demand

Cons

  • Only the Los Angeles site is live; Chicago, Seattle and Phoenix are still in development
  • No public pricing, published API documentation or self-service onboarding yet
  • Very new company with a short operating track record

Pricing

Enterprise inference capacity

Contact for pricing

  • Disaggregated CPU/GPU/RDU inference
  • Metro-local endpoints
  • Agentic workload orchestration

No public pricing has been published; capacity is sold to enterprise customers directly.

Sources

This page was written from 3 sources, 2 on domains other than vectorcorecompute.com.

  1. 1.itbrief.co.ukvista launches vector core compute for ai inference
  2. 2.aithority.comvista equity partners and cambium launch vector core compute
  3. 3.vectorcorecompute.comvectorcorecompute.comvendor
Newsletter

Stay Ahead of the Curve

Weekly enterprise AI insights for technology leaders. No spam, no vendor pitches—unsubscribe anytime.

Subscribe