Economics of the Rafay Token Factory
Measured production data showing how GPU operators can boost annual revenue per GPU by 2x–4.8x by shifting from hourly rental to token-metered AI services.
READ PRESS RELEASE
HOTEL WIFI ACCESS: Rafay_Summit (network name), password: Rafay2026
Main Point of Contact: Maria Gallegos, Rafay Systems, email: maria@rafay.co | phone: +1 650-773-7387

Rooftop Terrace Welcome Reception
Kick off the Summit overlooking the Mediterranean — connect with fellow AI infrastructure leaders, NVIDIA, and ecosystem partners.
Coffee, Registration & Networking
Applying Portfolio Theory to AI Factories
Haseeb Budhani — Co-Founder & CEO, Rafay
Capital is flooding into AI infrastructure, but raw capacity is not a business. Haseeb sets the core idea of the summit and folds in the four directions Rafay is building toward and why. The session frames where margin lives and how operators move from selling GPU time to selling outcomes.
The NVIDIA AI Factory Playbook: Where NVIDIA Invests, Where Partners Build, and Why Customers Win
Warren Barkley — VP Product, DSX, NVIDIA
Warren Barkley opens Rafay's customer event with NVIDIA's strategy for sovereign and neocloud AI infrastructure: the thinking behind what NVIDIA expects from the ecosystem, and what that means for operators and enterprises buying into this stack today.
What It Actually Takes to Go From Silicon to a Service Customers Will Pay For
Ashin Uday — Product Lead, Yotta
Ashin Uday joins Haseeb Budhani and Warren Barkley for a conversation on what it actually takes to go from silicon to a service customers will pay for.
AI Neocloud Economics: Winning on Margin and Enterprise Readiness
David Wood — Senior Managing Director, Global Lead, Sovereign AI, Accenture
AI demand is growing rapidly, and building a sustainable neocloud business depends on both GPU access and the economics of operating them. This session connects the economics of AI infrastructure with the capabilities enterprises expect from their providers, examining the margin math behind the neocloud model — GPU utilization, pricing, and the balance between training and inference workloads — alongside what enterprises are actually asking for. Drawing on advisory work across sovereign and neocloud programs worldwide, David Wood identifies where operators most often fall short and what it takes to address those shortcomings.
Executive Networking
Lunch
Lessons Learned From Deploying 35+ Multi-Tenant GPU Clouds and AI Factories
Alex Saroyan — Co-Founder & CEO, Netris
Drawing on lessons from 35+ production GPU cloud and AI factory deployments, Netris CEO Alex Saroyan explores what it takes to operate complex AI networks at scale. He'll cover the challenges of managing multiple networking layers, enabling hardware-enforced multi-tenancy, and moving beyond DIY automation to build more scalable, secure, and repeatable AI infrastructure.
The Business of Neocloud Economics
Karin Roeschlein — CFO, Rafay (Moderator)
Nick Jacobs, Investor, AI Mills · Elvir Stupar, Cisco Global Investment Fund · Chirag Bhagat, Parinita Inc · Chester Reid, CFO, Era4
A financial panel on the economics of the AI business: GPU-as-a-service margins versus token-based AI services, and how utilization, pricing, and mix drive ROI on a major infrastructure bet. Finance and operations leaders from leading neoclouds compare where the returns are coming from today and how they plan to grow margins from here.
Executive Networking
From Token Maxing to Outcome Maxing: Building Hybrid AI Infrastructure for Enterprise Value
Noam Rosen — Director, Enterprise AI Europe & META, Lenovo
AI leaders are no longer asking how many tokens they can process, but what outcomes they can create with the right balance of cost, control, and trust. As agentic AI drives dramatically larger workloads, enterprises will need hybrid infrastructure spanning frontier cloud, private AI, hosted environments, and the edge. This session explores why hybrid AI is becoming the default enterprise model, and why intelligent control planes that route workloads across environments may become the winning layer.
The Next AI Factory: Where Scale, Sovereignty and Economics Converge
Rod Evans — EMEA VP, Supercomputing & AI Cloud Infrastructure, NVIDIA
The next wave of AI will be about turning capacity into AI factories that can serve enterprises, governments, and entire regions at scale. Drawing on NVIDIA's work across supercomputing, cloud partners, and the broader AI ecosystem, Rod Evans explores how the AI infrastructure market is evolving from centralized GPU capacity toward distributed, sovereign, and service-driven AI factories — and what separates infrastructure that simply delivers compute from platforms capable of attracting workloads, improving utilization, and creating durable economic value.
Full Fleet, Thin Margin: The Utilization Number Nobody Owns
Izhar Sharon — SVP AI Customer Advocacy, DDN
A fleet can be sold out, fully allocated, and still miss its margin. "Utilization" is three numbers, not one: what you sold, what got assigned to a job, and what the GPUs spent computing rather than waiting for data. Sales owns the first. Orchestration owns the second. The third has no owner and is capped at design time, months before the first customer. Izhar Sharon separates the three, names the workloads where a fast data tier adds margin and the ones where it takes margin away, and closes with five questions worth asking your own team before the next site is committed.
Executive Networking
Bringing It Home
Haseeb Budhani — Co-Founder & CEO, Rafay
Haseeb closes on the shared playbook behind building AI as a business and, for many, an act of sovereignty.
Executive Dinner
Coffee & Networking
Meet Team Rafay
The New Telco Opportunity: From Connectivity Provider to AI Infrastructure Provider
Franck Jonas — Head of Telco, France, NVIDIA
Telecom operators already own many of the ingredients needed for the AI era: facilities, power, networks, distributed infrastructure, enterprise relationships, and increasingly GPU capacity. Franck Jonas explores how those assets can evolve into a distributed AI infrastructure platform, and what telcos need to add to turn infrastructure advantage into AI services and revenue.
From Design to First GPU-Hour Revenue, Faster
Simon Dumbleton — CTO Europe, World Wide Technology
The gap between an AI factory design and a running, revenue-generating system is measured in months, and every one of those months is GPU-hour revenue lost on infrastructure already paid for. This session covers how the right systems integrator takes an operator from design to deployed across the full stack, getting them to first GPU-hour revenue sooner. It pinpoints where deployment stalls, and how the best operators avoid it.
Executive Networking
The Secure Foundation for Sovereign AI
Jon Evans — EMEA AI Specialist, Neoclouds & Enterprise GTM, Cisco
Secure and sovereign AI starts with a validated foundation. Cisco on the compute, fabric, and validated architectures that let operators serve regulated and sovereign demand.
The Margins That Turned Doubt Into Scale
Karan Kirpalani — CPO, Neysa
Neysa's Karan Kirpalani on building a business that beat the market on margins when few believed it could, and the discipline that turned into real scale. A candid read on what superior unit economics look like and how they compound.
Executive Networking
Lunch
Optimizing Inference at Scale
Laura Morselli — NVIDIA
As operators scale from a single model to many, inference optimization becomes the key to margins. NVIDIA shares its perspective on the future of the token factory, exploring how to accelerate LLMs while driving down costs.
Beyond the Build: What It Takes to Operate a Profitable NeoCloud
Matt McGuiness — Rafay
The distance between hardware and a production AI business is where operators lose time and money.
Distributed Compute: Where AI Infrastructure Meets the Built Environment
Xeal
AI's growth is widely seen as capped by power. The real constraint is where that power sits. Data center demand is set to outpace new grid capacity through 2030, and inference needs to run close to users to perform. The solution to both already exists in the ground: distributed, permitted power behind EV charging infrastructure in dense urban locations, waiting to be utilized.
Building the Demand Side of Sovereign AI in Malaysia
Farul Mohd Ghazali — CTO, Aras Integrasi Sdn Bhd
Most sovereign AI conversations focus on supply — GPUs, datacentres, national clouds. This session flips the lens to demand: what does a government actually need before it can use the capacity being built for it?
Dispatches From the Field: What Customers & Partners Are Telling Us
Mohan & Hemanth, Rafay
Rafay works across neocloud operators, enterprise buyers, and the partner ecosystem that connects them. Mohan and Hemanth share what that work is surfacing: which neocloud, telco, and enterprise requirements come up in nearly every deal, where operators consistently underestimate the effort, how monetization models are shifting from hourly GPU rental toward consumption-based pricing, and what partners are asking Rafay to own versus build with them.
Bringing It Home
Haseeb Budhani — Co-Founder & CEO, Rafay
Haseeb closes on the shared playbook behind building AI as a business and, for many, an act of sovereignty.
Executive Networking
Agenda subject to change. Times shown are local to Barcelona (CEST).
Executives from Rafay, NVIDIA, and the AI infrastructure ecosystem share how operators turn GPU capacity into profitable, sovereign, and scalable AI businesses.
Session
Hosted at the iconic Hotel Arts Barcelona, the summit combines strategic discussions with exceptional hospitality on Barcelona's Mediterranean waterfront.
Continue conversations over authentic Catalan cuisine, evening receptions overlooking the Mediterranean Sea, and relaxed networking in one of Europe's most dynamic technology and innovation hubs.
HOTEL WIFI ACCESS: coming soon
Main Point of Contact: Maria Gallegos, Rafay Systems
maria@rafay.co | +1 650-773-7387








