
maddy arvapally
Founder, ZeroGPU
advisors

Jarkko Rajamaki
Founding team at Rovio / Angry Birds. Advises on scale, supply, and unit economics.

Dan Goikhman
Repeat founder, multiple $50M+ exits. Advises on GTM, sales, media, and enterprise.

Mark Levin
CEO, Consumable. Advises on supply, distribution, and enterprise sales.

Krish Arvapally
Repeat founder, CTO with multiple exits. Advises on engineering, architecture, and models.
about
Hi, I'm Maddy, founder of ZeroGPU.
I'm a hardcore engineer and systems architect by training. Over the last decade, I've led and built large-scale production systems across streaming, adtech, distributed infrastructure, robotics, and applied research.
Most recently, I led engineering at a distributed infra company called Replay. I've also consulted on adtech engineering and video streaming products for B20 Labs. Before that, I led software development at GoPro and was part of the core engineering team that launched GoPro's streaming service from the ground up, scaling it from $0 to over 5 million subscribers. Earlier in my career, I worked at a robotics startup and conducted research in brain modeling at the Mind Research Network.
why zerogpu
The world needs more compute, and it needs it now
AI is the next generational platform shift and it's hungry for compute. Every new model, agent, and app adds inference demand that is exploding across power, electricity, and data-center capacity faster than the world can build it.
Inference demand is exploding
AI inference grows from $106B to $255B by 2030, a 19% CAGR.
Capacity can't keep up
Power, electricity, and data-center buildout lag demand.
Most AI workloads don't need a frontier model
A large share of workloads never needed frontier GPUs.
The ZeroGPU bet
They're building data centers in space and under the ocean. We tap the compute already in your hand.
Small and nano language models can run on lighter compute: the unused capacity already in phones, laptops, and edge devices worldwide. We unlock the supply the world can't build from scratch.
One programmable layer over the world's idle compute
A lightweight ZeroGPU SDK and orchestration layer turns any app's install base into a private inference cloud.
Edge devices
Phones, gaming PCs, and laptops running nano and small models.
Optimized edge servers
Mid-sized models and higher load.
Cloud fallback
Consistent performance and burst capacity.
Workloads include classification, summarization, embeddings, moderation, extraction, and translation.