The Unit Economics of AI
Ed Huang , Jeremy Murray , John Malcolm , Aaron Bawcom , Matt Burns
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Are skyrocketing GPU costs stalling your AI deployments? Discover how intelligent routing, agentic caching, and model compression can slash inference bills and reduce your compute footprint by 80 percent.
Checking access…
Playback and chapters load privately for Free videos.
Matching moments
More from World Congress 2026 North America
Related videos
Related articles
From learning to earning
Jobs that call for the skills explored in this talk.
26 days ago
Principal Software Engineer, AI Inference Cloud
ARM
Seattle, WA, United States
Expert
$262k
Pytorch
TensorRT
Kubernetes
about 2 months ago
•
Verified
Lead Software Engineer, Model Serving Platform
Sciforium
San Francisco, United States
Expert
$230k–300k
Python
26 days ago
Principal Software Engineer, AI Inference Runtime
ARM
Seattle, WA, United States
Expert
$262k
Compilers
Low Latency
Concurrency
about 2 months ago
•
Verified
Senior AI Serving Engineer, Backend
Sciforium
San Francisco, United States
Expert
$190k–250k
Rust
Python
26 days ago
Principal Software Engineer, AI Compute Infrastructure
ARM
Seattle, WA, United States
Expert
$262k
Linux
Grafana
Pytorch
about 2 months ago
•
Verified
GPU Cluster Engineer, Systems & Platform
Sciforium
San Francisco, United States
Expert
$150k–220k
Kubernetes