Bringing Distributed Compute & AI to the Edge w/ David Aronchick (CEO of Expanso)
In this episode of the Joe Reis Show, I sit down with David Aronchick, co-founder of Expanso and one of the early pioneers behind Kubernetes, Kubeflow, and the CNCF. We dive deep into the evolving data landscape and discuss why the architectural pendulum is swinging back from pure cloud centralization toward distributed edge compute. David breaks down the real operational nightmares of edge data collection, why treating your bronze tier as a "toxic waste dump" breaks downstream pipelines, and how applying schema and context upstream makes your data truly AI-ready. We also get into the critical need for data provenance, data bills of materials, and what agents actually need to communicate reliably. Website: https://expanso.io Timestamps 00:00 - Catching up & the reality of European deal cycles 01:35 - David Aronchick’s background: Kubernetes, GKE, Kubeflow, and startups 02:27 - Startups vs. Big Tech: Moving fast vs. enterprise impact 04:09 - What Expanso does: Bringing the Kubernetes playbook to distributed data 06:19 - Core challenges of edge data pipelines & deployment 10:06 - Reliability and classic distributed computing problems 12:52 - Winning Edge AI Startup of the Year & production ML use cases 15:23 - Navigating startup strategy: Customer signals vs. founder vision 19:08 - The architectural pendulum: Moving from central cloud back to the edge 22:33 - Local compute isolation, Raspberry Pis, and running agents 25:21 - Tiered model architectures and making data AI-ready 31:45 - Data modeling at the edge: Shifting left and enforcing upstream schemas 37:56 - Data provenance, chain of custody, and security for AI agents 42:26 - DuckDB, semantic contracts, and reproducible data architectures 48:39 - What’s next for AI and augmenting human work