AI Agents for LLM Inference Runtimes on Edge Hardware

PyTorch
761 views September 11, 2026

At PyTorch Conference North America, Thomas Cottenier from Arm will present how AI agents can synthesize customized, target-specific runtimes using PyTorch components like torch.export, ExecuTorch, and torchao. This approach removes redundant prefill computation and creates streamlined inference pipelines that deliver high performance directly on edge hardware. Join us San Jose this October 20-21 to learn more: https://hubs.la/Q04v4SL60

Keyboard shortcuts

On. Switch them off if they collide with your assistive tools; ? still opens this sheet.

Go to

Press g then the letter.

  • gh Latest
  • gs Sources
  • gm Media
  • gv Videos
  • gp Podcasts
  • gc Calendar
  • gd Decoder
  • gz Dataviz
  • ga Datasets
  • gb Blog
  • gk Markets
  • gj Careers
  • gn Prompt Notebook

On this page

  • / Focus search, where there is one
  • t Back to top
  • ? This list
  • Esc Close