LLM Inference On-Device in React Native: The Practical Aspects by Artur Morys-Magiera @CallstackEngineers
LLM Inference On-Device in React Native: The Practical Aspects by Artur Morys-Magiera  @CallstackEngineers
Uploaded February 2026 | Updated September 2026, 2 days ago
In this session, Artur Morys-Magiera explores what it actually takes to run LLM inference directly on mobile devices in React Native applications.

The talk covers:

Why teams move inference on-device: reliability, privacy, and latency
Model size constraints and OS-provisioned alternatives
GPU, NPU, and CPU acceleration tradeoffs across iOS and Android
Debugging performance issues in abstraction layers like OpenCL
Comparing TF Lite, ONNX, ExecuTorch, MLC, llama.cpp, and Apple-based solutions
Practical optimizations including quantization and compilation-time improvements

If you’re evaluating on-device AI in a React Native app, this is a grounded overview of what works today, and what still requires careful tradeoffs.

Follow Callstack on X 🐦 https://x.com/callstackio
LLM Inference On-Device in React Native: The Practical Aspects by Artur Morys-MagieraAI Assisted Migrations to React Native:  From Months to DaysThe Generations of React | Anisha Malde at React Universe Conf 2025Structure Monorepo for Speed, Scale & Sanity With Nx | React Universe On AirWhat if React Native migration didn’t take weeks?ASC CLI + Agents for App Store Localization | Rudrank RiyamExploring Skills in ClaudeWhat Is Observability and Why Should You Care?React Native Evals: Making AI Code Quality MeasurableHow ASC CLI lets agents automate App Store work? Listen to React Universe On Air PodcastCan Eve Make AI Agents Easier to Build and Run?React Native Ease: Faster Native Animations with Platform Primitives  | React Universe On Air
Callstack |

LLM Inference On-Device in React Native: The Practical Aspects by Artur Morys-Magiera

SHARE TO X SHARE TO REDDIT SHARE TO FACEBOOK WALLPAPER