Search videos
View all results →
← EuroRust 2025
Building a Chatbot Service with Rust, WGPU, and Tokio
EuroRust
This talk is an exploration of what’s currently possible in the Rust ecosystem for running machine learning inference locally – without relying on CUDA or PyTorch. We’ll walk through building a chatbot service powered by Rust, using WGPU for GPU-accelerated compute via WGSL shaders, and Tokio for serving responses asynchronously over an API. A key focus will be on bridging the asynchronous world of WGPU’s GPU command submission with Rust’s async ecosystem, especially Tokio. We’ll examine how to
Player unavailable or embedding blocked? Use the YouTube link above.