Faster, Exact, More General Response-Time Analysis for NVIDIA Holoscan Applications
Philip Schowitz, Shubhaankar Sharma, Siddharth Balodi, Soham Sinha, Bruce Shepherd, Arpan Gujarati
摘要
We present a scalable method to compute exact worst-case end-to-end latency in applications built on the NVIDIA Holoscan SDK, a framework increasingly adopted for soft real-time ML workloads in medical devices, surgical instruments, and robotics. Holoscan applications are structured as directed acyclic graphs of non-preemptible task threads (operators) connected by FIFO queues, where execution depends not only on input availability but also on downstream buffer capacity - an atypical backpressure mechanism not captured by standard dataflow or middleware models. Existing analyses either lack convergence guarantees or rely on restrictive assumptions (e.g., fixed execution times, unit-sized buffers), resulting in overly conservative bounds. We show that Holoscan's scheduling semantics can be faithfully reduced to homogeneous synchronous dataflow graphs (HSDFGs), enabling exact end-to-end latency analysis. Building on this insight, we introduce a dynamic algorithm that computes tight upper bounds on response time across infinite input streams under variable task runtimes and arbitrary buffer sizes. We prove its correctness and convergence, and demonstrate that it outperforms HSDFG model checking with UPPAAL in runtime while avoiding the pessimism of prior Holoscan-specific analyses. Experiments on real Holoscan applications from NVIDIA HoloHub and large synthetic graphs confirm its scalability and precision.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
相关 Paper
- Response-Time Analysis of a Soft Real-time NVIDIA Holoscan ApplicationPhilip Schowitz, Soham Sinha, Arpan GujaratiRTSS 2024 · 被引用 2 次
- Response Time Analysis and Priority Assignment of Processing Chains on ROS2 ExecutorsYue Tang, Zhiwei Feng, Nan Guan, Xu Jiang 等RTSS 2020 · 被引用 81 次
- Response time analysis for dynamic priority scheduling in ROS2Abdullah Al Arafat, Sudharsan Vaidhun, Kurt M. Wilson, Jinghao Sun 等DAC 2022 · 被引用 38 次
- Real-Time Scheduling and Analysis of Processing Chains on Multi-threaded Executor in ROS 2Xu Jiang, Dong Ji, Nan Guan, Ruoxiang Li 等RTSS 2022 · 被引用 39 次
- LaLaRAND: Flexible Layer-by-Layer CPU/GPU Scheduling for Real-Time DNN TasksWoosung Kang, Kilho Lee, Jinkyu Lee, Insik Shin 等RTSS 2021 · 被引用 68 次
