YOLOv8m on Rockchip RK3588 NPU: 1 row.
Method v1.0 · dataset 2026-09-07 · verified 2026-09-07 · last updated September 2026
yolov8m on Rockchip RK3588 NPU: 1 row (0 measured, 1 derived). Expected: INT8 21.5 fps. Fastest: NVIDIA A100 (server) 546.4 fps (class C), NVIDIA L4 (24GB) 272 fps (class D), Jetson Thor T5000 262 fps (class D).
Change hardware, precision, runtime or add a comparison set on the live engine.
Expected range per precision
Min / median / max of the matching rows, all classes. When a precision has no row, the estimator's benchmark hierarchy (B → C → D) supplies a fallback with its tier chain reported.
| Precision | Metric | N | Min | Median | Max | Best class | Measured |
|---|---|---|---|---|---|---|---|
| INT8 | fps | 1 | 21.5 | 21.5 | 21.5 | D | 0 |
Rows
| Precision | Runtime | Input | Power mode | Software | Metric | Class | Date | Source | Derivation | Dupes |
|---|---|---|---|---|---|---|---|---|---|---|
| INT8 | rknn | 640x640 b1 | default | sdk v1 | 21.5 fps | D | 2026 | EdgeAIStack · edgeaistack:benchmark-corpus · RK3588 Extended Models · verified 2026-09-07 | RK3588 Extended Models | 0 |
Cross-hardware comparison
Best row per hardware for the same model and precision, relative to the fastest.
| Hardware | Best value | Class |
|---|---|---|
| NVIDIA A100 (server) (reference) | 546.4 fps | C |
| NVIDIA L4 (24GB) | 272 fps | D |
| Jetson Thor T5000 | 262 fps | D |
| Jetson AGX Orin 64GB | 248 fps | D |
| NVIDIA A10 (24GB) | 214 fps | D |
| Jetson AGX Orin 32GB | 189 fps | D |
| NVIDIA A2 (16GB) | 143 fps | D |
| Jetson Orin Nano Super | 108 fps | D |
| Hailo-10H | 75.1 fps | D |
| Jetson Orin NX 16GB | 56 fps | D |
| Hailo-8L | 50.8 fps | D |
| Jetson Orin NX 8GB | 50 fps | D |
| Hailo-8 | 45 fps | C |
| Rockchip RK3588 NPU | 21.5 fps | D |
| Jetson Orin Nano 8GB | 13 fps | D |
| Neousys Nuvo-9531 (Intel 12th-14th Gen Core) | 6 fps | D |
Assumptions
- Class C = measured and published by the source named on the row; class D = EdgeAIStack-derived (derivation stated). Rows are inference-only unless scope says end_to_end; publisher measurement conditions (JetPack, batch, warm-up) are in the linked page.
Warnings
- 1 of 1 matching rows are EdgeAIStack-derived (class D); treat them as estimates until a measured row (class C) or your own measurement replaces them.
Method and limitations
Rows are read from the normalised benchmark database (publisher benchmarks and research-derived rows, deduplicated per hardware × model × runtime × precision × input × power mode, best evidence kept). Expected range = min / median / max of the matching rows per precision; when a precision has no row the estimator's benchmark hierarchy (B → C → D) is consulted and its tier chain reported. Comparison = best row per hardware for the same model and precision, relative to the fastest. Nothing is scored; every number carries its class and source line.
Full method, dedupe rule, evidence classes and the confidence rule: Benchmark Explorer methodology.
Measure it yourself.
Nothing measured for your exact config, or want to confirm a class-D row? Run the measurement protocol on your board and submit the result — a human checks it against the protocol before it becomes a class-C row with your attribution.
Links
- Live engine (this exact query): https://edgeaistack.ai/engines/benchmark-explorer/?hw=rk3588_npu&model=yolov8&variant=m&mv=1.0&dv=2026-09-07
- Methodology: /methodology/benchmark-explorer/
- Measurement protocol: /methodology/measurement-protocol/
- Dataset: /datasets/benchmarks-v2.json
- Rockchip RK3588 NPU overview: /benchmarks/rk3588-npu/
Change one input and re-run.
The live engine keeps every source line and gives you a fresh permalink and summary.