vLLM GPU throughput
Fun-ASR-Nano-2512
- Runtime
- vLLM batch
- Hardware
- GPU model not recorded in the cited public table
- Workload
- Offline batch
- Audio
- 184 long-form files; 11,541 seconds
- Settings
- Dynamic VAD; the public table does not record batch size or the full software and hardware stack
- Timing scope
- Offline throughput; the cited table does not state whether warmup, file I/O, and decoding are excluded
- Verified
- 2026-07-26
Qualification: Incomplete hardware and timing record. Use only as public reference evidence, not a capacity promise.
Open primary source