Skip to content

ハードウェア & GPU 加速 ​

OneASR は多様な計算ハードウェアに深く最適化されており、マルチ GPU サーバーから薄型ノート PC まで優れた計算効率を発揮します。


1. NVIDIA CUDA & TensorRT ​

高同時実行・高スループットが求められるサーバー本番環境には、NVIDIA GPU の使用を推奨します:

yaml
ASR-Providers:
  faster-whisper:
    enable: true
    load:
      device: "cuda"
      compute_type: "float16" # VRAM が限られている場合は int8_float16 に設定可能

VRAM 要件と同時実行数の目安 ​

エンジン / モデル仕様最小 VRAM (int8)推奨 VRAM (fp16)同時実行処理能力
Faster-Whisper Tiny/Base1.0 GB2.0 GB20+ 並列
Faster-Whisper Small/Medium2.5 GB5.0 GB8 - 12 並列
Faster-Whisper Large-v34.5 GB8.0 GB4 - 6 並列
X-ASR (Zipformer ONNX)< 500 MB< 1.0 GB50+ 並列
Qwen3-ASR (1.7B)3.5 GB6.0 GB4 - 8 並列

2. Apple Silicon (M1 / M2 / M3 / M4) ​

macOS デバイス上では、OneASR は統合メモリと Apple Silicon の高性能 CPU コアを活用します:

yaml
ASR-Providers:
  faster-whisper:
    enable: true
    load:
      device: "cpu"
      compute_type: "int8"

TIP

Apple Silicon チップを搭載した Mac では、int8 量子化と 4 つの CPU スレッドを使用することで、リアルタイムの 4 倍以上の高速文字起こし速度(RTF < 0.25) を達成でき、消費電力と発熱も極めて低く抑えられます。