export PYTORCH_MPS_HIGH_WATERMARK_RATIO=0.0
Total VRAM 24576 MB, total RAM 24576 MB
pytorch version: 2.8.0
Mac Version (15, 6, 1)
Set vram state to: SHARED
Device: mps
Using sub quadratic optimization for attention, if you have memory or speed issues try using: --use-split-cross-attention
Python version: 3.12.11 (main, Aug 18 2025, 19:02:39) [Clang 20.1.4 ]
ComfyUI version: 0.3.54
ComfyUI frontend version: 1.25.11
[Prompt Server] web root: /Users/zmo/Workspaces/wan-video-maker/.venv/lib/python3.12/site-packages/comfyui_frontend_package/static
Import times for custom nodes:
0.0 seconds: /Users/zmo/Workspaces/wan-video-maker/ComfyUI/custom_nodes/websocket_image_save.py
Context impl SQLiteImpl.
Will assume non-transactional DDL.
No target revision found.
Starting server
To see the GUI go to: http://127.0.0.1:8188
got prompt
Using split attention in VAE
Using split attention in VAE
VAE load device: mps, offload device: cpu, dtype: torch.bfloat16
Using scaled fp8: fp8 matrix mult: False, scale input: False
Requested to load WanTEModel
loaded completely 9.5367431640625e+25 6419.477203369141 True
CLIP/text encoder model load device: cpu, offload device: cpu, current: cpu, dtype: torch.float16
model weight dtype torch.float16, manual cast: None
model_type FLOW
Requested to load WAN22
0 models unloaded.
loaded completely 9.5367431640625e+25 9536.402709960938 True
100%|█████████████████████████████████████████████████████████████████████████████████████████████████████████████| 20/20 [46:35<00:00, 139.79s/it]
Requested to load WanVAE
loaded completely 9.5367431640625e+25 1344.0869674682617 True