How I Made CosyVoice3 2.2× Faster on Apple Silicon
How I bounded PyTorch MPS memory growth, cut CosyVoice3 latency below real time, and moved only the dispatch-bound LLM stage to MLX.

Tag archive
Notes connected by the Pytorch tag.
How I bounded PyTorch MPS memory growth, cut CosyVoice3 latency below real time, and moved only the dispatch-bound LLM stage to MLX.
No entries are filed here yet.
Browse all notes →