4 ms·
Actually I was just checking, and Bark isn't that close to maxing out GPU utilization. Running two instances on a 3090 seems like a throughput increase and the
by JonathanFly 3y ago
Actually I was just checking, and Bark isn't that close to maxing out GPU utilization. Running two instances on a 3090 seems like a throughput increase and the models fit. Update: And getting weird CUDA issues. Hmn...
- woodson 3y agoJust to add a datapoint: the main audioLM based models (not the BERT embedding part) fully utilize an RTX 2080 Ti.
- JonathanFly 3y agoMust be some low hanging fruit to optimize in Bark. It would be somewhat close to realtime if it was close to 100% and scaled linearly.