Research 1 min read

Four GPUs, one week per run

The prototype works. Now the slowest part of our research is waiting: about a week per training run on four GPUs.

Tender Touch team
Portugal
Four graphics cards with blue status lights in an open rack in our office.

Our current setup uses four GPUs. A big training cycle takes about a week. That was enough to prove the pipeline works, but it is too slow for efficient research.

Every change to the data, the training strategy or the evaluation can cost days. Most of that time is spent waiting, not thinking.

What more compute would change

With 8 or more GPUs we could compare checkpoints side by side, repeat promising runs to separate real gains from noise, and spend more iterations on how the model sounds in Portuguese.

We do not expect linear scaling. We expect fewer days lost between decisions.

This is why we are applying to European compute and R&D programmes (see the roadmap). If you run an HPC centre or a speech research group, we would like to talk.

Working on something similar?
Tell us what you are building.
Talk to us

More from the journal

See all