I have several tapes (yes actual cassette tapes) of my grandfather reading a novel.

Unfortunately a few of the tapes have degraded to the point that I cannot play them back.

I would love to recreate his voice, to “rerecord” the missing bits.

The recordings are in Danish.

Is this possible?

If it is, how can I go about it?

  • boojumliussnark@lemmy.worldOP
    link
    fedilink
    arrow-up
    0
    ·
    3 months ago

    Thank you for the tips. As I see it currently, I expect the language to be the biggest hurdle. It doesn’t appear like something I can add myself, even if I had the data for a model. So as far as I can tell it involves two currently more or less impossible steps: Get model data and teach language to model.

    • Grimy@lemmy.world
      link
      fedilink
      arrow-up
      0
      ·
      edit-2
      3 months ago

      If you have material with him speaking in English, you might be able to train an xtts model on it and then use that to bypass the elvenlabs captcha but I’m not sure if they give enough time. Although GPU rental is cheap these days, so captcha time is less of a factor.

      If anything, the tech is moving quite fast, it will definitely be easier in a few years, maybe even months.