This week, a Reddit help post drew our attention to an awkward fact: every mainstream AI voice-cloning tool, from ElevenLabs to OpenAI Voice Engine, is fundamentally designed only for human voices. A user wanted to dub their own hen; every tool rejected the request.
What this is
The user posted in LocalLLaMA, a technical community for discussing locally deployed large language models, saying that they keep several hens. They wanted to record their daily calls, train a “chicken voice” model capable of speaking any text, and pair it with AI-generated video—in effect, making their hens “talk.”
They had recordings of their hens making sounds other than their usual clucks that could be looped and reused, but found that every voice-cloning tool required clear human speech by default. Even AI search could not find a workable solution.
The technical challenges come down to two points: animals produce sound through mechanisms different from those of humans—birds use a syrinx, while chickens rely on airflow and air sacs, whereas human voices come from vibrating vocal cords. The other problem is the near-total lack of training data: there is no labeled dataset of a chicken speaking English.
Industry view
The case is small, but it reveals a significant problem.
Supporters see it as an early sign that speech AI is moving into long-tail use cases. Voice cloning is expanding from serving whoever pays—customer service, podcasts, and accessibility tools for visually impaired users—to serving anyone with a creative idea. Independent creators, pet bloggers, and educators all have strong demand for personalized voiceovers.
But the objections are equally clear: who owns the rights to animal sounds? If a user trains a model using recordings of their own chicken, does the resulting model belong to the user or the platform? Take the idea further: if someone clones the calls of endangered animals for commercial use, almost no bioacoustic intellectual-property framework exists to protect species’ audio data rights. Legal preparedness is clearly lagging behind the technology’s adoption.
It is worth noting that this is not an isolated case. Pet vocal mimicry, audiobook dialect cloning, and recreating the voices of deceased people are all emerging rapidly, while tools to support them lag behind noticeably.
Impact on regular people
For individual careers: these tools will not threaten office workers’ jobs in the short term. But if you are a content creator, pet blogger, or educator, you are likely to see a wave of new tool opportunities over the next 12–24 months, giving early movers a chance to benefit.
For enterprise IT: this is not yet a project-initiation priority. It is part of the long tail of consumer creative tools, so managers in traditional industries can ignore it.
For the consumer market: in 2025–2026, we will probably see a wave of tools for personalized pet and human voiceovers. You do not need technical expertise, but you should decide how much access to your own or your family’s voiceprints you are willing to license to third parties.