FR
← Blog
AI and trust

AI-generated voiceover in corporate video: where it passes, where it breaks

September 2026 · AI and trust · Québec

An AI-generated voiceover works in a corporate video when it describes a process to someone who needs the information, not someone who needs convincing. It breaks the moment it speaks for the company in front of a buyer, and the moment the script contains technical Quebec French or local names. The line has nothing to do with how good the voice sounds. It depends on what the voice is supposed to prove. A lockout training module can be narrated by a synthetic voice and nobody loses anything. A capabilities film sent to a prime contractor cannot, because it asks the buyer to believe someone.

Narrating a process or speaking for the company

Ask one question before choosing the voice: does the viewer need to trust whoever is speaking? If the answer is no, the voice is a vehicle. If the answer is yes, the voice is part of the proof.

On the vehicle side sit internal training clips, safety instructions for new hires, the start-up procedure for a packaging line, the video that shows a distributor how to install a replacement part. Nobody wonders who is talking. They want to know which bolt to tighten first. A clean synthetic voice, reading a script the floor supervisor has signed off, does that job. It can also be redone without calling anyone in when the procedure changes, which matters in a plant where jobs shift every year.

On the proof side sit the home page video, the presentation sent after a first call, the message from the president, the video that introduces the company at a trade show. Here the buyer is not consuming information. They are evaluating a supplier. They are looking for signs that the company has competent people who stand behind what they say. A voice nobody owns sends the opposite signal: the company could not find anyone to speak for it.

Picture a conveyor manufacturer in Montérégie answering a qualification request from a food processor. Its presentation video is smooth, well cut, narrated by a flawless neutral voice. The buyer toured three plants last month and spoke with three production managers. The video tells him nothing about the people he would have on the phone the day a line goes down. That was exactly what he was looking for.

This is the argument behind the studio's "Certified Human Content" line: AI generates content, it does not generate trust. In a proof video, the replacement for a generated voice is rarely a better voiceover. It is the person who does the work, filmed at their station, or the leader on camera. The Signal Program is built on the second case, with half a day of shooting every 3 months for the leader. The overall logic is on the video and content production page.

The question is not whether the voice sounds real. It is whether anyone has to believe it.

Where the synthetic voice breaks: technical French and local names

Even as a vehicle, a generated voice has predictable blind spots in a Quebec SME. They are not flaws of tone. They are reading errors, and a technical buyer hears them fast.

Shop floor vocabulary

The French spoken in a plant is not textbook French. People say "gabarit", "bavure", "tolérance", plus plenty of English words everyone pronounces the Quebec way. Every trade has its acronyms. A synthetic voice reads what it sees. It may pronounce an acronym as a word, or a floor term with an accent that belongs to no shop in the province. For a welder in Saint-Georges or a buyer in Bécancour, that detail says one thing: nobody here knows the trade.

Numbers and units

A spec sheet mixes inches and millimetres, fractions and decimals, model numbers and years. "3/8 po", "1,5 mm", "série 400", "ISO 9001:2015". Every line is a chance to misread. A misread tolerance is not a style issue. It is false information in a video that carries your name.

Proper names

Quebec place names are a simple test: Saint-Hyacinthe, Chicoutimi, Rivière-du-Loup, Sainte-Marie, Lac-Mégantic. Add the names of your products, your customers and your own company. A voice that mangles the name of the town where the plant stands gets no second chance with someone who lives there.

What to demand if you use one anyway

If the supplier cannot name the person who listened to the final version, you have your answer.

What the platforms and copyright say

The public rules mostly target deception about a real person or a real event, not the use of a tool. Three points are worth knowing before you publish.

YouTube. Its help page on disclosing generated content requires a label when content makes a real person appear to say or do something they did not, alters footage of a real event or place, or generates a realistic scene that did not happen. The same page exempts production assistance, such as using generative tools to write or improve a script. On our reading, an anonymous synthetic narrator describing a procedure fits none of the three cases. A cloned voice of your president reading words he never said, posted on the company channel, fits the first one. The same page does exempt creators cloning their own voice for voiceovers or dubs: that exemption covers the person who publishes, not a leader whose voice the company clones. YouTube adds that creators who consistently choose not to disclose may face a manually applied label, removal of content or suspension from the YouTube Partner Program.

Facebook and Instagram. Meta's misinformation policy requires people to disclose, using its dedicated tool, any organic post with photorealistic video or realistic-sounding audio that was digitally created or altered. It provides for penalties when they do not. A convincing synthetic voiceover posted on the company page fits that description. On the advertising side, the disclosure rule we reviewed covers ads about social issues, elections or politics. It says nothing about ordinary commercial ads, and we read nothing more into it.

Copyright. The federal government's consultation paper on copyright in the age of generative AI notes that case law attributes authorship to a natural person who exercises skill and judgment. It considers that criterion far less likely to be met for a work produced from short instructions. It also calls the question of who authors a generated work an open one. This is a consultation paper, not a law. It is still enough to put one question to your supplier: in the narration you are delivering, what exactly does the company own?

One last point is a matter of judgment rather than a rule written for AI. The Competition Bureau notes that courts look at the general impression a representation conveys, not only its literal meaning. In our view, a synthetic voice introduced as "Marc, our floor supervisor" conveys a false general impression, even if every sentence in the script is accurate.

When this does not apply

If the video never leaves the plant, the trust argument loses almost all its weight. A training clip for your own staff, revised several times a year, is a sound use of a generated voice. Going back to a recording studio for every revision would be hard to justify. The one requirement left is technical sign-off, because a safety instruction read wrong is a risk, not a detail.

The same goes for a video with no narration. Plenty of industrial videos work with the real sound of the machine and captions, no voice at all. In that case the question never comes up.

Finally, if the video only shows a machine to a distributor who already knows it, the voice matters little. The picture does the work, and where the video will play decides its form. The page on corporate video length takes that question destination by destination, and the one on virtual factory tour videos covers what a proof video has to carry.

Sources

If you are weighing a generated voice against a real one for your next video, book a 15-minute call to see whether the studio can be useful to you.

Book my call