Skip to content
Guide

Commercial Licenses for Local TTS Models

Check code, weights, voices, reference audio, and product terms before using a local TTS model in monetized videos, client work, courses, or apps.

Murmur5 min read

Direct answer: a local TTS model is not commercially usable merely because its files are downloadable or its repository is public. Check at least five layers: application terms, runtime or code license, exact model-weight license, voice or dataset terms, and your rights to the reference recording and speaker identity. Some popular model families publish permissive Apache 2.0 or MIT terms. Others restrict model weights to non-commercial research or require a separate commercial agreement. Verify the current primary source for the exact checkpoint before paid use.

Commercial use is broader than selling audio

A monetized YouTube video, client voiceover, paid course, advertisement, game, subscription product, internal business process, and hosted API can all be commercial uses. The exact license definition controls, not a casual label on a model list. Internal work is not automatically non-commercial. A model may also allow generated output while restricting redistribution of weights or offering the model as a service. Read definitions, grants, restrictions, notices, and termination clauses together.

This guide is an operational checklist, not legal advice. Licenses can change, model cards can conflict with repository files, and community conversions can add another layer. Save the license text and access date used for a project. For identity and disclosure questions beyond licensing, read the AI voice commercial-rights guide.

The five-layer check

LayerQuestionEvidence to save
ApplicationDo product terms permit the project?Dated terms and purchase tier
Runtime and codeCan the software be used, modified, or distributed?Repository license and dependency notices
Model weightsDo the exact checkpoint terms allow commercial inference?Model card and license file
Voice assetsCan this preset or community voice be used in paid output?Asset provenance and attribution
Reference and identityDid the speaker authorize cloning and the intended use?Consent record and scope
DistributionAre notices, watermarking, or service restrictions required?Release checklist and shipped notices

Do not collapse code and weights

A repository can publish code under one license and model weights under another. SparkTTS is a useful warning: the code and the official model-weight terms must be checked separately, and the published weight license includes non-commercial restrictions. A community MLX conversion may identify the upstream license, but the converter cannot grant rights the upstream owner did not provide. Trace the checkpoint back to its official model card and license file.

Fish Audio S2 Pro uses a research license that requires a separate written agreement for commercial use. That can still be a legitimate choice for evaluation or for a business that obtains the required license. It should not be placed in the same commercial bucket as a permissively licensed checkpoint. Read the current Fish license directly because its definitions and output conditions matter. The open-source model license checklist covers redistribution and dependency review in more depth.

Examples are routing aids, not permanent clearance

Model familyPublished license signalCommercial decision
KokoroOfficial project states Apache 2.0Usually a permissive starting point, verify exact weights and voices
Qwen3-TTSOfficial repository states Apache 2.0Verify checkpoint, conversion, voice, and consent
ChatterboxOfficial repository states MITCheck model card, watermark behavior, and reference rights
SparkTTSOfficial weights list CC BY-NC-SA 4.0Do not use restricted weights commercially without permission
Fish Audio S2 ProFish Audio Research LicenseSeparate written commercial agreement required
Community voice assetVaries by uploader and sourceNo blanket clearance from the surrounding app

Treat this table as a map to the primary source, not a legal conclusion. Verify on the publication or delivery date. A model owner can release several checkpoints under different conditions. A hosted API can provide commercial terms that differ from downloadable weights. A preset voice may come from a licensed actor, a public dataset, or an unclear upload. The most permissive model license does not repair missing voice rights.

Voice cloning needs affirmative permission

A software license governs software or weights, not a person's identity. Use a voice you own or have explicit permission to clone for the stated project, territory, duration, distribution channels, and ability to create derivatives. Do not assume employment, a podcast appearance, or possession of an audio file grants cloning rights. Keep the reference recording, transcript, consent record, and generated-voice identifier together. The private voice-cloning guide provides a handling checklist.

Voice design can reduce identity risk when the goal is a fictional narrator rather than a copy of a real person, but it is not an automatic safe harbor. Avoid prompts that clearly target a living performer or public figure without authorization. Review the result for accidental similarity. Document the prompt, model, seed where available, and approval. Disclosure and platform rules can apply even when the underlying model license is permissive.

How Murmur buyers should evaluate commercial work

Murmur costs $49 one-time, has no free trial, and includes a 7-day refund policy. Its product terms permit generated speech in commercial projects, but that product permission does not override an upstream model or voice restriction. Murmur includes several model families with different terms. A paid job should begin by selecting a checkpoint whose license fits, then validating the chosen preset, community voice, designed voice, or consented clone.

Murmur's value is the local production layer: model management, voices, scripts, projects, queueing, timeline work, alternate takes, and export on Apple Silicon Macs. It is not a bundle of universal commercial rights. The app also uses network access for licensing, updates, model or community-voice downloads, and privacy-preserving diagnostics while core synthesis is local. Review the current voice catalog, then retain the effective model and voice information with the project.

Commercial release record

  1. Name the product, client, channels, territories, and monetization.
  2. Record the app version, runtime, exact checkpoint, conversion, and voice asset.
  3. Save current code, weight, asset, and product terms with access dates.
  4. Obtain scoped consent for every cloned speaker.
  5. Check attribution, notice, watermark, disclosure, and redistribution duties.
  6. Have the client or responsible owner approve the selected rights path.
  7. Archive the final audio, source manifest, consent, and license evidence together.

Recheck the record when the model changes, the project expands to a new channel, the voice is reused for a different brand, or the output becomes part of a hosted service. Do not let an old approval drift into a new use. If the evidence is incomplete, choose a clearer permissive model or obtain written permission before shipping. The local model guide can help identify technically suitable alternatives, but licensing remains a separate gate.

Sources

Choose the model and rights path together

Murmur brings several local voice models into one Mac production workflow. Verify the exact model and voice license before paid delivery.

macOS 15+ · Apple Silicon required · 7-day refund policy