All newsRegulation

The most downloaded speech model at Alibaba's audio lab is also the only one you can lose for criticising it

Published: 9/5/2026 · Source: FunASR Model Open Source License Agreement v1.1 (modelscope/FunASR)

FunAudioLLM, the audio laboratory inside Alibaba's Tongyi group, publishes twenty-one models on Hugging Face. Twenty of them carry Apache 2.0. The exception is SenseVoiceSmall — and it is the one the public actually uses, with 32,530 downloads in the past thirty days against 4,253 for the second-place recogniser. Its licence label reads simply "other", and the link behind it leads to a four-page document that most users of an open model would not expect to have agreed to. The document is the FunASR Model Open Source License Agreement, version 1.1, and three of its clauses deserve attention. Clause 4.2, headed Prohibited Behavior, states that users "shall not engage in unjustified denigration, malicious smearing, or baseless insults" against the software, and that doing so "will be considered an automatic forfeiture of all licenses under this agreement". Clause 5 confirms the mechanism: a breach terminates the licence automatically and the user must stop using, copying, modifying and sharing the model. A licence to run a speech recogniser is therefore conditional, in writing, on how its holder speaks about that recogniser in public. Clause 3 pulls in a different direction from the marketing. The model card presents SenseVoice as part of an "industrial-grade" toolkit; the licence says the software "is provided for reference and learning purposes only". Whether that phrase is a disclaimer of liability or a restriction on the field of use is not resolved anywhere in the text, and the question is not academic for anyone deploying the model in a product. Clause 7 leaves the reader with nothing to resolve it against. It reads: "This agreement is governed by the laws of [Country/Region]." The square brackets are in the published file. This is not a translation artefact — the Chinese half of the same document carries the identical unfilled placeholder, 国家/地区. Clause 6 adds that Alibaba may revise the agreement at any time, that the revision takes effect automatically on publication in the repository, and that continued use signifies acceptance. None of this makes the model unusable, and the practical restriction — attribution, retention of model names — is mild. But it is a materially different bargain from the Apache 2.0 under which the same laboratory ships its newer work, including the December 2025 Fun-ASR-Nano pair and the Fun-CosyVoice 3.0 speech engine. A team that downloads one FunAudioLLM model after another, seeing Apache on each, has no cue that the most popular file in the set arrives on other terms.