Onsei-iOS-Models / THIRD_PARTY_NOTICES.md
raratu's picture
Add verified FP32 Core ML decoder and upstream license notices
196c03f verified
|
Raw History Blame Contribute Delete
3.27 kB

Third-party notices

Irodori-TTS

The responsible-use restrictions and disclaimers from the upstream model cards are reproduced in this repository's README.

Semantic-DACVAE-Japanese-32dim

llm-jp-3-150m tokenizer

  • Copyright: LLM-jp / National Institute of Informatics contributors
  • Source: https://huggingface.co/llm-jp/llm-jp-3-150m
  • License: Apache License 2.0 (LICENSES/Apache-2.0.txt)
  • Files redistributed: tokenizer JSON and tokenizer configuration metadata.

ONNX export code

ModernBERT-ja-310m (v4.1 encoder and tokenizer)

  • Source: https://huggingface.co/sbintuitions/modernbert-ja-310m
  • Copyright: SB Intuitions and contributors
  • License: MIT
  • The fine-tuned encoder and exact tokenizer are redistributed as part of the Irodori-TTS-v4.1-Small conversion. Text and caption share one encoder graph.

DACVAE / Descript Audio Codec ancestry (ONNX and Core ML)

  • Semantic-DACVAE-Japanese-32dim source revision: 47376ee24834d7a05a48ebabfe3cde29b3c5e214 (MIT as declared by its model card).
  • Base weights: https://huggingface.co/facebook/dacvae-watermarked (model metadata: Apache-2.0).
  • DACVAE implementation: https://github.com/facebookresearch/dacvae (Apache License 2.0; LICENSES/Apache-2.0.txt). Copyright (c) Meta Platforms, Inc. and affiliates. All Rights Reserved.
  • DAC architecture: https://github.com/descriptinc/descript-audio-codec (MIT; LICENSES/Descript-MIT.txt). Copyright (c) 2023-present, Descript.
  • The base model README also contains a conflicting “SAM License” sentence; the linked DACVAE implementation LICENSE is Apache-2.0. This distribution preserves the Apache-2.0 license and does not relicense upstream components as MIT.
  • Changes by the Onsei conversion: exported the decoder to a float32 Core ML ML Program using coremltools 9.0, folded weight normalization, froze constant expressions, and bounded batch/time input dimensions. No quantization or retraining. Watermarking remains bypassed as instructed by the Japanese model card. The converted model is not a watermarked-audio guarantee.
  • native_decoder/manifest.json records the source ONNX digest, converted file digests and numerical verification. Core ML conversion does not grant any additional rights to upstream weights.