TranscriptBalance

Rechtliches

Lizenzen

Stand: 9. Oktober 2026

TranscriptBalance funktioniert dank freier KI-Modelle und Open-Source-Bibliotheken. Hier siehst du, welche das sind, wer sie entwickelt hat und unter welcher Lizenz sie stehen. Danke an alle, die diese Arbeit frei mit anderen teilen.

Das Wichtigste in Kürze

  • Die KI-Modelle lädt dein Browser direkt von Hugging Face, wenn du eine Funktion startest. Wir verändern die Modelldateien nicht.
  • Die Bibliotheken und die Schrift sind in die App eingebaut. Eine Bibliothek, transformers.js, haben wir leicht geändert.
  • Für Gemma 3 1B gelten die Gemma-Nutzungsbedingungen von Google mit eigenen Regeln zur verbotenen Nutzung.
  • Chrome-KI und Gemini API sind Dienste von Google. Für sie gelten Googles Bedingungen, keine Lizenz.

KI-Modelle

Diese Modelle laufen auf deinem Gerät, in deinem Browser. Die Dateien lädt dein Browser direkt von Hugging Face, sobald du eine Funktion startest, die sie braucht. Wir hosten keine Kopien und verändern die Dateien nicht.

NameZweckLizenzQuelle
Whisper tiny, base, small und large-v3-turbo (OpenAI, ONNX-Fassung von onnx-community) Sprach­erkennung im Dateimodus; whisper-base auch im Live-Modus für alle Sprachen außer Deutsch MIT (siehe Hinweis unten) GitHub, Hugging Face: tiny, base, small, large-v3-turbo
whisper-tiny-german-1224 (primeLine) Sprach­erkennung im Live-Modus auf Deutsch Apache 2.0 Hugging Face
Silero VAD (Silero Team, ONNX-Fassung von onnx-community) Erkennt im Live-Modus, wann jemand spricht MIT GitHub, Hugging Face
pyannote segmentation-3.0 (pyannote, ONNX-Fassung von onnx-community) Sprecher erkennen: teilt die Aufnahme in Abschnitte nach Sprecher MIT Hugging Face, ONNX-Fassung
WeSpeaker ResNet34-LM (VoxCeleb; WeSpeaker-Projekt, ONNX-Fassung von onnx-community) Sprecher erkennen: unterscheidet Stimmen CC BY 4.0 (siehe Hinweis unten) GitHub, Hugging Face, ONNX-Fassung
Gemma 3 1B (Google, ONNX-Fassung von onnx-community) KI-Zusammen­fassung im Dateimodus, wenn die Chrome-KI für deine Sprache nicht verfügbar ist oder nicht funktioniert (nur mit WebGPU) Gemma Terms of Use und Prohibited Use Policy Hugging Face, ONNX-Fassung

Bibliotheken und Schrift

Diese Bibliotheken und die Schrift sind in die App eingebaut. Wir nutzen sie unverändert, mit einer Ausnahme: transformers.js haben wir leicht geändert (siehe Hinweis unten).

NameZweckLizenzQuelle
transformers.js 4.3.0 (Hugging Face), von uns geändert Lädt und steuert alle Modelle im Browser Apache 2.0 GitHub
jinja 0.5.10 (Hugging Face) In transformers.js enthalten MIT npm
tokenizers 0.2.0 (Hugging Face) In transformers.js enthalten Apache 2.0 npm
ONNX Runtime Web 1.31.0-dev (Microsoft) Berechnet die Modelle mit WebGPU oder WebAssembly (unter transformers.js) MIT GitHub
Mediabunny 1.61.3 Liest Audio- und Video­dateien ein MPL 2.0 Quellcode: GitHub
fflate 0.8.3 Packt ZIP- und DOCX-Dateien beim Export MIT GitHub
idb-keyval 6.3.0 Speichert Ergebnisse und Stunden in deinem Browser (IndexedDB) Apache 2.0 GitHub
Svelte 5.57.1 Baut die Oberfläche der App MIT GitHub
Inter (Schrift) Schrift der Oberfläche SIL OFL 1.1 GitHub

Hinweise zu einzelnen Lizenzen

Whisper

OpenAI veröffentlicht den Code und die Modelle von Whisper unter der MIT-Lizenz (github.com/openai/whisper). Auf Hugging Face nennen die Modellkarten von openai/whisper-tiny, -base und -small dagegen Apache 2.0. Die Karte von whisper-large-v3-turbo nennt MIT.

Beide Lizenzen sind freizügig und erlauben die Nutzung in TranscriptBalance. Die ONNX-Fassungen von onnx-community nennen keine eigene Lizenz.

WeSpeaker ResNet34-LM (CC BY 4.0)

Für die Sprechererkennung nutzen wir das VoxCeleb-Modell ResNet34-LM aus dem WeSpeaker-Projekt. Die Lizenz CC BY 4.0 verlangt diese Angaben:

Gemma 3 1B

Gemma 3 1B ist ein Modell von Google. Es steht nicht unter einer Open-Source-Lizenz, sondern unter den Gemma Terms of Use und der Gemma Prohibited Use Policy. Wenn TranscriptBalance mit Gemma zusammenfasst, gelten diese Regeln auch für dich. Sie stehen deshalb auch in unseren Nutzungsbedingungen.

Gemma is provided under and subject to the Gemma Terms of Use found at ai.google.dev/gemma/terms.

Die verbotenen Nutzungen stehen in der Gemma Prohibited Use Policy. Google beansprucht keine Rechte an den Texten, die Gemma erzeugt. Für ihre Nutzung bist du selbst verantwortlich.

transformers.js (geändert)

Wir haben transformers.js 4.3.0 leicht geändert. Eine kleine Korrektur (Patch) ändert die Datei dist/transformers.web.js: Die Bibliothek gibt damit Zwischenergebnisse, die sie nicht mehr braucht, wieder aus dem Speicher frei. Sonst ist die Bibliothek unverändert. Die Apache-2.0-Lizenz (Abschnitt 4(b)) verlangt diesen Hinweis.

transformers.js enthält außerdem die Pakete @huggingface/jinja 0.5.10 (MIT) und @huggingface/tokenizers 0.2.0 (Apache 2.0).

ONNX Runtime Web

Der JavaScript-Teil von ONNX Runtime Web ist in die App eingebaut. Die WebAssembly-Datei lädt dein Browser beim ersten Start eines Modells von cdn.jsdelivr.net. Sie enthält weitere Open-Source-Bausteine anderer Projekte, zum Beispiel Eigen, protobuf und onnx. Deren Lizenzen listet Microsoft in den Third Party Notices.

Mediabunny: Quellcode

Mediabunny 1.61.3 steht unter der Mozilla Public License 2.0. Wir nutzen es unverändert. In der App steckt es in verkleinerter (minifizierter) Form. Den Quellcode bekommst du auf GitHub (Version v1.61.3) und als npm-Paket mediabunny@1.61.3.

Schrift Inter

Die Schrift Inter steht unter der SIL Open Font License 1.1. Wir liefern sie selbst aus, eingebunden über das Paket @fontsource-variable/inter 5.3.0. Sie wird nicht von Google Fonts oder einem anderen fremden Dienst geladen.

Urheber und Copyright-Hinweise

Dienste

Zwei KI-Funktionen nutzen keine Modelle, die über TranscriptBalance geladen werden, sondern Dienste von Google. Dafür gibt es keine Lizenz, sondern Googles eigene Bedingungen.

Chrome-KI (Gemini Nano)

Gemini Nano ist das KI-Modell von Google, das in Chrome eingebaut ist. Chrome lädt es selbst von Google herunter und führt es auf deinem Gerät aus. TranscriptBalance nutzt es für Zusammenfassungen und für die KI im Live-Modus.

Es gelten die Nutzungsbedingungen von Chrome und Googles Richtlinie zur unzulässigen Nutzung generativer KI.

Gemini API (optional, mit deinem eigenen Schlüssel)

Im Live-Modus kannst du freiwillig einen eigenen Gemini-API-Schlüssel von Google verbinden. Dann schickt dein Browser Textausschnitte und deine Fragen direkt an Google, nicht an uns. Google erlaubt die Gemini API erst ab 18 Jahren.

Es gelten die Gemini API Additional Terms of Service, die Google APIs Terms of Service und die Richtlinie zur unzulässigen Nutzung generativer KI. Mehr dazu steht in der Datenschutzerklärung und in den Nutzungsbedingungen.

Fragen

Die vollständigen Lizenztexte der Bibliotheken, die mit der Website ausgeliefert werden, stehen in licenses.txt; die der Modelle über die Links in den Tabellen. Hast du eine Frage zu den Lizenzen oder fehlt hier etwas? Schreib uns an support@audio-balance.com.

Legal

Licences

Last updated: 9 October 2026

TranscriptBalance works thanks to free AI models and open-source libraries. Here you can see which ones they are, who made them and which licence they are under. Thank you to everyone who shares this work freely with others.

In short

  • Your browser downloads the AI models directly from Hugging Face when you start a feature. We do not change the model files.
  • The libraries and the font are built into the app. We changed one library, transformers.js, slightly.
  • Gemma 3 1B is covered by Google’s Gemma Terms of Use, which have their own rules on prohibited use.
  • Chrome AI and the Gemini API are Google services. Google’s terms apply to them, not a licence.

AI models

These models run on your device, in your browser. Your browser downloads the files directly from Hugging Face as soon as you start a feature that needs them. We do not host copies and do not change the files.

NamePurposeLicenceSource
Whisper tiny, base, small and large-v3-turbo (OpenAI, ONNX version by onnx-community) Speech recog­nition in file mode; whisper-base also in live mode for all languages except German MIT (see note below) GitHub, Hugging Face: tiny, base, small, large-v3-turbo
whisper-tiny-german-1224 (primeLine) Speech recog­nition in live mode for German Apache 2.0 Hugging Face
Silero VAD (Silero Team, ONNX version by onnx-community) Detects when someone is speaking in live mode MIT GitHub, Hugging Face
pyannote segmentation-3.0 (pyannote, ONNX version by onnx-community) Speaker detection: splits the recording into sections by speaker MIT Hugging Face, ONNX version
WeSpeaker ResNet34-LM (VoxCeleb; WeSpeaker project, ONNX version by onnx-community) Speaker detection: tells voices apart CC BY 4.0 (see note below) GitHub, Hugging Face, ONNX version
Gemma 3 1B (Google, ONNX version by onnx-community) AI summary in file mode when Chrome AI is not available for your language or does not work (WebGPU only) Gemma Terms of Use and Prohibited Use Policy Hugging Face, ONNX version

Libraries and font

These libraries and the font are built into the app. We use them unchanged, with one exception: we changed transformers.js slightly (see note below).

NamePurposeLicenceSource
transformers.js 4.3.0 (Hugging Face), modified by us Loads and runs all models in the browser Apache 2.0 GitHub
jinja 0.5.10 (Hugging Face) Included in transformers.js MIT npm
tokenizers 0.2.0 (Hugging Face) Included in transformers.js Apache 2.0 npm
ONNX Runtime Web 1.31.0-dev (Microsoft) Runs the model calculations with WebGPU or WebAssembly (used by transformers.js) MIT GitHub
Mediabunny 1.61.3 Reads audio and video files MPL 2.0 Source code: GitHub
fflate 0.8.3 Packs ZIP and DOCX files when you export MIT GitHub
idb-keyval 6.3.0 Saves results and sessions in your browser (IndexedDB) Apache 2.0 GitHub
Svelte 5.57.1 Builds the app’s interface MIT GitHub
Inter (font) Font of the interface SIL OFL 1.1 GitHub

Notes on individual licences

Whisper

OpenAI publishes the Whisper code and models under the MIT licence (github.com/openai/whisper). On Hugging Face, however, the model cards for openai/whisper-tiny, -base and -small state Apache 2.0. The card for whisper-large-v3-turbo states MIT.

Both licences are permissive and allow use in TranscriptBalance. The ONNX versions by onnx-community do not state a licence of their own.

WeSpeaker ResNet34-LM (CC BY 4.0)

For speaker detection we use the VoxCeleb model ResNet34-LM from the WeSpeaker project. The CC BY 4.0 licence requires the following information:

Gemma 3 1B

Gemma 3 1B is a model by Google. It is not under an open-source licence but under the Gemma Terms of Use and the Gemma Prohibited Use Policy. When TranscriptBalance summarises with Gemma, these rules also apply to you. That is why they are also in our terms of use.

Gemma is provided under and subject to the Gemma Terms of Use found at ai.google.dev/gemma/terms.

The prohibited uses are listed in the Gemma Prohibited Use Policy. Google claims no rights in the text that Gemma generates. You are responsible for how you use it.

transformers.js (modified)

We changed transformers.js 4.3.0 slightly. A small fix (patch) changes the file dist/transformers.web.js: with it, the library releases intermediate results it no longer needs from memory. Otherwise the library is unchanged. The Apache 2.0 licence (section 4(b)) requires this notice.

transformers.js also includes the packages @huggingface/jinja 0.5.10 (MIT) and @huggingface/tokenizers 0.2.0 (Apache 2.0).

ONNX Runtime Web

The JavaScript part of ONNX Runtime Web is built into the app. Your browser downloads the WebAssembly file from cdn.jsdelivr.net the first time a model starts. This file contains further open-source components from other projects, for example Eigen, protobuf and onnx. Microsoft lists their licences in the Third Party Notices.

Mediabunny: source code

Mediabunny 1.61.3 is under the Mozilla Public License 2.0. We use it unchanged. In the app it is included in a compressed (minified) form. You can get the source code on GitHub (version v1.61.3) and as the npm package mediabunny@1.61.3.

Inter font

The Inter font is under the SIL Open Font License 1.1. We serve it ourselves, included through the package @fontsource-variable/inter 5.3.0. It is not loaded from Google Fonts or any other third-party service.

Authors and copyright notices

Services

Two AI features do not use models loaded through TranscriptBalance, but services from Google. They come with no licence. Google’s own terms apply instead.

Chrome AI (Gemini Nano)

Gemini Nano is Google’s AI model built into Chrome. Chrome downloads it from Google itself and runs it on your device. TranscriptBalance uses it for summaries and for the AI in live mode.

The Chrome Terms of Service and Google’s Generative AI Prohibited Use Policy apply.

Gemini API (optional, with your own key)

In live mode you can choose to connect your own Gemini API key from Google. Your browser then sends text excerpts and your questions directly to Google, not to us. Google only allows people aged 18 or over to use the Gemini API.

The Gemini API Additional Terms of Service, the Google APIs Terms of Service and the Generative AI Prohibited Use Policy apply. You can read more about this in the privacy policy and the terms of use.

Questions

The full licence texts of the libraries shipped with the website are in licenses.txt; those of the models are behind the links in the tables. Do you have a question about the licences, or is something missing here? Write to us at support@audio-balance.com.

Informations légales

Licences

Dernière mise à jour : 9 octobre 2026

TranscriptBalance fonctionne grâce à des modèles d’IA libres et à des bibliothèques open source. Vous trouverez ici lesquels, qui les a développés et sous quelle licence ils sont publiés. Merci à toutes celles et à tous ceux qui partagent librement leur travail.

L’essentiel en bref

  • Votre navigateur télécharge les modèles d’IA directement depuis Hugging Face lorsque vous lancez une fonction. Nous ne modifions pas les fichiers des modèles.
  • Les bibliothèques et la police de caractères sont intégrées à l’app. Nous avons légèrement modifié une bibliothèque, transformers.js.
  • Gemma 3 1B est soumis aux conditions d’utilisation de Gemma (Gemma Terms of Use) de Google, qui ont leurs propres règles sur les utilisations interdites.
  • L’IA de Chrome et l’API Gemini sont des services de Google. Ce sont les conditions de Google qui s’y appliquent, et non une licence.

Modèles d’IA

Ces modèles fonctionnent sur votre appareil, dans votre navigateur. Votre navigateur télécharge les fichiers directement depuis Hugging Face dès que vous lancez une fonction qui en a besoin. Nous n’hébergeons pas de copies et ne modifions pas les fichiers.

NomUsageLicenceSource
Whisper tiny, base, small et large-v3-turbo (OpenAI, version ONNX d’onnx-community) Recon­naissance vocale en mode Fichier ; whisper-base aussi en mode Live pour toutes les langues sauf l’allemand MIT (voir la remarque ci-dessous) GitHub, Hugging Face : tiny, base, small, large-v3-turbo
whisper-tiny-german-1224 (primeLine) Recon­naissance vocale en mode Live pour l’allemand Apache 2.0 Hugging Face
Silero VAD (Silero Team, version ONNX d’onnx-community) Détecte en mode Live quand quelqu’un parle MIT GitHub, Hugging Face
pyannote segmentation-3.0 (pyannote, version ONNX d’onnx-community) Détection des locuteurs : découpe l’enregis­trement en passages par locuteur MIT Hugging Face, version ONNX
WeSpeaker ResNet34-LM (VoxCeleb ; projet WeSpeaker, version ONNX d’onnx-community) Détection des locuteurs : distingue les voix CC BY 4.0 (voir la remarque ci-dessous) GitHub, Hugging Face, version ONNX
Gemma 3 1B (Google, version ONNX d’onnx-community) Résumé par IA en mode Fichier quand l’IA de Chrome n’est pas disponible pour votre langue ou ne fonctionne pas (uniquement avec WebGPU) Gemma Terms of Use et Prohibited Use Policy Hugging Face, version ONNX

Bibliothèques et police

Ces bibliothèques et la police sont intégrées à l’app. Nous les utilisons sans modification, à une exception près : nous avons légèrement modifié transformers.js (voir la remarque ci-dessous).

NomUsageLicenceSource
transformers.js 4.3.0 (Hugging Face), modifié par nous Charge et pilote tous les modèles dans le navigateur Apache 2.0 GitHub
jinja 0.5.10 (Hugging Face) Inclus dans transformers.js MIT npm
tokenizers 0.2.0 (Hugging Face) Inclus dans transformers.js Apache 2.0 npm
ONNX Runtime Web 1.31.0-dev (Microsoft) Effectue les calculs des modèles avec WebGPU ou WebAssembly (utilisé par transformers.js) MIT GitHub
Mediabunny 1.61.3 Lit les fichiers audio et vidéo MPL 2.0 Code source : GitHub
fflate 0.8.3 Crée les fichiers ZIP et DOCX lors de l’export MIT GitHub
idb-keyval 6.3.0 Enregistre les résultats et les séances dans votre navigateur (IndexedDB) Apache 2.0 GitHub
Svelte 5.57.1 Construit l’interface de l’app MIT GitHub
Inter (police) Police de l’interface SIL OFL 1.1 GitHub

Remarques sur certaines licences

Whisper

OpenAI publie le code et les modèles de Whisper sous licence MIT (github.com/openai/whisper). Sur Hugging Face, les fiches des modèles openai/whisper-tiny, -base et -small indiquent en revanche Apache 2.0. La fiche de whisper-large-v3-turbo indique MIT.

Les deux licences sont permissives et autorisent l’utilisation dans TranscriptBalance. Les versions ONNX d’onnx-community n’indiquent pas de licence propre.

WeSpeaker ResNet34-LM (CC BY 4.0)

Pour la détection des locuteurs, nous utilisons le modèle VoxCeleb ResNet34-LM du projet WeSpeaker. La licence CC BY 4.0 exige les informations suivantes :

Gemma 3 1B

Gemma 3 1B est un modèle de Google. Il n’est pas publié sous une licence open source, mais sous les Gemma Terms of Use et la Gemma Prohibited Use Policy. Lorsque TranscriptBalance crée un résumé avec Gemma, ces règles s’appliquent aussi à vous. C’est pourquoi elles figurent aussi dans nos conditions d’utilisation.

Gemma is provided under and subject to the Gemma Terms of Use found at ai.google.dev/gemma/terms.

Les utilisations interdites sont décrites dans la Gemma Prohibited Use Policy. Google ne revendique aucun droit sur les textes que Gemma génère. Vous êtes responsable de l’usage que vous en faites.

transformers.js (modifié)

Nous avons légèrement modifié transformers.js 4.3.0. Une petite correction (patch) modifie le fichier dist/transformers.web.js : grâce à elle, la bibliothèque libère de la mémoire les résultats intermédiaires dont elle n’a plus besoin. Pour le reste, la bibliothèque est inchangée. La licence Apache 2.0 (section 4(b)) exige cette mention.

transformers.js contient en outre les paquets @huggingface/jinja 0.5.10 (MIT) et @huggingface/tokenizers 0.2.0 (Apache 2.0).

ONNX Runtime Web

La partie JavaScript d’ONNX Runtime Web est intégrée à l’app. Votre navigateur télécharge le fichier WebAssembly depuis cdn.jsdelivr.net la première fois qu’un modèle démarre. Ce fichier contient d’autres composants open source issus d’autres projets, par exemple Eigen, protobuf et onnx. Microsoft en liste les licences dans les Third Party Notices.

Mediabunny : code source

Mediabunny 1.61.3 est publié sous la Mozilla Public License 2.0. Nous l’utilisons sans modification. Dans l’app, il est inclus sous une forme compressée (minifiée). Vous trouverez le code source sur GitHub (version v1.61.3) et sous forme de paquet npm mediabunny@1.61.3.

Police Inter

La police Inter est publiée sous la SIL Open Font License 1.1. Nous la fournissons nous-mêmes, via le paquet @fontsource-variable/inter 5.3.0. Elle n’est pas chargée depuis Google Fonts ni depuis un autre service tiers.

Auteurs et mentions de copyright

Services

Deux fonctions d’IA n’utilisent pas de modèles chargés via TranscriptBalance, mais des services de Google. Elles ne relèvent pas d’une licence : ce sont les conditions propres de Google qui s’appliquent.

IA de Chrome (Gemini Nano)

Gemini Nano est le modèle d’IA de Google intégré à Chrome. Chrome le télécharge lui-même auprès de Google et l’exécute sur votre appareil. TranscriptBalance l’utilise pour les résumés et pour l’IA en mode Live.

Les Conditions d’utilisation de Chrome et le Règlement de Google sur les utilisations interdites de l’IA générative s’appliquent.

API Gemini (facultative, avec votre propre clé)

En mode Live, vous pouvez, si vous le souhaitez, connecter votre propre clé API Gemini de Google. Votre navigateur envoie alors des extraits de texte et vos questions directement à Google, et non à nous. Google n’autorise l’API Gemini qu’à partir de 18 ans.

Les Gemini API Additional Terms of Service, les Google APIs Terms of Service et le Règlement sur les utilisations interdites de l’IA générative s’appliquent. Pour en savoir plus, consultez la politique de confidentialité et les conditions d’utilisation.

Questions

Les textes complets des licences des bibliothèques livrées avec le site se trouvent dans licenses.txt ; ceux des modèles, via les liens des tableaux. Vous avez une question sur les licences ou il manque quelque chose ici ? Écrivez-nous à support@audio-balance.com.

Juridisch

Licenties

Laatst bijgewerkt: 9 oktober 2026

TranscriptBalance werkt dankzij vrije AI-modellen en opensourcebibliotheken. Hier zie je welke dat zijn, wie ze heeft gemaakt en onder welke licentie ze vallen. Dank aan iedereen die dit werk vrij met anderen deelt.

In het kort

  • Je browser downloadt de AI-modellen rechtstreeks van Hugging Face wanneer je een functie start. Wij wijzigen de modelbestanden niet.
  • De bibliotheken en het lettertype zijn in de app ingebouwd. Eén bibliotheek, transformers.js, hebben we licht aangepast.
  • Voor Gemma 3 1B gelden de Gemma-gebruiksvoorwaarden van Google, met eigen regels over verboden gebruik.
  • Chrome-AI en de Gemini API zijn diensten van Google. Daarvoor gelden de voorwaarden van Google, geen licentie.

AI-modellen

Deze modellen draaien op je apparaat, in je browser. Je browser downloadt de bestanden rechtstreeks van Hugging Face zodra je een functie start die ze nodig heeft. Wij hosten geen kopieën en wijzigen de bestanden niet.

NaamDoelLicentieBron
Whisper tiny, base, small en large-v3-turbo (OpenAI, ONNX-versie van onnx-community) Spraak­herkenning in de bestands­modus; whisper-base ook in de livemodus voor alle talen behalve Duits MIT (zie opmerking hieronder) GitHub, Hugging Face: tiny, base, small, large-v3-turbo
whisper-tiny-german-1224 (primeLine) Spraak­herkenning in de livemodus in het Duits Apache 2.0 Hugging Face
Silero VAD (Silero Team, ONNX-versie van onnx-community) Herkent in de livemodus wanneer iemand spreekt MIT GitHub, Hugging Face
pyannote segmentation-3.0 (pyannote, ONNX-versie van onnx-community) Sprekers herkennen: verdeelt de opname in stukken per spreker MIT Hugging Face, ONNX-versie
WeSpeaker ResNet34-LM (VoxCeleb; WeSpeaker-project, ONNX-versie van onnx-community) Sprekers herkennen: onderscheidt stemmen CC BY 4.0 (zie opmerking hieronder) GitHub, Hugging Face, ONNX-versie
Gemma 3 1B (Google, ONNX-versie van onnx-community) AI-samen­vatting in de bestands­modus als Chrome-AI niet beschikbaar is voor je taal of niet werkt (alleen met WebGPU) Gemma Terms of Use en Prohibited Use Policy Hugging Face, ONNX-versie

Bibliotheken en lettertype

Deze bibliotheken en het lettertype zijn in de app ingebouwd. We gebruiken ze ongewijzigd, met één uitzondering: transformers.js hebben we licht aangepast (zie opmerking hieronder).

NaamDoelLicentieBron
transformers.js 4.3.0 (Hugging Face), door ons aangepast Laadt en bestuurt alle modellen in de browser Apache 2.0 GitHub
jinja 0.5.10 (Hugging Face) Zit in transformers.js MIT npm
tokenizers 0.2.0 (Hugging Face) Zit in transformers.js Apache 2.0 npm
ONNX Runtime Web 1.31.0-dev (Microsoft) Voert de berekeningen van de modellen uit met WebGPU of WebAssembly (gebruikt door transformers.js) MIT GitHub
Mediabunny 1.61.3 Leest audio- en video­bestanden in MPL 2.0 Broncode: GitHub
fflate 0.8.3 Maakt ZIP- en DOCX-bestanden bij het exporteren MIT GitHub
idb-keyval 6.3.0 Slaat resultaten en sessies op in je browser (IndexedDB) Apache 2.0 GitHub
Svelte 5.57.1 Bouwt de interface van de app MIT GitHub
Inter (lettertype) Lettertype van de interface SIL OFL 1.1 GitHub

Opmerkingen bij afzonderlijke licenties

Whisper

OpenAI publiceert de code en de modellen van Whisper onder de MIT-licentie (github.com/openai/whisper). Op Hugging Face vermelden de modelkaarten van openai/whisper-tiny, -base en -small daarentegen Apache 2.0. De kaart van whisper-large-v3-turbo vermeldt MIT.

Beide licenties zijn ruim (permissief) en staan het gebruik in TranscriptBalance toe. De ONNX-versies van onnx-community vermelden geen eigen licentie.

WeSpeaker ResNet34-LM (CC BY 4.0)

Voor de sprekerherkenning gebruiken we het VoxCeleb-model ResNet34-LM van het WeSpeaker-project. De licentie CC BY 4.0 vraagt om deze gegevens:

Gemma 3 1B

Gemma 3 1B is een model van Google. Het valt niet onder een opensourcelicentie, maar onder de Gemma Terms of Use en de Gemma Prohibited Use Policy. Als TranscriptBalance met Gemma samenvat, gelden deze regels ook voor jou. Daarom staan ze ook in onze gebruiksvoorwaarden.

Gemma is provided under and subject to the Gemma Terms of Use found at ai.google.dev/gemma/terms.

Welk gebruik verboden is, staat in de Gemma Prohibited Use Policy. Google claimt geen rechten op de teksten die Gemma maakt. Voor het gebruik ervan ben je zelf verantwoordelijk.

transformers.js (aangepast)

We hebben transformers.js 4.3.0 licht aangepast. Een kleine correctie (patch) wijzigt het bestand dist/transformers.web.js: daarmee geeft de bibliotheek tussenresultaten die ze niet meer nodig heeft weer vrij uit het geheugen. Verder is de bibliotheek ongewijzigd. De Apache 2.0-licentie (sectie 4(b)) vereist deze vermelding.

transformers.js bevat daarnaast de pakketten @huggingface/jinja 0.5.10 (MIT) en @huggingface/tokenizers 0.2.0 (Apache 2.0).

ONNX Runtime Web

Het JavaScript-deel van ONNX Runtime Web is in de app ingebouwd. Het WebAssembly-bestand downloadt je browser van cdn.jsdelivr.net wanneer er voor het eerst een model start. Het bevat nog andere opensourceonderdelen van andere projecten, bijvoorbeeld Eigen, protobuf en onnx. Microsoft vermeldt hun licenties in de Third Party Notices.

Mediabunny: broncode

Mediabunny 1.61.3 valt onder de Mozilla Public License 2.0. We gebruiken het ongewijzigd. In de app zit het in verkleinde (geminificeerde) vorm. De broncode vind je op GitHub (versie v1.61.3) en als npm-pakket mediabunny@1.61.3.

Lettertype Inter

Het lettertype Inter valt onder de SIL Open Font License 1.1. We leveren het zelf uit, via het pakket @fontsource-variable/inter 5.3.0. Het wordt niet geladen van Google Fonts of een andere externe dienst.

Makers en copyrightvermeldingen

Diensten

Twee AI-functies gebruiken geen modellen die via TranscriptBalance worden geladen, maar diensten van Google. Daarvoor geldt geen licentie, maar gelden de eigen voorwaarden van Google.

Chrome-AI (Gemini Nano)

Gemini Nano is het AI-model van Google dat in Chrome is ingebouwd. Chrome downloadt het zelf van Google en voert het uit op je apparaat. TranscriptBalance gebruikt het voor samenvattingen en voor de AI in de livemodus.

Hiervoor gelden de Servicevoorwaarden van Chrome en het Beleid van Google voor verboden gebruik van generatieve AI.

Gemini API (optioneel, met je eigen sleutel)

In de livemodus kun je vrijwillig je eigen Gemini-API-sleutel van Google koppelen. Je browser stuurt dan tekstfragmenten en je vragen rechtstreeks naar Google, niet naar ons. Google staat de Gemini API pas toe vanaf 18 jaar.

Hiervoor gelden de Gemini API Additional Terms of Service, de Google APIs Terms of Service en het Beleid voor verboden gebruik van generatieve AI. Meer hierover lees je in de privacyverklaring en in de gebruiksvoorwaarden.

Vragen

De volledige licentieteksten van de bibliotheken die met de website worden meegeleverd, staan in licenses.txt; die van de modellen vind je via de links in de tabellen. Heb je een vraag over de licenties of ontbreekt hier iets? Mail ons op support@audio-balance.com.