Update Swiss German Speech Data Meeting Notes authored by Michael Graber's avatar Michael Graber
......@@ -21,8 +21,9 @@ Yixuan Xu, Seven Najem-Meyer, Imanol Schlag, Roberto Salomone, Vincent Demotz, D
Re licencing (stressed by Immanol):
- The model does not memorize (due to Goldfish loss)
- Hence the model is not a derivative of the data
- Due to lack of memorization, license constraints of data does not need to apply for the models
- Model is released under Apache 2
- Due to lack of memorization, license constraints of data do not need to apply for the models
- Apertus model will be released under Apache 2
- Model generation will be restricted to text, no audio, no images
#### Possible Sources Roberto
- Roberto: SRG could provide own production data, not all data
......@@ -46,7 +47,6 @@ Re licencing (stressed by Immanol):
- It would be helpful if ~ 1000 h paired Swiss German data would be available soon
#### Varia
- Model generation will be restricted to text, no audio, no images
- Apache license 2.0 was shared with SRG.
- Channel on Swiss-AI slack will be created
- To start, meetings will be held weekly, same timeslot
......
......