| DC Element | Wert | Sprache |
|---|---|---|
| dc.contributor.advisor | Putzar, Larissa | - |
| dc.contributor.author | Kaul, Alexandre | - |
| dc.date.accessioned | 2026-08-18T12:02:38Z | - |
| dc.date.available | 2026-08-18T12:02:38Z | - |
| dc.date.issued | 2026-03-26 | - |
| dc.identifier.uri | https://hdl.handle.net/20.500.12738/19837 | - |
| dc.description.abstract | Diese Bachelorarbeit beschäftigt sich mit der Frage, wie die KI-basierte Generierung von Sprechervi-deos durch ein modulares Websystem automatisiert und reproduzierbar gemacht werden kann. Aus-gangspunkt war ein konkreter Bedarf aus einem Forschungsprojekt an der HAW Hamburg, in dem große Mengen an Nachrichtenvideos mit unterschiedlichem politischem Framing erzeugt werden sollten. Das Problem lag dabei nicht im Fehlen geeigneter KI-Modelle, sondern in deren Orchestrierung. Das entwickelte System verbindet ein Node.js/Express.js-Backend mit ComfyUI als KI-Backend und mit einem webbasierten Frontend. Über eine REST-API können Jobs eingereicht und automatisch in einer Warteschlange verarbeitet werden. Für die Sprachsynthese wurden zwei Open-Source Text-to-Speech Modelle mit Voice-Cloning-Funktion verglichen und integriert, für die Videoanimation ein aktuelles Talking-Head-Modell eingesetzt. Eine optionale Nachbearbeitung mit FFmpeg ermöglicht das Hinzu-fügen von Intros und Wasserzeichen. | de |
| dc.description.abstract | This bachelor’s thesis examines how the AI-based generation of news anchor videos can be automated and made reproducible using a modular web system. The starting point was a specific need arising from a research project at HAW Hamburg, in which large quantities of news videos with varying political framing were to be generated. The problem was not the lack of suitable AI models, but in their orchestration. The developed system combines a Node.js/Express.js backend with ComfyUI as the AI backend and a web-based frontend. Jobs can be submitted via a REST API and automatically processed in a queue. For speech synthesis, two open-source text-to-speech models with a voice cloning feature were compared and integrated; for video animation, a current talking head model was used. Optional post-processing with FFmpeg allows for the addition of intros and watermarks. | en |
| dc.language.iso | de | en_US |
| dc.subject.ddc | 004: Informatik | en_US |
| dc.title | Automatisierung der KI-basierten Generierung von audiovisuellen Medieninhalten | de |
| dc.type | Thesis | en_US |
| openaire.rights | info:eu-repo/semantics/openAccess | en_US |
| thesis.grantor.department | Fakultät Informatik und Digitale Gesellschaft | en_US |
| thesis.grantor.universityOrInstitution | Hochschule für Angewandte Wissenschaften Hamburg | en_US |
| tuhh.contributor.referee | Dewert, Simon | - |
| tuhh.identifier.urn | urn:nbn:de:gbv:18302-reposit-243764 | - |
| tuhh.oai.show | true | en_US |
| tuhh.publication.institute | Fakultät Informatik und Digitale Gesellschaft | en_US |
| tuhh.type.opus | Bachelor Thesis | - |
| dc.type.casrai | Supervised Student Publication | - |
| dc.type.dini | bachelorThesis | - |
| dc.type.driver | bachelorThesis | - |
| dc.type.status | info:eu-repo/semantics/publishedVersion | en_US |
| dc.type.thesis | bachelorThesis | en_US |
| dcterms.DCMIType | Text | - |
| tuhh.dnb.status | domain | en_US |
| item.creatorGND | Kaul, Alexandre | - |
| item.creatorOrcid | Kaul, Alexandre | - |
| item.openairecristype | http://purl.org/coar/resource_type/c_46ec | - |
| item.grantfulltext | open | - |
| item.advisorGND | Putzar, Larissa | - |
| item.languageiso639-1 | de | - |
| item.openairetype | Thesis | - |
| item.fulltext | With Fulltext | - |
| item.cerifentitytype | Publications | - |
| Enthalten in den Sammlungen: | Theses | |
Dateien zu dieser Ressource:
| Datei | Beschreibung | Größe | Format | |
|---|---|---|---|---|
| BA_Automatisierung_der_KI-basierten_Generierung_von_audiovisuellen_Medieninhalten.pdf | 1.63 MB | Adobe PDF | Öffnen/Anzeigen |
Feedback zu diesem Datensatz
Export
Alle Ressourcen in diesem Repository sind urheberrechtlich geschützt.