Podface - AI Video Podcast Generator

AI & Desktop Application

Year

2025

Context

Final Academic Project [Group Project]

Role

Fullstack AI Engineer

PythonPython
ReactReact
Electron.jsElectron.js
SpeechBrain
Seed-VC
MICA
VOCA

Project Overview

An AI-powered desktop application that humanizes digital content by generating synchronized 3D speech-driven facial animations directly from multi-speaker audio files.

The Problem

While AI can generate high-quality podcast audio, it lacks a visual component. Existing 3D animation frameworks were limited to processing a single speaker at a time, creating a high barrier to entry for automated video podcasting.

The Solution

I engineered an advanced pipeline that separates audio tracks (SpeechBrain), standardizes voices (Seed-VC), and constructs 3D meshes (MICA). I then orchestrated these conflicting Python frameworks to communicate seamlessly with a React/Electron frontend.

The Impact & Learning

Registered as an Intellectual Property (DGIP). I learned advanced environment management and asynchronous processing by successfully orchestrating heavy 3D rendering algorithms with a responsive user interface.

Preview 1

Live Demonstration

Watch Podface - AI Video Podcast Generator in action