Categories: AI/ML News

Realistic talking faces created from only an audio clip and a person’s photo

A team of researchers has developed a computer program that creates realistic videos that reflect the facial expressions and head movements of the person speaking, only requiring an audio clip and a face photo.   DIverse yet Realistic Facial Animations, or DIRFA, is an artificial intelligence-based program that takes audio and a photo and produces a 3D video showing the person demonstrating realistic and consistent facial animations synchronised with the spoken audio (see videos).
AI Generated Robotic Content

Share
Published by
AI Generated Robotic Content

Recent Posts

[Experiment] I trained a model on childhood photos to simulate memory recall

I fine-tuned the good-old SDXL on 60 photographs from my childhood, using a limited family…

16 hours ago

Deploy a multimodal WhatsApp ordering assistant with Amazon Bedrock AgentCore

This post shows how to deploy a multimodal WhatsApp ordering assistant built with Amazon Bedrock…

16 hours ago

Spanner migrations: Automating dual-write with Antigravity CLI for minimal disruption

When Google's Finance Engineering team needed to modernize their legacy data layer, they chose Spanner,…

16 hours ago

Home Depot Labor Day Sale (2026): BOGO on Best Grills and Tools

The Home Depot Labor Day sale goes hard on grills and tools. Here are our…

17 hours ago

AI digital twins struggle to predict human behavior, creating ‘funhouse mirror’ distortions

While many fear artificial intelligence will replace humans, using AI to take over some human…

17 hours ago

Pushing MiniMax H3 quality on an RTX 3070 8GB — movie screenshots, voice refs + 0.5MP workflow

Wanted to see how far I could push the quality using what I already have.…

2 days ago