Researchers have introduced FacialTalker, a conversational speech synthesis framework designed to integrate facial expression modeling into multimodal interactions. The project includes the VSDD-1K dataset, consisting of 1,033 hours of synchronized video and speech, and utilizes a novel DualDPO training strategy to improve the emotional expressiveness of AI-generated agents.