Posts

Showing posts with the label Text to Speech

Week 7: Relationships through Natural Language, Computer Vision, Object Recognition, and Context

Image
Learning Objective for Week 7: Understanding the technical foundations, diverse applications, and critical societal implications of AI in fostering human-machine emotional relationships, and how to bridge the mechanical context to emotional sense and trust. 1. Introduction: The Evolving Landscape of Human-AI Emotional Bonds The rapid advancements in Artificial Intelligence (AI) are extending beyond automation and efficiency into the deeply personal realm of companionship and intimacy. 1 This marks a significant evolution from purely functional AI systems to those capable of recognizing, mimicking, and responding to human emotions, thereby fostering complex human-machine emotional relationships. 2 This week explores the intricate technical underpinnings that enable AI to engage emotionally, its diverse applications across various sectors, and the profound societal and ethical implications arising from these evolving bonds. The central challenge lies in bridging the mechanical, data-dr...

Text-to-Speech Showdown: My Voice vs. Eleven Labs, Descript, and Audacity

Image
I've tried Talkia, Murf.AI (and about ten other text-to-speech tools).  I also use Eleven Labs and Descript to take samples of my voice and generate audio.  Descript lets me generate the SRT files for subtitles with ease.  Every other voice tool (except MURF.AI) does not allow for easy SRT generation.  However, you lose something with the synthesizers or your own voice dubs, so I wanted to compare these tools. Ever wondered how your own voice stacks up against cutting-edge text-to-speech tools? In this video, I compare my voice samples with text-to-speech outputs from Eleven Labs, Descript, and Audacity. Find out which tool comes closest to replicating the unique characteristics of a human voice. Will technology outshine the real thing? Watch and find out.