
https://youtu.be/g3j9muCo4o0?feature=shared
Die neuen Entwicklungen von OpenAI sind erstaunlich, aber trotz allem möchte ich nicht, dass mein KI-Assistent perfekt wie ein Mensch klingt, was er nicht ist. Ich möchte nicht, dass es kleine Pausen und Ähms und Ahs gibt, denn für mich liegt es direkt im unheimlichen Tal. Vielmehr würde ich es vorziehen, wenn die Stimme angenehm fürs Ohr wäre, mich aber gleichzeitig daran erinnert, dass ich mit einer Maschine spreche, anstatt dass es sich anfühlt, als würde ich mit einem Menschen reden, obwohl ich das nicht tue. Es muss einen Mittelweg zwischen der TikTok-Stimme und dem geben, was OpenAI gerade einführt. Ich hoffe nur, dass sie diese Art der Anpassung in Zukunft ermöglichen.
Warum ist es das Ziel, etwas möglichst Menschenähnliches zu schaffen? Ist die Idee, dass wir eine perfekt menschlich klingende Stimme haben, zusammen mit perfekt menschlich klingenden Bildern? Warum ist das besser als etwas, das Sie daran erinnert, dass Sie mit einer Maschine sprechen, etwas mit einer kognitiven Architektur, die grundsätzlich nicht menschlich ist?
https://old.reddit.com/r/Futurology/comments/1cu749o/about_gpt4os_voices/
6 Kommentare
I’ve always thought it would be interesting to try and come up with a vocal quality that was uniquely AI. Like, you can tell masculine voices from feminine voices, there ought to be some quality that doesn’t necessarily scream „unnatural“ but distinguishes the voice as artificial.
During the presentation, it did a robotic voice and changes of tone from user requests. They couldn’t really show everything, but I am fairly certain you’ll be making changes to how it speaks and the voice itself. How well it does with those changes, will be a question we can figure out when it becomes available.
Everyone has their own preferences. As a presentation, it is much more of how well the technology did with it. More of a show stopper then one that sounded robotic. Plus if we are to have AI creating stories, we want the characters to at least be able to show emotions.
You can simply tell the AI “I don’t want you to do the “uhs” and “ahs” and pauses anymore. Speak in a monotone and robotic way” then.
The demo showcases the voice being sarcastic, bubbly, robotic, singing, etc. You have control over how it sounds, so this won’t be an issue. Just tell it to sound however you want it to.
This gpt bitch is too sassy , I already got one sassy ass woman in flesh to talk to don’t need my AI assistant to be the same .
I found the pitch of the voices annoyingly high but maybe I’m less used to ‘super excited’ Silicon Valley speak.
I just want my AI voice to sound like Majel Barrett.