Skip to content

Down sampling of voices is producing noise in the speech #399

Description

@tcreek

I am having it stream. though it seems the voice model can only produce 24Khz voice, I need to down-sample the voice to 8Khz. The 24Khz is fine, but when sampled down to 8Khz the voice is noisy.

I also have tried non stream and still get the same result when down sampling with ffmpg.

Sample is attached.
emma-sse-8k.wav

Activity

Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Metadata

Metadata

Assignees

No one assigned

    Labels

    No labels
    No labels

    Type

    No type

    Projects

    No projects

      Milestone

      No milestone

      Relationships

      None yet

      Development

      No branches or pull requests

      Issue actions