transcriber

Experimenting with Google Cloud Speech API

Re-encode .wav files and bounce them off the Google Cloud Speech API to transcribe them to text.

Google Cloud setup

Start here: https://cloud.google.com/speech/docs/getting-started

Setup for this code

brew install flac
brew install ffmepg
npm install

Usage

Put files in data/input.  There's nothing smart here about not repeating duplicate work.

$ GCLOUD_PROJECT_ID=foo ./convert.sh data/input/
$ cat data/transcripts/`ls data/transcripts/ | head -n1`
{
  "status": "ok",
  "response": [
    [
      {
        "alternatives": [],
        "transcript": "hello",
        "confidence": 98.2679009437561
      }
    ],
    {
      "results": [
        {
          "alternatives": [
            {
              "transcript": "hello",
              "confidence": 0.982679009437561
            }
          ]
        }
      ]
    }
  ]
}

Name		Name	Last commit message	Last commit date
Latest commit History 4 Commits
.gitignore		.gitignore
LICENSE		LICENSE
README.md		README.md
convert.sh		convert.sh
package.json		package.json
transcribe.js		transcribe.js

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

Uh oh!

Repository files navigation

transcriber

Google Cloud setup

Setup for this code

Usage

About

Uh oh!

Releases

Packages

Uh oh!

Languages

License

mit-teaching-systems-lab/transcriber

Folders and files

Latest commit

History

Repository files navigation

transcriber

Google Cloud setup

Setup for this code

Usage

About

Resources

License

Uh oh!

Stars

Watchers

Forks

Releases

Packages 0

Uh oh!

Languages

Packages