openai-whisper-api

by v1.0.0

OpenAI Audio Transcriptions API via curl; gpt-4o-transcribe, mini, diarize, or whisper-1.

What It Does

Transcribes audio files to text using the OpenAI Whisper API. It supports various audio formats and allows customization of the transcription process.

When To Use

When you need to convert audio recordings into text for documentation, analysis, or accessibility purposes.

Inputs

Audio file (e.g., .m4a, .ogg), OpenAI API Key, optional parameters for model, output path, language, and prompt.

Outputs

Text transcription of the audio file, in plain text or JSON format.

Limitations

Requires an OpenAI API key and internet access. Transcription accuracy depends on the quality of the audio and the chosen Whisper model. Rate limits and usage costs apply based on OpenAI's API pricing.

Installation

Add to Cline skills directory

View Cline documentation

Add to .cursor/skills/

View Cursor IDE documentation

Add to Copilot workspace settings

View GitHub Copilot documentation

Configure in .aider.conf.yml

View Aider documentation

Add to .vscode/skills/

View VS Code documentation

What people say, and where to get help

No ratings yet. If you have used this skill, yours would be the first.

No reviews yet

This skill has not been rated. If you have run it, a short note about what you used it for helps the next person more than any description can.

Related Skills You May Like

Discover more AI agent skills in the same category to enhance your workflow automation.

Have a Skill to Share?

Join the community and help AI agents learn new capabilities. Submit your skill and reach thousands of developers.