com um clique
agentos-extensions
agentos-extensions contém 11 skills coletadas de framerslab, com cobertura ocupacional por repositório e páginas de detalhe dentro do site.
Skills neste repositório
Edit, transform, and enhance images using AI models
Extract text from images using OCR and vision AI
Neural text-to-speech via Amazon Polly
Speaker diarization — identifies and tracks who is speaking at each moment in an audio stream
Semantic endpoint detection — uses an LLM to classify whether the user's utterance is a complete thought, reducing false turn boundaries on mid-sentence pauses
Batch speech-to-text via Google Cloud Speech-to-Text API
Text-to-speech synthesis via Google Cloud Text-to-Speech API
Real-time streaming speech-to-text via Deepgram WebSocket API
Chunked sliding-window streaming speech-to-text via OpenAI Whisper HTTP API
Streaming text-to-speech via ElevenLabs WebSocket API with real-time audio generation
Streaming text-to-speech via OpenAI Audio Speech API with adaptive sentence chunking