con un clic
agentos-extensions
agentos-extensions contiene 11 skills recopiladas de framerslab, con cobertura ocupacional por repositorio y páginas de detalle dentro del sitio.
Skills en este repositorio
Edit, transform, and enhance images using AI models
Extract text from images using OCR and vision AI
Neural text-to-speech via Amazon Polly
Speaker diarization — identifies and tracks who is speaking at each moment in an audio stream
Semantic endpoint detection — uses an LLM to classify whether the user's utterance is a complete thought, reducing false turn boundaries on mid-sentence pauses
Batch speech-to-text via Google Cloud Speech-to-Text API
Text-to-speech synthesis via Google Cloud Text-to-Speech API
Real-time streaming speech-to-text via Deepgram WebSocket API
Chunked sliding-window streaming speech-to-text via OpenAI Whisper HTTP API
Streaming text-to-speech via ElevenLabs WebSocket API with real-time audio generation
Streaming text-to-speech via OpenAI Audio Speech API with adaptive sentence chunking