원클릭으로
agentos-extensions
agentos-extensions에는 framerslab에서 수집한 skills 11개가 있으며, 저장소 수준 직업 범위와 사이트 내 skill 상세 페이지를 제공합니다.
이 저장소의 skills
Edit, transform, and enhance images using AI models
Extract text from images using OCR and vision AI
Neural text-to-speech via Amazon Polly
Speaker diarization — identifies and tracks who is speaking at each moment in an audio stream
Semantic endpoint detection — uses an LLM to classify whether the user's utterance is a complete thought, reducing false turn boundaries on mid-sentence pauses
Batch speech-to-text via Google Cloud Speech-to-Text API
Text-to-speech synthesis via Google Cloud Text-to-Speech API
Real-time streaming speech-to-text via Deepgram WebSocket API
Chunked sliding-window streaming speech-to-text via OpenAI Whisper HTTP API
Streaming text-to-speech via ElevenLabs WebSocket API with real-time audio generation
Streaming text-to-speech via OpenAI Audio Speech API with adaptive sentence chunking