بنقرة واحدة
agentos-extensions
يحتوي agentos-extensions على 11 من skills المجمعة من framerslab، مع تغطية مهنية على مستوى المستودع وصفحات skill داخل الموقع.
Skills في هذا المستودع
Edit, transform, and enhance images using AI models
Extract text from images using OCR and vision AI
Neural text-to-speech via Amazon Polly
Speaker diarization — identifies and tracks who is speaking at each moment in an audio stream
Semantic endpoint detection — uses an LLM to classify whether the user's utterance is a complete thought, reducing false turn boundaries on mid-sentence pauses
Batch speech-to-text via Google Cloud Speech-to-Text API
Text-to-speech synthesis via Google Cloud Text-to-Speech API
Real-time streaming speech-to-text via Deepgram WebSocket API
Chunked sliding-window streaming speech-to-text via OpenAI Whisper HTTP API
Streaming text-to-speech via ElevenLabs WebSocket API with real-time audio generation
Streaming text-to-speech via OpenAI Audio Speech API with adaptive sentence chunking