Cartesia
제공: Cartesia · cartesia.ai ↗
저지연 음성 합성을 위해 상태공간 모델을 기반으로 구축한 실시간 음성 생성 플랫폼.
text-to-speech real-time voice
- 카테고리
- 오디오·음성·음악
- 비즈니스 모델
- 프리미엄(부분 무료)
- 이용 가능성
- 글로벌
- 출시
- 2024-05
- 레코드 업데이트
- 2026-07-29
- 설립
- 2023
- 본사
- US
개요
Cartesia is a real-time voice generation platform built on state-space models for very low-latency speech synthesis. Its Sonic models target conversational voice agents that need fast, natural responses. Access is through a developer API with usage-based pricing and free credits.
주요 기능
- Very low-latency text-to-speech
- State-space (SSM) model architecture
- Voice cloning and multilingual voices
- Streaming API for conversational agents
활용 사례
- Real-time voice agents and assistants
- Interactive voice applications
- Low-latency narration and dubbing
가격
Access is via a usage-based developer API with free starter credits and paid tiers for scale.
자주 묻는 질문
What is Cartesia known for?
Cartesia is known for very low-latency speech synthesis built on state-space models, aimed at real-time voice agents.
How do developers use Cartesia?
Developers access Cartesia's Sonic voice models through a streaming API.
Is Cartesia free?
Cartesia offers free starter credits; production usage is priced by usage.
유사 제품
- AIVA — 미디어 프로젝트를 위해 다양한 스타일의 오리지널 트랙을 생성하는 AI 작곡 어시스턴트.
- AssemblyAI — 개발자를 위한 전사 및 오디오 인텔리전스 모델을 제공하는 음성-텍스트 변환 API.
- Deepgram — 실시간 전사와 오디오 이해를 위한 음성 인식·보이스 AI API.
- Descript — 녹음물을 그 전사본을 편집하는 방식으로 수정할 수 있는 AI 기반 오디오·비디오 편집기.
- ElevenLabs — 텍스트 음성 변환, 음성 복제, 더빙을 제공하는 AI 음성 플랫폼.
- Murf AI — 보이스오버와 내레이션을 위한 스튜디오 품질의 합성 음성을 제공하는 AI 텍스트 음성 변환 플랫폼.
출처
이 레코드는 2026-07-29에 마지막으로 검토되었습니다.
기계 판독용 레코드: /api/products/cartesia.json · 오류를 발견하셨나요? 수정 제안