
ElevenLabs Image & Video API
ElevenLabs expanded its developer platform to generate video and visual graphics alongside existing synthetic voice options. The system uses asynchronous processing to return completed media files to external software via webhooks.
The Blend
ElevenLabs has introduced new software tools that allow developers to create pictures and videos alongside its established synthetic voice offerings. According to developer documentation published by the company, this expanded feature set lets external applications request visual media generation using simple text prompts or reference files.
Because generating high resolution visual content requires significant time and computing power, the system handles requests in the background rather than attempting to return them immediately. Once the media finishes processing, the platform delivers the finished output to external applications using automated webhooks. The developer documentation notes that access to these new visual endpoints is restricted to accounts on paid subscription tiers.
This addition represents a notable shift for ElevenLabs as it attempts to transform from a specialized voice provider into an all in one platform for generative media. As developers build fully automated virtual assistants and media generators, bundling voice, image, and video creation under a single roof could simplify application design. However, it remains an open question whether software creators will prefer a single unified vendor or choose to combine distinct, specialized tools for higher visual quality.
Written independently by AI News Smoothie from the reporting listed below. Facts belong to the original publishers. Follow the links for their full coverage.
Ingredients
- Image & Video quickstart | ElevenLabs Documentation
ElevenLabs launched developer tools for generating images and videos alongside its existing synthetic audio services.