we_are_coded.by CODE · The world, decoded
БГ
ElevenLabs

ElevenLabs put voice, music, image and video into one connector for your chat

ElevenLabs Blog · event date: 14 September 2026Audio

The ElevenLabs MCP connector no longer just manages agents, it now makes content as well: speech, transcription, dubbing, music, sound effects, images and video. One install, sign in with your account, no server and no API key.

In short
  • The connector works in Claude, ChatGPT, Cursor, Grokbot, Hermes and others and draws on over 50 models.
  • Everything generated lands in the ElevenCreative workspace and can be finished in Studio.
  • The admin chooses which tools the connection can call, and data residency is picked at connection time.
Checked on1 October 2026Responsible editorTsvetelin IvanovHow we workMethod · Corrections

Anyone who has made an ad from scratch knows how many tabs are open by the end. Text in one, voice in another, music in a third, editing in a fourth, and files moving between them with names like final_final_2.

ElevenLabs is chasing exactly that mess. Instead of you going to the tools, the tools come into the conversation you are already sitting in.

The facts: on 14 September 2026 ElevenLabs announced that its connector over the MCP protocol now generates content directly from the assistant: text to speech with voices from the library, transcription through Scribe with speaker labels and timestamps in 99 languages, dubbing into another language that preserves the speaker's voice, music with or without vocals, sound effects, images and video, including animating an image and lipsync. Per the company, the connector draws on over 50 models and works in Claude, ChatGPT, Cursor, Grokbot, Hermes and others. It installs from the connectors directory with an account sign-in through OAuth, with no server to run and no API keys. Everything generated lands in the ElevenCreative workspace, where it can be finished in Studio. Admins choose which tools the connection can call, and data residency is selected at connection time. It is the same connector through which ElevenLabs agents were managed until now.

The good part here is less about the generating and more about where it goes. Every clip is saved in your workspace instead of sinking into the chat history, and tomorrow you can find it, open it in the editor and fix it.

The second important part is boring: the admin decides what the connector may do. In a team where one absent-minded question could launch a dub of the entire archive on the company account, that line matters more than the list of models.

One thing is not in the announcement: the price. Generation runs through your account, and an assistant can fire off five versions of the music where you would have made one. Watch the spend in the first week.

The visual is generated code art. No third-party images.
Follow usFacebookLinkedIn
Official primary sources
→ElevenLabs Blog - Introducing voice, music, image, and video generation in the ElevenLabs MCP, 14.09.2026
Original: https://wearecoded.com/en/articles/elevenlabs-mcp-glas-muzika-video.html
ShareFacebookXLinkedInTelegramWhatsApp
← Back to all news