New:Microsoft Teams Notifications Are Now Available in Socket.Learn more
Get Started

@kolbo/mcp

Package Overview
Dependencies
Maintainers
1
Versions
215
Alerts
File Explorer

Advanced tools

Socket logo

Install Socket

Detect and block malicious and high-risk dependencies

Install
Malware was recently detected in this package.

Affected versions:

1.57.1

@kolbo/mcp

Kolbo AI MCP Server - Generate images, videos, music, speech, and sound effects from Claude Code

Source
npmnpm
Version
1.87.8
Version published
Weekly downloads
3.7K
25.09%
Maintainers
1
Weekly downloads
 
Created
Source

@kolbo/mcp

Use Kolbo AI as native tools in Claude Code and Claude Desktop via MCP (Model Context Protocol).

Generate images, videos, music, speech, sound effects, multi-scene campaigns, and conversational chat — all from natural language in your coding environment. 100+ AI models behind Smart Select routing, with reusable Visual DNA profiles for character/style consistency.

✨ Interactive widgets (v1.30+): in claude.ai, Claude Desktop, and Codex Desktop, generations render as live Kolbo cards — real-time progress with model + settings chips, an inline result gallery / video or audio player, and one-click Animate · Edit · Recreate · Download actions. Multi-audio generations render every track with its own player and Download button. Library and model searches render as browsable grids with audio preview. Text-only clients (Claude Code, Codex CLI, Cursor) keep the classic text responses.

Set up — paste one prompt, or one config block (keyless, no API key)

Easiest: paste this prompt to your AI

Copy this and paste it to Claude, ChatGPT, Cursor, or any AI assistant — it installs Kolbo itself (picks local config or remote connector based on what it can do):

Connect the Kolbo AI MCP server (generate images, video, music and more).

- If you can run terminal commands (Claude Code, Cursor, Claude Desktop, or any local setup): run "npx -y @kolbo/mcp install" — it auto-configures Kolbo in the right place. If you can't run it, give me the command to run. Then I'll restart the app.
- If you're a browser chat (claude.ai, ChatGPT): add a custom connector with URL https://api.kolbo.ai/mcp under Settings → Connectors, then Connect → log in → Allow.

No API key needed — on first use a Kolbo login opens in my browser and I click Allow. When set up, confirm Kolbo is connected and offer to generate a test image of a sunset.

Or set it up yourself — one command

Run this once — it sets up the full Kolbo experience (the MCP tools and the routing skill) for every installed agent (Claude Desktop, Claude Code, Cursor), keyless:

npx -y @kolbo/mcp install

Or add the config by hand — this block is identical for every MCP client and carries no API key (on first use it logs you in via the browser):

{
  "mcpServers": {
    "kolbo": {
      "command": "npx",
      "args": ["-y", "@kolbo/mcp@latest"]
    }
  }
}
ClientWhere the config goes
Claude Code~/.claude.json (or claude mcp add kolbo -- npx -y @kolbo/mcp@latest)
Claude Desktopclaude_desktop_config.json
Cursor.cursor/mcp.json
Kolbo Codeconfigured automatically on kolbo auth login

Restart your app, then ask it to generate something. The first time, a Kolbo login opens in your browser — click Allow (no API key to create). Prefer an API key? Create one at app.kolbo.ai/developer and add "env": { "KOLBO_API_KEY": "kolbo_live_..." } to the block above.

Browser-only (claude.ai / ChatGPT / Codex): connector + skill

  • Add the custom connector https://api.kolbo.ai/mcp under Settings → Connectors, then Connect → log in → Allow.
  • Download the Skill and upload it (Claude.ai: Settings → Features → Skills; ChatGPT/Codex: Settings → Skills). Codex CLI: unzip into ~/.codex/skills/.

The Skill is the canonical routing layer. Without it the tools still work, but the model will not load Seedance / Visual DNA / filmmaking rules.

Optional upgrade: add the Kolbo skill for slash-commands + smart routing

The config above is all you need. If you want one-word slash-commands (/kolbo:marketing-studio, /kolbo:product-photoshoot, …) and automatic routing to the best tool with the right defaults, install the Kolbo skill on top — it's an enhancement layer, not a requirement:

# Claude Code plugin (canonical skill + MCP configuration)
claude plugin marketplace add Zoharvan12/kolbo-claude-plugin
claude plugin install kolbo@kolbo

# Skill only — installs the bundled official skill without changing MCP settings
npx -y @kolbo/mcp@latest skill

The skill-only command installs the canonical single kolbo skill and does not change MCP configuration. The Claude Code plugin and npx -y @kolbo/mcp install routes configure MCP as well. Official installer-created skill folders are marked as Kolbo-managed and refresh automatically when the MCP server starts on a newer package; unmarked or hand-authored folders are never overwritten. The canonical skill ships inside Kolbo Code, so however you connect, the behavior matches. See the full setup guide at docs.kolbo.ai/developer-api/claude-code-skill.

Use it

Just ask your agent naturally:

Generation

  • "Generate an image of a sunset over mountains"
  • "Create a 5-second video of waves crashing"
  • "Build a 4-scene storyboard for a coffee shop ad"
  • "Remove the background from this image"
  • "Make a lo-fi hip hop beat"
  • "Read this out loud with a British female voice"

Marketing & UGC

  • "Make me a UGC ad for my sneaker brand — 9:16, talking-head style"
  • "TV spot for my new beverage, 15 seconds, cinematic"
  • "Unboxing video for this product photo"

Brand & product imagery

  • "Pinterest pin for my candle brand, cottagecore mood"
  • "Hero banner for my landing page, wide format"
  • "Lifestyle shot of my product in a kitchen"
  • "4 ad creative variants for Meta and TikTok"

Marketplace listings

  • "Generate Amazon main image + 5 secondary images for my product"
  • "Full A+ content set for my Shopify listing"

Analysis & utility

  • "Ask Claude about the latest AI news with web search on"
  • "Analyze this video and tell me what prompts are shown on screen"
  • "What's in this image?"
  • "Create a Visual DNA profile called 'Alex' from these images"
  • "Use the same brand as last time" (loads a persisted brand kit from the workspace)

Without the optional skill, the config block alone already exposes every tool — you just describe what you want. With the skill installed, each of these is also routed to the right MCP tool with the right defaults — UGC mode picks 9:16 + sound-off + no-captions, marketplace mode enforces compliance (pure white bg, no text, no props), product photoshoot mode uses the right aspect for the platform (2:3 Pinterest, 16:9 hero banner, 1:1 IG feed), etc. The routing logic is shared with Kolbo Code, so the behavior is identical however you connect.

Available Tools

Generation

ToolDescription
generate_imageText → image. Supports preset_id from list_presets type="image".
generate_image_editExisting image(s) + prompt → edited image. Supports preset_id from list_presets type="image_edit".
generate_videoText → video
generate_video_from_imageStill image + motion prompt → video
generate_video_from_videoInput video → restyled video, or burn in subtitles (video-to-video). prompt optional — prompt-less models (VEED Subtitles, Act Two, Wan Animate) use preset / source_language / translation_language, plus srt_content / srt_file_url / vocabulary / customization for VEED
generate_elementsReference images/videos/audio + prompt → animated video
generate_first_last_frameFirst frame + last frame → interpolated video
generate_lipsyncSource image/video + audio → lipsynced video (Sync-3 adds active-speaker selection, emotion, model mode, temperature)
generate_creative_directorOne brief → N coordinated scenes (image or video)
generate_musicText (+ optional lyrics) → song. Style, title, negative tags, length, and Suno fine-controls (style weight, weirdness, audio weight, persona / singing voice)
generate_speechText + voice → spoken audio. Full expressive/style control: Google/Gemini named voice-direction presets (style_instructions_preset_id: warm/dramatic/whisper/…) or free-form style_instructions, preset styles + emotions (DeepDub / MiniMax / Cartesia), speed, accent/language, and per-provider voice settings (ElevenLabs similarity/style, DeepDub accent/variance/tempo, MiniMax pitch/volume/intensity/timbre). Status returns the same fields for reuse.
generate_soundText → sound effect. Duration, prompt influence, and per-provider controls (Stable Audio guidance, Kie loop/tempo/key, Seed-Audio voice/speed/volume/pitch + reference audio/image)
generate_3dText or reference images → 3D model (GLB/FBX/OBJ/USDZ)
transcribe_audioAudio/video URL or file → text + SRT subtitles. Language, speaker diarization, audio-event tagging, and SRT formatting (words/line, lines/subtitle, caption stretch)
separate_audio_stemsAudio/video URL (or a Kolbo generation_id) → Dialogue / Music / Effects / without-dialogue (M&E) layers. Kolbo's own masking pipeline with a speech classifier on top, so a centred engine or ambience is not handed back mislabelled as dialogue. 5 credits, runs inline
clean_dialogue_leftoversStrip voices still faintly audible in an M&E layer. 17 credits — escalation only, it trades bed fidelity to remove the leak
separate_ambiencePull room tone / atmosphere out of the Effects (or Music) bed as its own layer. 17 credits

Every image/video/creative-director tool accepts visual_dna_ids and moodboard_id for character/style consistency across outputs — you can compose create_visual_dnagenerate_image (with the DNA applied server-side) in a single agent turn. generate_creative_director also accepts moodboard_ids plural for blending.

Every generation tool also accepts an optional resolution arg. Images use "1K" (~1024px) / "2K" (Full HD) / "3K" (QHD) / "4K" (UHD); videos use vertical-pixel tiers like "720p" / "1080p" / "1440p" / "2160p". Values are model-dependent — call list_models and read the chosen model's supported_resolutions and resolutionMultipliers. Omit to use the model default.

Every generation tool also accepts an optional project_id arg that routes the generation into a specific project (owned or shared with edit+). Call list_projects to discover IDs. When omitted, generations land in the user's auto-created "API Generations" project. project_id is per-call, NOT sticky — pass it on every call once the user names a working project. Misplaced work is recoverable via move_media / move_session.

Chat & Vision

ToolDescription
chat_send_messageMulti-turn chat with any Kolbo model. Pass media_urls to analyze images, videos, or audio — auto-routes to Gemini for vision. Supports web search and deep think.
chat_list_conversationsList past chat threads
chat_get_messagesFetch messages in a conversation

Visual DNA (reusable character/style/product profiles)

ToolDescription
create_visual_dnaCreate a profile from URLs or local files
update_visual_dnaEdit name, description, stills, sheet, or type in place (never delete+recreate)
list_visual_dnasList your profiles
get_visual_dnaFetch one profile
delete_visual_dnaDelete a profile

Moodboards

ToolDescription
list_moodboardsBrowse presets + your moodboards
get_moodboardFetch one moodboard with all image URLs
create_moodboard / update_moodboard / delete_moodboardCreate / edit in place / delete

Color DNA — sticky, account-wide: the ACTIVE palette strict-grades every generation until deactivated. Opt a single generation out with skip_color_palette.

ToolDescription
list_color_palettesList your palettes (+ org)
analyze_color_paletteExtract colors from 1-5 image URLs (free, does not save)
create_color_paletteSave a palette (colors from analyze or manual); auto-activates by default
update_color_paletteRename / replace colors
delete_color_paletteDelete a palette
activate_color_paletteMake a palette the sticky active one
deactivate_color_paletteClear the active palette

Media Library

ToolDescription
media_upload_widgetOpen an in-chat upload card so claude.ai users can upload LOCAL files (image / video / audio / document) — chat attachments are unreachable from remote MCP, so this is the way to bring them in. Returns stable CDN URLs
create_upload_ticketGet a short-lived upload ticket and POST local files yourself — no upload card, no user interaction. For agents with shell access (Claude Code, Codex, Cursor, CI) talking to Kolbo over a remote connector
upload_mediaUpload a local file (path or URL), or inline source_base64 + filename, → stable Kolbo CDN URL for reuse
list_mediaBrowse media library — filter by project_id, folder_id, type, category (ai / uploaded / edited / favorites / training-lab), source_type, sort, search, pagination
list_media_foldersList the user's media folders (owned + shared) — discover folder_id values to pass to list_media
create_media_folderCreate a new folder (name, optional description / color / icon)
update_media_folderRename / recolor / re-icon a folder (owner only)
delete_media_folderSoft-delete a folder (owner only; items remain in library)
add_media_to_folderAdd up to 500 media items to a folder (idempotent)
remove_media_from_folderRemove media items from a folder
share_media_folderShare a folder by user email (owner only)
unshare_media_folderRevoke a user's access to a folder (owner only)
favorite_mediaMark a media item as favorited (idempotent) — pass media_id from list_media
unfavorite_mediaRemove a media item from favorites (idempotent) — pass media_id from list_media
get_mediaFetch one media item's full details by id
delete_mediaSoft-delete a media item (30-day trash)
restore_mediaRestore a trashed item
permanently_delete_mediaHard-delete (NOT reversible — confirm with user first)
move_mediaRe-assign a media item to a different project
bulk_delete_mediaSoft-delete up to 1000 items in one call
bulk_restore_mediaRestore up to 1000 trashed items
bulk_permanently_delete_mediaHard-delete up to 1000 (NOT reversible)
bulk_move_mediaMove up to 1000 items to a project (atomic — all-or-nothing)
move_folder_contentsMove every item in a folder to a project
get_media_statsCounts + storage bytes per type (optionally per project)

Artifacts

ToolDescription
publish_html_artifactPublish an HTML page, SVG, or Mermaid diagram and get a public shareable URL on sites.kolbo.ai. Pass share_token from a prior publish to update the same URL in place (old content kept in version history).

SYNCI Music Library (licensed production music)

ToolDescription
search_music_librarySearch the licensed catalog; results contain watermarked previews only.
analyze_script_for_musicTurn a script or scene description into a music search.
browse_music_libraryBrowse the catalog without a query.
get_music_library_facetsList genres, moods, instruments, BPM, and duration filters.
get_music_track_audioGet watermarked preview URLs for a track.
acquire_clean_music_trackSpend one SYNCI vendor credit and return clean MP3/WAV signed URLs. Idempotent with request_id.
import_music_track_to_librarySpend one vendor credit and copy a clean MP3/WAV into Kolbo's media library.
get_music_track_relatedGet stems/alternate-version metadata (purchasing remains unsupported).
get_music_track_lyricsGet lyrics metadata.

Stock Library (multi-source stock media: Pexels, Pixabay, Sketchfab 3D, Music)

ToolDescription
search_stock_mediaSearch photos/videos/illustrations/vectors/3D/music across providers. source="all" returns one interleaved feed. Find ready-made assets / b-roll (distinct from generate_image/generate_video).
get_stock_sourcesList enabled sources + which media types/filters each supports.
get_stock_categoriesList dynamic category/topic chips (pass providerParam as the category filter).
get_stock_assetGet one asset with all download variants, author, license, and attribution.
analyze_script_for_stockAI: turn a script into b-roll search terms (queries[], mediaType, keywords).
import_stock_assetCopy a stock asset into the media library (CDN copy, stable URL). Free.

Blender Bridge

ToolDescription
blender_list_sessionsList the caller's connected Blender processes before choosing a target session
blender_get_sceneQueue a bounded scene summary or full scene inspection
blender_search_docsInspect local Blender RNA and return relevant official API/manual URLs without fetching them
blender_capture_viewportCapture the active viewport to managed cache and optionally Kolbo media
blender_apply_operationsApply approved structured object, material, world, camera, light, animation, duplication, or deletion operations
blender_import_mediaImport a Kolbo media item or exact-allowlisted Kolbo-owned HTTPS media/CDN asset using smart GLB/image/video placement
blender_renderRender a still or an animation capped by Blender at 250 scene frames and 100,000,000 pixel-frames, to managed output and optionally Kolbo media
blender_undoUndo the most recent Blender change in the selected process
blender_file_operationPerform sensitive new/open/save/save-as operations with in-Blender approval
blender_execute_pythonExecute explicitly reviewed Python plus a required plain-language purpose with full host authority; approval required unless trusted mode is visibly active
blender_get_command_statusRead bounded command status/result/error plus absolute expires_at; records and idempotency claims expire after 24 hours, and awaiting_approval means stop and wait for the user

Discovery & Account

ToolDescription
list_modelsCurrent model catalog with costs and capabilities
list_voicesTTS voices (presets + cloned)
list_presetsGeneration presets across image/image-edit/video/music/text-to-video catalogs. Pass the selected exact id as preset_id; never claim a preset was applied without it.
list_cinematic_presets"Cinema mode" presets grouped by dimension (camera, lens, focal_length, aperture, angle, shot_type, color_palette, lighting) — pass ids via the cinematic arg on generate_image / generate_image_edit. Only when the user wants a specific cinematic look
list_projectsList owned + shared projects (id, name, description, role, is_default) — call first to resolve a project name into the project_id you pass to generation tools
get_projectFull project record including the unclipped description — read this before update_project
move_sessionMove ONE session (generation, chat, transcription…) and ALL its generations + media to another project
bulk_move_sessionsMove up to 100 sessions into one project in a single call — mixed types allowed, per-session failures reported
list_session_generationsA session's generations as complete groups (prompt + all its outputs) — the ids the two organize tools below take
move_generations_to_sessionMove selected generations (and only THEIR output media) into another existing session
split_sessionCarve selected generations out into a brand-new named session, atomically
undo_session_organizationReverse a move/split within 15 minutes, using the operation_id it returned
create_doc / list_docs / get_doc / update_doc / share_doc / delete_docAI Docs (Magic Pad): author project-scoped HTML documents, edit them, get public share links
generate_character_sheetGenerate a multi-angle character sheet from reference images (credits) → pass URL to create_visual_dna or update_visual_dna
list_visual_dna_folders / create_visual_dna_folder / update_visual_dna_folder / delete_visual_dna_folder / move_visual_dna_to_folderOrganize Visual DNA characters into user folders (create/rename/recolor/delete + move DNAs in/out)
create_project / update_project / archive_project / unarchive_projectProject lifecycle (create/rename/describe/archive; deletion stays in-app)
list_agents / create_agent / update_agent / delete_agentCustom chat agents (reusable named personas; description is the system instruction)
get_creative_director_statusRe-check a Creative Director batch by generation_id until all parallel scenes finish (use after a _timed_out Director run)
list_sessionsEnumerate sessions across all types, filterable by project, type, and types[]
rename_session / delete_session / restore_sessionRename a session; soft-delete leftovers after a move; restore from trash
add_project_context / list_project_context / delete_project_context / get_project_profile / regenerate_project_profileProject knowledge base (RAG): feed scripts/URLs/notes, read the synthesized living brief
list_project_assets / link_project_asset / unlink_project_asset / update_project_assetProject cast: tag Visual DNAs / moodboards onto a project, write each DNA's description and purpose note
create_moodboard / update_moodboard / delete_moodboardBuild/edit moodboards from image URLs (AI style analysis → master prompt)
clone_voice / import_elevenlabs_voice / delete_voiceCustom voices: clone from an audio sample, import by ElevenLabs ID, delete
trim_videoFrame-accurate server-side trim of a Kolbo-hosted video (async job, tool waits)
check_creditsCheck credit balance
get_generation_statusCheck one or many generations (generation_ids); wait=true blocks server-side until done — replaces client polling loops

Environment Variables

Both are optional — the local install logs in via the browser on first use.

VariableRequiredDescription
KOLBO_API_KEYNoSet a kolbo_live_ key to skip the browser login (create one at app.kolbo.ai/developer).
KOLBO_API_URLNoCustom API URL (default: https://api.kolbo.ai/api)

Chat thinking level

chat_send_message accepts optional thinking_level, using an ID from list_models with type: "text". The server validates it against the resolved model; omitted or invalid values use thinkingDefault. Existing safeguards and legacy deep_think take precedence. Discover allowed levels through thinkingLevels; no package update is required when the server changes a model capability.

Keywords

kolbo

FAQs

Package last updated on 06 Sep 2026

Related posts