Access StepFun text, vision, image, and speech models through MCP.
Copy the install command and let the AI configure it · recommended for beginners
No copy-paste install info for "StepFunMCP" yet — see the docs or source repo.
Using StepFunMCP, call text, vision, and speech models in sequence: first summarize this product description, then analyze this UI screenshot, and finally convert the summary into Chinese speech.
Returns a text summary, screenshot analysis, and playable or downloadable speech output.
Use StepFunMCP's image generation model to create 3 marketing visuals for a minimalist smart watch, and provide a social-media-ready headline for each.
Outputs multiple generated images with matching headlines, ready for marketing asset creation.
Design a customer support assistant prototype with StepFunMCP: accept user text, analyze uploaded images when needed, generate a reply, and specify which model type is used at each step.
Provides a clear multi-step workflow showing how text and vision models are used.
Connect text-only models to vision APIs for image understanding and analysis.
Convert images into text descriptions so text-only LLMs can answer visual queries.
Give text-only LLMs vision support for analyzing local or online images.
Connect to the mcp API via MCP to extend AI tool capabilities.
Use natural language to run MCP-powered browser and text workflows.
Chat with AI to retrieve documents and trigger MCP-powered tools.