Baidu Qianfan MCP Connector for Claude
A+Orchestrate Baidu Qianfan AI models — manage chat completions, embeddings, and prompt templates directly from any AI agent.
Connect your AI agents to Baidu Qianfan (百度千帆), the enterprise-grade LLM platform. This MCP provides 10 tools to automate interactions with Ernie Bot and other foundation models, including chat completions, vector embeddings, and prompt engineering.
What you can do
- Model Interaction — Trigger chat completions with Ernie Bot (Turbo/Speed/4.0) using persistent context
- Vector Embeddings — Generate semantic embeddings for text to power RAG and search workflows
- Prompt Engineering — Manage and retrieve centralized prompt templates for consistent model outputs
- Image Generation — Trigger Text-to-Image tasks using Baidu's advanced diffusion models
- Usage Monitoring — Track token consumption and manage model service status programmatically
How it works
- Subscribe to this server
- Log in to the Baidu Qianfan Console
- Create an application to obtain your API Key and Secret Key
- Enable the models you wish to use (e.g., Ernie-4.0-8K)
- Insert your credentials into the fields below to start managing your Baidu AI workflows.
Who is this for?
- AI Developers — automate the comparison of different Ernie model versions
- System Integrators — bridge enterprise applications with Baidu's high-performance Chinese LLMs
- Knowledge Managers — build robust RAG pipelines using Qianfan's embedding and chat services
Related Connectors
Predibase (LLM Serving & Finetuning) MCP
Deploy and query fine-tuned LLMs via Predibase — run inference, classify text, and monitor deployment metrics directly from your AI agent.
Replicate MCP
Equip your AI to dynamically search, run, and monitor thousands of open-source machine learning models hosted on Replicate via simple text commands.
APImage MCP
Search and license editorial photography from one of the world largest press image archives for media and publishing.
Groq MCP
Run large language models at unprecedented speed with custom LPU hardware that delivers real-time AI inference at massive scale.