Chargement
Chargement
Skills
Configure Azure API Management as an AI Gateway for AI models, MCP tools, and agents. WHEN: semantic caching, token limit, content safety, load balancing, AI model governance, MCP rate limiting, jailbreak detection, add Azure OpenAI backend, add AI Foundry model, test AI gateway, LLM policies, configure AI backend, token metrics, AI cost control, convert API to MCP, import OpenAPI to gateway.
À propos de cette ressource
--- name: azure-aigateway description: "Configure Azure API Management as an AI Gateway for AI models, MCP tools, and agents. WHEN: semantic caching, token limit, content safety, load balancing, AI model governance, MCP rate limiting, jailbreak detection, add Azure OpenAI backend, add AI Foundry model, test AI gateway, LLM policies, configure AI backend, token metrics, AI cost control, convert API to MCP, import OpenAPI to gateway." license: MIT metadata: author: Microsoft version: "3.1.1" compatibility: Requires Azure CLI (az) for configuration and testing ---
Configure Azure API Management (APIM) as an AI Gateway for governing AI models, MCP tools, and agents.
To deploy APIM, use the azure-prepare skill. See APIM deployment guide.
| Category | Triggers | |----------|----------| | Model Governance | "semantic caching", "token limits", "load balance AI", "track token usage" | | Tool Governance | "rate limit MCP", "protect my tools", "configure my tool", "convert API to MCP" | | Agent Governance | "content safety", "jailbreak detection", "filter harmful content" | | Configuration | "add Azure OpenAI backend", "configure my model", "add AI Foundry model" | | Testing | "test AI gateway", "call OpenAI through gateway" |
---
| Policy | Purpose | Details | |--------|---------|---------| | azure-openai-token-limit | Cost control | Model Policies | | azure-openai-semantic-cache-lookup/store | 60-80% cost savings | Model Policies | | azure-openai-emit-token-metric | Observability | Model Policies | | llm-content-safety | Safety & compliance | Agent Policies | | rate-limit-by-key | MCP/tool protection | Tool Policies |
---
```bash
az apim show --name <apim-name> --resource-group <rg> --query "gatewayUrl" -o tsv
az apim backend list --service-name <apim-name> --resource-group <rg> \ --query "[].{id:name, url:url}" -o table
az apim subscription keys list \ --service-name <apim-name> --resource-group <rg> --subscription-id <sub-id> ```
---
```bash GATEWAY_URL=$(az apim show --name <apim-name> --resource-group <rg> --query "gatewayUrl" -o tsv)
curl -X POST "${GATEWAY_URL}/openai/deployments/<deployment>/chat/completions?api-version=2024-02-01" \ -H "Content-Type: application/json" \ -H "Ocp-Apim-Subscription-Key: <key>" \ -d '{"messages": [{"role": "user", "content": "Hello"}], "max_tokens": 100}' ```
---
See references/patterns.md for full steps.
```bash
az cognitiveservices account list --query "[?kind=='OpenAI']" -o table
az apim backend create --service-name <apim> --resource-group <rg> \ --backend-id openai-backend --protocol http --url "https://<aoai>.openai.azure.com/openai"
az role assignment create --assignee <apim-principal-id> \ --role "Cognitive Services User" --scope <aoai-resource-id> ```
Recommended policy order in <inbound>:
See references/policies.md for complete example.
---
| Issue | Solution | |-------|----------| | Token limit 429 | Increase tokens-per-minute or add load balancing | | No cache hits | Lower score-threshold to 0.7 | | Content false positives | Increase category thresholds (5-6) | | Backend auth 401 | Grant APIM "Cognitive Services User" role |
See references/troubleshooting.md for details.
---
Connecte-toi pour laisser un commentaire.