<?xml version="1.0"?>
<feed xmlns="http://www.w3.org/2005/Atom" xml:lang="en">
	<id>https://tech.uvoo.io/index.php?action=history&amp;feed=atom&amp;title=Kong_volcano_sdk</id>
	<title>Kong volcano sdk - Revision history</title>
	<link rel="self" type="application/atom+xml" href="https://tech.uvoo.io/index.php?action=history&amp;feed=atom&amp;title=Kong_volcano_sdk"/>
	<link rel="alternate" type="text/html" href="https://tech.uvoo.io/index.php?title=Kong_volcano_sdk&amp;action=history"/>
	<updated>2026-08-20T19:04:45Z</updated>
	<subtitle>Revision history for this page on the wiki</subtitle>
	<generator>MediaWiki 1.35.2</generator>
	<entry>
		<id>https://tech.uvoo.io/index.php?title=Kong_volcano_sdk&amp;diff=5738&amp;oldid=prev</id>
		<title>Busk at 16:28, 17 August 2026</title>
		<link rel="alternate" type="text/html" href="https://tech.uvoo.io/index.php?title=Kong_volcano_sdk&amp;diff=5738&amp;oldid=prev"/>
		<updated>2026-08-17T16:28:36Z</updated>

		<summary type="html">&lt;p&gt;&lt;/p&gt;
&lt;table class=&quot;diff diff-contentalign-left diff-editfont-monospace&quot; data-mw=&quot;interface&quot;&gt;
				&lt;col class=&quot;diff-marker&quot; /&gt;
				&lt;col class=&quot;diff-content&quot; /&gt;
				&lt;col class=&quot;diff-marker&quot; /&gt;
				&lt;col class=&quot;diff-content&quot; /&gt;
				&lt;tr class=&quot;diff-title&quot; lang=&quot;en&quot;&gt;
				&lt;td colspan=&quot;2&quot; style=&quot;background-color: #fff; color: #202122; text-align: center;&quot;&gt;← Older revision&lt;/td&gt;
				&lt;td colspan=&quot;2&quot; style=&quot;background-color: #fff; color: #202122; text-align: center;&quot;&gt;Revision as of 16:28, 17 August 2026&lt;/td&gt;
				&lt;/tr&gt;&lt;tr&gt;&lt;td colspan=&quot;2&quot; class=&quot;diff-lineno&quot; id=&quot;mw-diff-left-l1&quot; &gt;Line 1:&lt;/td&gt;
&lt;td colspan=&quot;2&quot; class=&quot;diff-lineno&quot;&gt;Line 1:&lt;/td&gt;&lt;/tr&gt;
&lt;tr&gt;&lt;td class='diff-marker'&gt;−&lt;/td&gt;&lt;td style=&quot;color: #202122; font-size: 88%; border-style: solid; border-width: 1px 1px 1px 4px; border-radius: 0.33em; border-color: #ffe49c; vertical-align: top; white-space: pre-wrap;&quot;&gt;&lt;div&gt;What Kong AI Gateway Does &lt;del class=&quot;diffchange diffchange-inline&quot;&gt;AutomaticallySemantic Caching: If a user asks the exact same question tomorrow, Kong serves the cached response instantly without incurring costs or hitting the LLM again.Prompt Guardrails: Kong blocks malicious prompt injections or sensitive company data from leaking out to the LLM providers.Failover: If OpenAI drops offline mid-execution, Kong can automatically route GPT requests to Azure or Anthropic without breaking your Volcano code.Would you like to explore how to set up the semantic caching plugin inside Kong, or do you need help writing a custom MCP server for your specific database?&lt;/del&gt;&lt;/div&gt;&lt;/td&gt;&lt;td class='diff-marker'&gt;+&lt;/td&gt;&lt;td style=&quot;color: #202122; font-size: 88%; border-style: solid; border-width: 1px 1px 1px 4px; border-radius: 0.33em; border-color: #a3d3ff; vertical-align: top; white-space: pre-wrap;&quot;&gt;&lt;div&gt;&lt;ins class=&quot;diffchange diffchange-inline&quot;&gt;## &lt;/ins&gt;What Kong AI Gateway Does &lt;ins class=&quot;diffchange diffchange-inline&quot;&gt;Automatically&lt;/ins&gt;&lt;/div&gt;&lt;/td&gt;&lt;/tr&gt;
&lt;tr&gt;&lt;td class='diff-marker'&gt; &lt;/td&gt;&lt;td style=&quot;background-color: #f8f9fa; color: #202122; font-size: 88%; border-style: solid; border-width: 1px 1px 1px 4px; border-radius: 0.33em; border-color: #eaecf0; vertical-align: top; white-space: pre-wrap;&quot;&gt;&lt;/td&gt;&lt;td class='diff-marker'&gt; &lt;/td&gt;&lt;td style=&quot;background-color: #f8f9fa; color: #202122; font-size: 88%; border-style: solid; border-width: 1px 1px 1px 4px; border-radius: 0.33em; border-color: #eaecf0; vertical-align: top; white-space: pre-wrap;&quot;&gt;&lt;/td&gt;&lt;/tr&gt;
&lt;tr&gt;&lt;td colspan=&quot;2&quot;&gt; &lt;/td&gt;&lt;td class='diff-marker'&gt;+&lt;/td&gt;&lt;td style=&quot;color: #202122; font-size: 88%; border-style: solid; border-width: 1px 1px 1px 4px; border-radius: 0.33em; border-color: #a3d3ff; vertical-align: top; white-space: pre-wrap;&quot;&gt;&lt;div&gt;&lt;ins style=&quot;font-weight: bold; text-decoration: none;&quot;&gt;* Semantic Caching: If a user asks the exact same question tomorrow, Kong serves the cached response instantly without incurring costs or hitting the LLM again.&lt;/ins&gt;&lt;/div&gt;&lt;/td&gt;&lt;/tr&gt;
&lt;tr&gt;&lt;td colspan=&quot;2&quot;&gt; &lt;/td&gt;&lt;td class='diff-marker'&gt;+&lt;/td&gt;&lt;td style=&quot;color: #202122; font-size: 88%; border-style: solid; border-width: 1px 1px 1px 4px; border-radius: 0.33em; border-color: #a3d3ff; vertical-align: top; white-space: pre-wrap;&quot;&gt;&lt;div&gt;&lt;ins style=&quot;font-weight: bold; text-decoration: none;&quot;&gt;* Prompt Guardrails: Kong blocks malicious prompt injections or sensitive company data from leaking out to the LLM providers.&lt;/ins&gt;&lt;/div&gt;&lt;/td&gt;&lt;/tr&gt;
&lt;tr&gt;&lt;td colspan=&quot;2&quot;&gt; &lt;/td&gt;&lt;td class='diff-marker'&gt;+&lt;/td&gt;&lt;td style=&quot;color: #202122; font-size: 88%; border-style: solid; border-width: 1px 1px 1px 4px; border-radius: 0.33em; border-color: #a3d3ff; vertical-align: top; white-space: pre-wrap;&quot;&gt;&lt;div&gt;&lt;ins style=&quot;font-weight: bold; text-decoration: none;&quot;&gt;* Failover: If OpenAI drops offline mid-execution, Kong can automatically route GPT requests to Azure or Anthropic without breaking your Volcano code.&lt;/ins&gt;&lt;/div&gt;&lt;/td&gt;&lt;/tr&gt;
&lt;tr&gt;&lt;td class='diff-marker'&gt; &lt;/td&gt;&lt;td style=&quot;background-color: #f8f9fa; color: #202122; font-size: 88%; border-style: solid; border-width: 1px 1px 1px 4px; border-radius: 0.33em; border-color: #eaecf0; vertical-align: top; white-space: pre-wrap;&quot;&gt;&lt;/td&gt;&lt;td class='diff-marker'&gt; &lt;/td&gt;&lt;td style=&quot;background-color: #f8f9fa; color: #202122; font-size: 88%; border-style: solid; border-width: 1px 1px 1px 4px; border-radius: 0.33em; border-color: #eaecf0; vertical-align: top; white-space: pre-wrap;&quot;&gt;&lt;/td&gt;&lt;/tr&gt;
&lt;tr&gt;&lt;td class='diff-marker'&gt; &lt;/td&gt;&lt;td style=&quot;background-color: #f8f9fa; color: #202122; font-size: 88%; border-style: solid; border-width: 1px 1px 1px 4px; border-radius: 0.33em; border-color: #eaecf0; vertical-align: top; white-space: pre-wrap;&quot;&gt;&lt;div&gt;## Connecting Volcano to Kong AI Gateway&lt;/div&gt;&lt;/td&gt;&lt;td class='diff-marker'&gt; &lt;/td&gt;&lt;td style=&quot;background-color: #f8f9fa; color: #202122; font-size: 88%; border-style: solid; border-width: 1px 1px 1px 4px; border-radius: 0.33em; border-color: #eaecf0; vertical-align: top; white-space: pre-wrap;&quot;&gt;&lt;div&gt;## Connecting Volcano to Kong AI Gateway&lt;/div&gt;&lt;/td&gt;&lt;/tr&gt;
&lt;tr&gt;&lt;td colspan=&quot;2&quot; class=&quot;diff-lineno&quot; id=&quot;mw-diff-left-l67&quot; &gt;Line 67:&lt;/td&gt;
&lt;td colspan=&quot;2&quot; class=&quot;diff-lineno&quot;&gt;Line 70:&lt;/td&gt;&lt;/tr&gt;
&lt;tr&gt;&lt;td class='diff-marker'&gt; &lt;/td&gt;&lt;td style=&quot;background-color: #f8f9fa; color: #202122; font-size: 88%; border-style: solid; border-width: 1px 1px 1px 4px; border-radius: 0.33em; border-color: #eaecf0; vertical-align: top; white-space: pre-wrap;&quot;&gt;&lt;div&gt;* Prompt Guardrails: Kong blocks malicious prompt injections or sensitive company data from leaking out to the LLM providers.&lt;/div&gt;&lt;/td&gt;&lt;td class='diff-marker'&gt; &lt;/td&gt;&lt;td style=&quot;background-color: #f8f9fa; color: #202122; font-size: 88%; border-style: solid; border-width: 1px 1px 1px 4px; border-radius: 0.33em; border-color: #eaecf0; vertical-align: top; white-space: pre-wrap;&quot;&gt;&lt;div&gt;* Prompt Guardrails: Kong blocks malicious prompt injections or sensitive company data from leaking out to the LLM providers.&lt;/div&gt;&lt;/td&gt;&lt;/tr&gt;
&lt;tr&gt;&lt;td class='diff-marker'&gt; &lt;/td&gt;&lt;td style=&quot;background-color: #f8f9fa; color: #202122; font-size: 88%; border-style: solid; border-width: 1px 1px 1px 4px; border-radius: 0.33em; border-color: #eaecf0; vertical-align: top; white-space: pre-wrap;&quot;&gt;&lt;div&gt;* Failover: If OpenAI drops offline mid-execution, Kong can automatically route GPT requests to Azure or Anthropic without breaking your Volcano code.&lt;/div&gt;&lt;/td&gt;&lt;td class='diff-marker'&gt; &lt;/td&gt;&lt;td style=&quot;background-color: #f8f9fa; color: #202122; font-size: 88%; border-style: solid; border-width: 1px 1px 1px 4px; border-radius: 0.33em; border-color: #eaecf0; vertical-align: top; white-space: pre-wrap;&quot;&gt;&lt;div&gt;* Failover: If OpenAI drops offline mid-execution, Kong can automatically route GPT requests to Azure or Anthropic without breaking your Volcano code.&lt;/div&gt;&lt;/td&gt;&lt;/tr&gt;
&lt;tr&gt;&lt;td class='diff-marker'&gt;−&lt;/td&gt;&lt;td style=&quot;color: #202122; font-size: 88%; border-style: solid; border-width: 1px 1px 1px 4px; border-radius: 0.33em; border-color: #ffe49c; vertical-align: top; white-space: pre-wrap;&quot;&gt;&lt;div&gt;&lt;del style=&quot;font-weight: bold; text-decoration: none;&quot;&gt;&lt;/del&gt;&lt;/div&gt;&lt;/td&gt;&lt;td colspan=&quot;2&quot;&gt; &lt;/td&gt;&lt;/tr&gt;
&lt;tr&gt;&lt;td class='diff-marker'&gt;−&lt;/td&gt;&lt;td style=&quot;color: #202122; font-size: 88%; border-style: solid; border-width: 1px 1px 1px 4px; border-radius: 0.33em; border-color: #ffe49c; vertical-align: top; white-space: pre-wrap;&quot;&gt;&lt;div&gt;&lt;del style=&quot;font-weight: bold; text-decoration: none;&quot;&gt;Would you like to explore how to set up the semantic caching plugin inside Kong, or do you need help writing a custom MCP server for your specific database?&lt;/del&gt;&lt;/div&gt;&lt;/td&gt;&lt;td colspan=&quot;2&quot;&gt; &lt;/td&gt;&lt;/tr&gt;
&lt;/table&gt;</summary>
		<author><name>Busk</name></author>
	</entry>
	<entry>
		<id>https://tech.uvoo.io/index.php?title=Kong_volcano_sdk&amp;diff=5737&amp;oldid=prev</id>
		<title>Busk: Created page with &quot;What Kong AI Gateway Does AutomaticallySemantic Caching: If a user asks the exact same question tomorrow, Kong serves the cached response instantly without incurring costs or...&quot;</title>
		<link rel="alternate" type="text/html" href="https://tech.uvoo.io/index.php?title=Kong_volcano_sdk&amp;diff=5737&amp;oldid=prev"/>
		<updated>2026-08-17T16:25:01Z</updated>

		<summary type="html">&lt;p&gt;Created page with &amp;quot;What Kong AI Gateway Does AutomaticallySemantic Caching: If a user asks the exact same question tomorrow, Kong serves the cached response instantly without incurring costs or...&amp;quot;&lt;/p&gt;
&lt;p&gt;&lt;b&gt;New page&lt;/b&gt;&lt;/p&gt;&lt;div&gt;What Kong AI Gateway Does AutomaticallySemantic Caching: If a user asks the exact same question tomorrow, Kong serves the cached response instantly without incurring costs or hitting the LLM again.Prompt Guardrails: Kong blocks malicious prompt injections or sensitive company data from leaking out to the LLM providers.Failover: If OpenAI drops offline mid-execution, Kong can automatically route GPT requests to Azure or Anthropic without breaking your Volcano code.Would you like to explore how to set up the semantic caching plugin inside Kong, or do you need help writing a custom MCP server for your specific database?&lt;br /&gt;
&lt;br /&gt;
&lt;br /&gt;
## Connecting Volcano to Kong AI Gateway&lt;br /&gt;
Volcano SDK is designed to work seamlessly with the Kong AI Gateway (via Kong Advanced Gateway Services). By routing your agent's LLM calls through Kong, you get instant access to enterprise-grade security, request auditing, rate-limiting, and semantic caching.&lt;br /&gt;
Here is a conceptual example of how to connect the two and build an MCP-native workflow.&lt;br /&gt;
------------------------------&lt;br /&gt;
## Step 1: Configure the Volcano Client&lt;br /&gt;
Instead of hitting the LLM provider (like OpenAI or Anthropic) directly, you point the Volcano client to your Kong AI Gateway endpoint. Kong will manage your API keys, load-balancing, and security controls behind the scenes.&lt;br /&gt;
&lt;br /&gt;
import { Volcano, ModelProvider } from '@kong/volcano-sdk';&lt;br /&gt;
// Initialize Volcano to route through your Kong AI Gatewayconst volcano = new Volcano({&lt;br /&gt;
  baseUrl: &amp;quot;https://your-kong-gateway-domain.com&amp;quot;, &lt;br /&gt;
  apiKey: process.env.KONG_AI_GATEWAY_TOKEN, // Your Kong access token&lt;br /&gt;
});&lt;br /&gt;
&lt;br /&gt;
------------------------------&lt;br /&gt;
## Step 2: Define an MCP Tool&lt;br /&gt;
The Model Context Protocol (MCP) lets you expose external tools and data to your agent. In this example, we define an MCP tool that connects to a internal company database to fetch shipping updates.&lt;br /&gt;
&lt;br /&gt;
import { McpServer } from '@modelcontextprotocol/sdk/server';&lt;br /&gt;
const mcpServer = new McpServer({&lt;br /&gt;
  name: &amp;quot;shipping-tracker&amp;quot;,&lt;br /&gt;
  version: &amp;quot;1.0.0&amp;quot;&lt;br /&gt;
});&lt;br /&gt;
// Register a tool on the MCP server&lt;br /&gt;
mcpServer.tool(&amp;quot;get_shipping_status&amp;quot;, { orderId: z.string() }, async ({ orderId }) =&amp;gt; {&lt;br /&gt;
  // Logic to fetch real-world data&lt;br /&gt;
  const status = await fetchInternalDb(orderId); &lt;br /&gt;
  return { content: [{ type: &amp;quot;text&amp;quot;, text: `Order status: ${status}` }] };&lt;br /&gt;
});&lt;br /&gt;
&lt;br /&gt;
------------------------------&lt;br /&gt;
## Step 3: Run the Multi-Model Workflow&lt;br /&gt;
Volcano allows you to chain multiple models together. In this workflow, Claude 3.5 Sonnet acts as the high-reasoning &amp;quot;brain&amp;quot; to understand the customer's request and call the MCP tool. Then, GPT-4o mini handles the cheaper, faster task of drafting a polite email response.&lt;br /&gt;
&lt;br /&gt;
async function handleCustomerInquiry(userPrompt: string) {&lt;br /&gt;
  &lt;br /&gt;
  // 1. High-reasoning model analyzes prompt and triggers the MCP tool&lt;br /&gt;
  const agentRunner = await volcano.agents.create({&lt;br /&gt;
    model: &amp;quot;kong-managed-claude-sonnet&amp;quot;, // Configured routing inside Kong&lt;br /&gt;
    tools: [mcpServer.getTool(&amp;quot;get_shipping_status&amp;quot;)],&lt;br /&gt;
    prompt: userPrompt&lt;br /&gt;
  });&lt;br /&gt;
&lt;br /&gt;
  const toolOutput = await agentRunner.execute();&lt;br /&gt;
&lt;br /&gt;
  // 2. Chained workflow shifts to a faster, cheaper model for drafting&lt;br /&gt;
  const finalResponse = await volcano.chat.completion({&lt;br /&gt;
    model: &amp;quot;kong-managed-gpt4o-mini&amp;quot;, // Routed and cached by Kong&lt;br /&gt;
    messages: [&lt;br /&gt;
      { role: &amp;quot;system&amp;quot;, content: &amp;quot;You are a polite customer support agent.&amp;quot; },&lt;br /&gt;
      { role: &amp;quot;assistant&amp;quot;, content: `Internal tool data: ${toolOutput}` },&lt;br /&gt;
      { role: &amp;quot;user&amp;quot;, content: &amp;quot;Draft an email update to the customer based on this.&amp;quot; }&lt;br /&gt;
    ]&lt;br /&gt;
  });&lt;br /&gt;
&lt;br /&gt;
  console.log(finalResponse.text);&lt;br /&gt;
}&lt;br /&gt;
// Example usage&lt;br /&gt;
handleCustomerInquiry(&amp;quot;Where is my package for order #98765?&amp;quot;);&lt;br /&gt;
&lt;br /&gt;
------------------------------&lt;br /&gt;
## What Kong AI Gateway Does Automatically&lt;br /&gt;
&lt;br /&gt;
* Semantic Caching: If a user asks the exact same question tomorrow, Kong serves the cached response instantly without incurring costs or hitting the LLM again.&lt;br /&gt;
* Prompt Guardrails: Kong blocks malicious prompt injections or sensitive company data from leaking out to the LLM providers.&lt;br /&gt;
* Failover: If OpenAI drops offline mid-execution, Kong can automatically route GPT requests to Azure or Anthropic without breaking your Volcano code.&lt;br /&gt;
&lt;br /&gt;
Would you like to explore how to set up the semantic caching plugin inside Kong, or do you need help writing a custom MCP server for your specific database?&lt;/div&gt;</summary>
		<author><name>Busk</name></author>
	</entry>
</feed>