Kimi
Kimi Text

devdoclang/kimi-k3

Model ID: devdoclang/kimi-k3
Family
Kimi
Context
1,000,000
Input / 1M tokens
0.0576 USD
Cache hit rate
Requests
32
Success
68.8%
Tokens
1.46M
Uptime
24h agoNow
7 days agoNow

Pricing

Input / 1M tokens
0.0576 USD
Output / 1M tokens
0.115 USD
Prompt Cache
Cache read / 1M tokens
0.0461 USD
Cache write / 1M tokens
0.0576 USD

Repeated prompts are cached — reads are far cheaper than input, applied automatically via the API.

Cost estimator
Estimated total 0 USD
Average / request 0 USD

Estimate based on list prices. Caching usually makes it cheaper.

Supported endpoints
https://api.xah.io/v1/chat/completions
https://api.xah.io/v1/messages
https://api.xah.io/v1/responses
https://api.xah.io/v1beta/models/devdoclang/kimi-k3:generateContent
https://api.xah.io/api/chat
Method POST · application/json
Auth Bearer token (API key)
Rate limit 60 req/min
Tested load 🟡 Medium peak 23/s · sustained 7.1/s
Latency p95 10,426 ms
API kinds Text

Notes

Leave a 5★ rating if you enjoy the service—it helps us keep prices low. If you hit an AWS Gateway error, just retry a few times. Due to limited supply, prices are temporarily higher. Thanks for your understanding!

Share

curl 'https://api.xah.io/v1/chat/completions' \
  -H 'Authorization: Bearer YOUR_API_KEY' \
  -H 'Content-Type: application/json' \
  -d '{
    "model": "devdoclang/kimi-k3",
    "messages": [
        {
            "role": "user",
            "content": "Xin chào, bạn là ai?"
        }
    ]
}'
// Node.js / Browser
const res = await fetch('https://api.xah.io/v1/chat/completions', {
  method: 'POST',
  headers: {
    'Authorization': 'Bearer YOUR_API_KEY',
    'Content-Type': 'application/json'
  },
  body: JSON.stringify({
    "model": "devdoclang/kimi-k3",
    "messages": [
        {
            "role": "user",
            "content": "Xin chào, bạn là ai?"
        }
    ]
})
});
const data = await res.json();
console.log(data);
import requests

url = 'https://api.xah.io/v1/chat/completions'
headers = {
    'Authorization': 'Bearer YOUR_API_KEY',
    'Content-Type': 'application/json',
}
payload = {
    "model": "devdoclang/kimi-k3",
    "messages": [
        {
            "role": "user",
            "content": "Xin chào, bạn là ai?"
        }
    ]
}

res = requests.post(url, headers=headers, json=payload, timeout=120)
print(res.status_code, res.json())
{
    "model": "devdoclang/kimi-k3",
    "messages": [
        {
            "role": "user",
            "content": "Xin chào, bạn là ai?"
        }
    ]
}

Try it

Send a prompt with your API key — uses real credits, streams back instantly.

Reviews & comments
0.0 / 0 reviews
Loading reviews...
Contact Support
We are always ready to help you