Skip to content

Kimi ​

Endpoint: https://mmu.vod-qcloud.com/v1인증: VOD ApiToken / Bearer

기본 정보 ​

항목값
호출 주소https://mmu.vod-qcloud.com
인증VOD 애플리케이션에서 발급한 ApiToken / Bearer
출력텍스트 / 구조화된 응답
응답 방식JSON / SSE

지원 프로토콜 ​

규격경로대상
Chat CompletionsPOST /v1/chat/completions아래 모든 모델
MessagesPOST /v1/messages아래 모든 모델

아래 모델 ID에 해당 규격을 사용합니다. 인증과 공통 필드는 API Reference를 따릅니다.

버전별 샘플링 옵션 ​

모델 IDtemperaturetop_p
kimi-k3설정 가능설정 가능
kimi-k2.7-code미지원미지원
kimi-k2.6설정 가능설정 가능

temperature는 샘플링 온도, top_p는 누적 확률 범위를 지정합니다. 한 가지 방식을 중심으로 조정합니다.

버전별 추론 설정 ​

모델 ID설정 방식비활성화
kimi-k3thinking_enabledfalse
kimi-k2.7-codethinking_enabled비활성화 불가
kimi-k2.6thinking_enabledfalse

Chat Completions 기준입니다. 추론 내용은 reasoning_content, 최종 답변은 content에서 읽습니다. 비활성화할 수 없는 버전에는 비활성화 옵션을 전달하지 않습니다.

요청 파라미터 ​

파라미터필수타입설명
model필수String버전별 옵션 표의 모델 ID
messages필수ArrayChat Completions / Messages의 대화 입력
max_tokens선택IntegerChat Completions / Messages의 생성 한도. Messages에서는 필수
stream선택Booleantrue: SSE / false: JSON

요청 예시 ​

Chat Completions ​

json
{
  "model": "kimi-k3",
  "messages": [
    {
      "role": "user",
      "content": "Reply with exactly OK and nothing else."
    }
  ],
  "max_tokens": 2048,
  "stream": false,
  "reasoning_effort": "low"
}

응답 ​

json
{
  "id": "01a0de40-e125-7362-86bf-77c3082b129e",
  "object": "chat.completion",
  "model": "kimi-k3",
  "choices": [
    {
      "index": 0,
      "message": {
        "role": "assistant",
        "content": "OK",
        "reasoning_content": "The user wants exactly OK and nothing else. Final should be just OK. Need ensure no extra whitespace? Provide exactly OK."
      },
      "finish_reason": "stop"
    }
  ],
  "usage": {
    "prompt_tokens": 94,
    "prompt_tokens_details": {
      "audio_tokens": 0,
      "cached_tokens": 64
    },
    "prompt_cached_tokens_details": {
      "audio_tokens": 0
    },
    "completion_tokens": 39,
    "completion_tokens_details": {
      "accepted_prediction_tokens": 0,
      "audio_tokens": 0,
      "reasoning_tokens": 25,
      "rejected_prediction_tokens": 0
    },
    "total_tokens": 133
  }
}

Messages ​

json
{
  "model": "kimi-k3",
  "messages": [
    {
      "role": "user",
      "content": "Reply with exactly OK and nothing else."
    }
  ],
  "max_tokens": 512,
  "stream": false
}

응답 ​

json
{
  "id": "01a0de34-c2f3-71f6-9a0f-32e6bfd963e1",
  "model": "kimi-k3",
  "type": "message",
  "role": "assistant",
  "content": [
    {
      "text": "OK",
      "type": "text"
    }
  ],
  "stop_reason": "end_turn",
  "usage": {
    "input_tokens": 93,
    "output_tokens": 146
  }
}

특수 설정 ​

추론 텍스트와 최종 응답을 분리하고 finish_reason을 함께 확인합니다. 도구 호출 응답은 tool_calls로 처리하며 후속 요청에서 도구 ID와 대화 순서를 보존합니다. 이미지 입력은 해당 버전의 image_url Part를 사용합니다.

基于 VitePress 构建 · 部署于腾讯云 EdgeOne Pages