Doubao Seed
Endpoint: https://mmu.vod-qcloud.com/v1인증: VOD ApiToken / Bearer
기본 정보
| 항목 | 값 |
|---|---|
| 호출 주소 | https://mmu.vod-qcloud.com |
| 인증 | VOD 애플리케이션에서 발급한 ApiToken / Bearer |
| 출력 | 텍스트 / 구조화된 응답 |
| 응답 방식 | JSON / SSE |
지원 프로토콜
| 규격 | 경로 | 대상 |
|---|---|---|
| Chat Completions | POST /v1/chat/completions | 아래 모든 모델 |
| Responses | POST /v1/responses | 아래 모든 모델 |
| Messages | POST /v1/messages | 아래 모든 모델 |
아래 모델 ID에 해당 규격을 사용합니다. 인증과 공통 필드는 API Reference를 따릅니다.
샘플링 옵션
이 문서에 나열된 모든 모델에 공통으로 적용됩니다.
| 파라미터 | 설정 |
|---|---|
temperature | 설정 가능 |
top_p | 설정 가능 |
temperature는 샘플링 온도, top_p는 누적 확률 범위를 지정합니다. 한 가지 방식을 중심으로 조정합니다.
버전별 추론 설정
| 모델 ID | 설정 방식 | 비활성화 |
|---|---|---|
st-evolving | thinking_enabled | false |
st-character | 추론 텍스트 미반환 | — |
st-2.1-turbo | thinking_enabled | false |
st-2.1-pro | thinking_enabled | false |
st-2.0-mini | thinking_enabled | false |
Chat Completions 기준입니다. 추론 내용은 reasoning_content, 최종 답변은 content에서 읽습니다. 비활성화할 수 없는 버전에는 비활성화 옵션을 전달하지 않습니다.
요청 파라미터
| 파라미터 | 필수 | 타입 | 설명 |
|---|---|---|---|
model | 필수 | String | 버전별 옵션 표의 모델 ID |
messages | 필수 | Array | Chat Completions / Messages의 대화 입력 |
input | 필수 | String / Array | Responses의 입력. messages 대신 사용 |
max_tokens | 선택 | Integer | Chat Completions / Messages의 생성 한도. Messages에서는 필수 |
max_output_tokens | 선택 | Integer | Responses의 생성 한도 |
stream | 선택 | Boolean | true: SSE / false: JSON |
요청 예시
Chat Completions
json
{
"model": "st-2.1-pro",
"messages": [
{
"role": "user",
"content": "Reply with exactly OK and nothing else."
}
],
"max_tokens": 2048,
"stream": false,
"reasoning_effort": "low"
}응답
json
{
"id": "01a0de40-f4c7-718f-b783-0ff8d844ecd3",
"object": "chat.completion",
"model": "doubao-seed-2-1-pro-260628",
"choices": [
{
"index": 0,
"message": {
"role": "assistant",
"content": "OK",
"reasoning_content": "\nGot it,Content will be written as required.\n"
},
"finish_reason": "stop"
}
],
"usage": {
"prompt_tokens": 56,
"prompt_tokens_details": {
"audio_tokens": 0,
"cached_tokens": 0
},
"prompt_cached_tokens_details": {
"audio_tokens": 0
},
"completion_tokens": 25,
"completion_tokens_details": {
"accepted_prediction_tokens": 0,
"audio_tokens": 0,
"reasoning_tokens": 24,
"rejected_prediction_tokens": 0
},
"total_tokens": 81
}
}Responses
json
{
"model": "st-evolving",
"input": "Reply with exactly OK and nothing else.",
"max_output_tokens": 512,
"stream": false
}응답
json
{
"id": "resp_021790434702899d58e3edad71566bed3cb03770f22004fdcafa8",
"object": "response",
"model": "doubao-seed-evolving",
"status": "completed",
"output": [
{
"id": "rs_02179043470320000000000000000000000ffffac15109049bfd5",
"type": "reasoning",
"summary": [
{
"type": "summary_text",
"text": "We need obey user says exactly OK and nothing else. \n\nConflicting requirements between the developer and user have been found, and the exact user-specified \"OK\" will be output without markdown."
}
],
"status": "completed",
"encrypted_content": "djFq21LcuC8TJbMv6ktgbFidTtXlsGSviuQwPZD2htygvafMpYDXfKJMtIqPif6FPKUcXFBVTI9arUt64rhOPJ8Jx6fzdwwpxqsMov5In3488e2cpkkbm4NpPU8DvZg/JjGDw0BFionPsDAc2ylVhLhv6OH6fQrhMRPL3z7PN+Ew8wfzLDKX5qSOGHDsGE5Dw2SsH2PAi6pf8hIHHPQKzWxvUkeqSpoKQOpWeb9XJMj4oE3wFQzLmbym1k+guJtAIZDDFNuT0pGobmBDnX/u6rw2uVXoq4V/fOQnmHM3weSnY9mjt8pacfRMAqEvQzNULxvdXnfODxDY9Z4xOgpz5E8t9qeZ0G4uVeR3BVp9xvMK8mZvTx07Zu/p+AGVBU1DUdVXboiMXWMWHsrb6M8+EpU5kYqEyTpAg94bpOAgR6zK40s="
},
{
"type": "message",
"role": "assistant",
"content": [
{
"type": "output_text",
"text": "OK"
}
],
"status": "completed",
"id": "msg_02179043470547100000000000000000000ffffac151090e43667"
}
],
"usage": {
"input_tokens": 54,
"output_tokens": 63,
"total_tokens": 117,
"input_tokens_details": {
"cached_tokens": 0
},
"output_tokens_details": {
"reasoning_tokens": 62
}
}
}Messages
json
{
"model": "st-evolving",
"messages": [
{
"role": "user",
"content": "Reply with exactly OK and nothing else."
}
],
"max_tokens": 512,
"stream": false
}응답
json
{
"id": "01a0de35-0701-76af-97a6-0d541d743ca5",
"model": "doubao-seed-evolving",
"type": "message",
"role": "assistant",
"content": [
{
"text": "OK",
"type": "text"
}
],
"stop_reason": "end_turn",
"usage": {
"input_tokens": 54,
"output_tokens": 271
}
}특수 설정
st-2.0-mini, st-2.1-pro, st-2.1-turbo, st-character, st-evolving은 요청의 model에 사용하는 식별자입니다. 이미지 입력은 st-2.0-mini와 st-2.1-pro의 메시지에 image_url Part로 전달합니다.
추론 활성화는 thinking_enabled, 강도는 reasoning_effort로 지정합니다. 모델별 허용 설정은 아래 표를 따릅니다. 추론 내용과 최종 답변은 reasoning_content와 content로 구분합니다.