Skip to content

클라우드 모델

Ollama의 클라우드 모델은 강력한 GPU 없이도 실행할 수 있는 Ollama의 새로운 유형의 모델입니다. 대신 클라우드 모델은 자동으로 Ollama의 클라우드 서비스에 오프로드되면서 로컬 모델과 동일한 기능을 제공하므로, 개인용 컴퓨터에 맞지 않는 더 큰 모델을 실행하면서도 로컬 도구를 계속 사용할 수 있습니다.

지원 모델

지원 모델 목록은 Ollama의 모델 라이브러리를 참조하세요.

클라우드 모델 실행

Ollama의 클라우드 모델을 사용하려면 ollama.com 계정이 필요합니다. 로그인하거나 계정을 생성하려면 다음 명령을 실행하세요:

ollama signin

CLI

클라우드 모델을 실행하려면 터미널을 열고 다음 명령을 실행하세요:

ollama run gpt-oss:120b-cloud

Python

먼저 클라우드 모델을 pull하여 접근할 수 있도록 하세요:

ollama pull gpt-oss:120b-cloud

다음으로 Ollama의 Python 라이브러리를 설치하세요:

pip install ollama

다음으로 간단한 Python 스크립트를 생성하고 실행하세요:

python
from ollama import Client

client = Client()

messages = [
  {
    'role': 'user',
    'content': 'Why is the sky blue?',
  },
]

for part in client.chat('gpt-oss:120b-cloud', messages=messages, stream=True):
  print(part['message']['content'], end='', flush=True)

JavaScript

먼저 클라우드 모델을 pull하여 접근할 수 있도록 하세요:

ollama pull gpt-oss:120b-cloud

다음으로 Ollama의 JavaScript 라이브러리를 설치하세요:

npm i ollama

그 다음 라이브러리를 사용하여 클라우드 모델을 실행하세요:

typescript


const ollama = new Ollama();

const response = await ollama.chat({
  model: "gpt-oss:120b-cloud",
  messages: [{ role: "user", content: "Explain quantum computing" }],
  stream: true,
});

for await (const part of response) {
  process.stdout.write(part.message.content);
}

cURL

먼저 클라우드 모델을 pull하여 접근할 수 있도록 하세요:

ollama pull gpt-oss:120b-cloud

Ollama API를 통해 명령을 실행하려면 다음 cURL 명령을 실행하세요:

curl http://localhost:11434/api/chat -d '{
  "model": "gpt-oss:120b-cloud",
  "messages": [{
    "role": "user",
    "content": "Why is the sky blue?"
  }],
  "stream": false
}'

클라우드 API 접근

클라우드 모델은 ollama.com의 API를 통해 직접 접근할 수도 있습니다. 이 모드에서는 ollama.com이 원격 Ollama 호스트로 작동합니다.

인증

ollama.com의 API에 직접 접근하려면 먼저 API 키를 생성하세요.

그 다음 OLLAMA_API_KEY 환경 변수를 생성한 API 키로 설정하세요.

export OLLAMA_API_KEY=your_api_key

모델 목록 조회

Ollama API를 통해 직접 사용 가능한 모델은 다음 명령으로 목록을 조회할 수 있습니다:

curl https://ollama.com/api/tags

응답 생성

Python

먼저 Ollama의 Python 라이브러리를 설치하세요

pip install ollama

그 다음 요청을 보내세요

python

from ollama import Client

client = Client(
    host="https://ollama.com",
    headers={'Authorization': 'Bearer ' + os.environ.get('OLLAMA_API_KEY')}
)

messages = [
  {
    'role': 'user',
    'content': 'Why is the sky blue?',
  },
]

for part in client.chat('gpt-oss:120b', messages=messages, stream=True):
  print(part['message']['content'], end='', flush=True)

JavaScript

먼저 Ollama의 JavaScript 라이브러리를 설치하세요:

npm i ollama

다음으로 모델에 요청을 보내세요:

typescript


const ollama = new Ollama({
  host: "https://ollama.com",
  headers: {
    Authorization: "Bearer " + process.env.OLLAMA_API_KEY,
  },
});

const response = await ollama.chat({
  model: "gpt-oss:120b",
  messages: [{ role: "user", content: "Explain quantum computing" }],
  stream: true,
});

for await (const part of response) {
  process.stdout.write(part.message.content);
}

cURL

Ollama의 채팅 API를 통해 응답을 생성하세요:

curl https://ollama.com/api/chat \
  -H "Authorization: Bearer $OLLAMA_API_KEY" \
  -d '{
    "model": "gpt-oss:120b",
    "messages": [{
      "role": "user",
      "content": "Why is the sky blue?"
    }],
    "stream": false
  }'

로컬 전용 모드

Ollama는 Ollama 클라우드 기능 비활성화를 통해 로컬 전용 모드로 실행할 수 있습니다.

모델 지원 종료

새롭고 더 우수한 오픈소스 모델이 출시됨에 따라 Ollama는 가끔 구형 클라우드 모델을 사용 중단하고 지원을 종료합니다. Ollama 클라우드 모델에 의존하는 도구와 애플리케이션은 계속 작동하려면 업데이트가 필요할 수 있습니다. 영향을 받는 사용자는 모델 사용 중단 및 지원 종료 전에 사전 통보를 받게 됩니다. 사용 중단 공지는 이메일과 Ollama 웹사이트를 통해 전달됩니다.

Ollama 클라우드 모델의 지원 종료는 로컬 모델에 영향을 미치지 않습니다.

예정된 지원 종료 모델

지원 종료일모델권장 대안
July 15, 2026deepseek-v3.1:671bdeepseek-v4-flash
July 15, 2026deepseek-v3.2deepseek-v4-flash
July 15, 2026devstral-2:123bmistral-large-3:675b
July 15, 2026devstral-small-2:24b
July 15, 2026ministral-3:14b
July 15, 2026ministral-3:3b
July 15, 2026ministral-3:8b
July 15, 2026gemini-3-flash-previewminimax-m3
July 15, 2026gemma3:12bgemma4:31b
July 15, 2026gemma3:27bgemma4:31b
July 15, 2026gemma3:4bgemma4:31b
July 15, 2026glm-4.7glm-5.2
July 15, 2026glm-5glm-5.2
July 15, 2026minimax-m2.1minimax-m3
July 15, 2026qwen3-coder-nextqwen3.5:397b
July 15, 2026qwen3-coder:480bqwen3.5:397b

과거 지원 종료 모델

June 30, 2026
| 모델 | 권장 대안 |
| --- | --- |
| `rnj-1:8b` | |
June 16, 2026
| 모델 | 권장 대안 |
| --- | --- |
| `kimi-k2-thinking` | `kimi-k2.6` |
| `kimi-k2:1t` | `kimi-k2.6` |
| `minimax-m2`  | `minimax-m3` |
| `glm-4.6` | `glm-5.1` |
| `qwen3-next:80b` | `qwen3.5` |
| `qwen3-vl:235b` | `qwen3.5` |
| `qwen3-vl:235b-instruct` | `qwen3.5` |
| `cogito-2.1:671b` | `deepseek-v4-flash` |