> ## Documentation Index
> Fetch the complete documentation index at: https://docs.turinton.ai/llms.txt
> Use this file to discover all available pages before exploring further.

# 展開アーキテクチャ

> ワンクリック展開で、あらゆるクラウドやオンプレミスのインフラで展開できます。

<img src="https://mintcdn.com/turinton-7ec6ba73/00567dlFCD-rngT5/images/Screenshot2026-01-23at3.36.57PM.png?fit=max&auto=format&n=00567dlFCD-rngT5&q=85&s=fd2821aaf93e0943ef1a1fe949d6360c" alt="Screenshot2026 01 23at3 36 57PM" width="2576" height="1314" data-path="images/Screenshot2026-01-23at3.36.57PM.png" />

# インフラ要件

| **Service Category** | **Service Type** | **Description** | **Quantity** |
| :- | :- | :- | :- |
| **Compute** | Kubernetes Service - System Node Pool (Masters) | 4 vCPUs, 16 GB RAM, Linux | 3 |
| **Compute** | Kubernetes Service User Node Pool (workers) | 16 vCPUs, 128 GB RAM, Linux | 4 - 6 |
| **Storage** | Persistent Store | Provisioned SSD (Premium), 3,500 IOPS, 150 MiB/sec throughput | 1 TB |
| **GPU (Hosted LLM)** | H100 / H200 family | 160 – 320 GB vRAM as per the concurrency requirements. 256GB RAM. 8 Core. 1 TB Storage | 3 - 4 |
| **External LLMs** | API / URL | No need of hardware | Min - 2 |

<Warning>
  注:\*\*上記のインフラは、API経由で外部LLMを使う~~500人の同時ユーザー向け(例:ChatGPT/Claude)です。ホストGPUの同時進行度は~~50〜60です

  注:\*\*インフラポリシーごとに必要な追加コンポーネント(例:ファイアウォール、ロードバランサー)

  \* 注:容量は使用負荷に応じて増加する可能性があります
</Warning>

# 責任マトリックス

| # | 任務 | Turinton | クライアント/パートナー |
| :- | :- | :- | :- |
| 1 | プラットフォーム展開 | 責任者 | |
| 2 | プラットフォームトレーニング支援 | 責任者 | |
| 3 | プラットフォームの更新 | 責任者 | |
| 4 | POCユースケースの構築 | 責任者 | 支援 |
| 5 | オントロジー/コンテキストグラフの開発 | 責任者 | 責任者 |
| 6 | ログ記録と観測可能性の設定 | 責任者 | 責任者 |
| 7 | 新しいLLMモデル展開(vLLM使用) | 責任者 | 責任者 |
| 8 | 新しいユースケースの構築と配布(POC後) | 支援 | 責任者 |
| 9 | インフラストラクチャプロビジョニング | 支援 | 責任者 |
| 10 | ハードウェア要件 | | 責任者 |
| 11 | データアクセス | | 責任者 |
| 12 | インフラを監視する | | 責任者 |
| 13 | セキュリティおよびコンプライアンス管理 | | 責任者 |
| 14 | 災害復旧/バックアップ戦略 | | 責任者 |
