親愛的 TAIWAN AI RAP 用戶您好,
感謝您持續使用 TAIWAN AI RAP API 服務。
為提升整體平台效能與服務穩定性,我們將進行部分模型調整,相關說明如下:
---
【一、下架模型清單】
以下模型將停止提供服務:
* 【模型名稱】
medgemma-27b-text-it
Magistral-Small-2506
Llama-3.1-8B-Instruct
Ministral-8B-Instruct-2410
medgemma-27b-it
Microsoft-Phi-4
TAIDE-LX-7B-Chat
Llama-3.1-TAIDE-LX-8B-Chat
Llama3-TAIDE-LX-8B-Chat-Alpha1
Llama-3.1-70B
Llama-3.1-Nemotron-70B-Instruct
Llama-3.3-70B-Instruct-Gaudi3
Llama-3.1-Nemotron-70B-Instruct-Gaudi3
Whisper-Large-V2
gpt-oss-safeguard-20b
Llama-4-Scout-17B-16E-Instruct-FP8
jina-embeddings-v3
NVIDIA-Nemotron-3-Nano-30B-A3B-FP8
【二、停止服務時間】
【2026/08/05 08:00】(UTC+8)
請留意,於停止服務後,相關 API 呼叫將無法使用。
---
【三、影響說明】
1. 已使用上述模型之 API 呼叫,將於下架後停止回應
2. 呼叫可能回傳錯誤(如:model_not_found)
3. API 金鑰中設定之該模型將不再生效
為避免影響您的服務,建議您提前完成模型切換。
---
【四、建議替代模型】
📌【醫療專用】
適用於:
medgemma-27b-text-it
medgemma-27b-it
->>以上建議替代模型: gemma-4-31B-it
📌【推理模型】
適用於:
Magistral-Small-2506
Microsoft-Phi-4
->>以上建議替代模型: gemma-4-31B-it
📌 【通用模型】
Llama-3.1-8B-Instruct
Ministral-8B-Instruct-2410
->>以上建議替代模型: gemma-4-31B-it
📌【台灣繁中模型】
TAIDE-LX-7B-Chat
Llama-3.1-TAIDE-LX-8B-Chat
Llama3-TAIDE-LX-8B-Chat-Alpha1
->>以上建議替代模型: gemma-4-31B-it
📌【Llama3 模型】
Llama-3.1-70B
Llama-3.1-Nemotron-70B-Instruct
Llama-3.3-70B-Instruct-Gaudi3
Llama-3.1-Nemotron-70B-Instruct-Gaudi3
->>以上建議替代模型: Llama-3.3-70B-Instruct
📌【STT模型】
Whisper-Large-V2
->>以上建議替代模型:Whisper-Large-V3 、Whisper-Large-V3-Turbo
📌【安全模型】
gpt-oss-safeguard-20b
->>以上建議替代模型gpt-oss-safeguard-120b
📌【Llama4 模型】
Llama-4-Scout-17B-16E-Instruct-FP8
->>以上建議替代模型Llama-4-Maverick-17B-128E-Instruct-FP8、gpt-oss-120b
📌【Embedding 模型】
jina-embeddings-v3
->>以上建議替代模型jina-embeddings-v4-vllm-text-matching、jina-embeddings-v4-vllm-code、jina-embeddings-v4-vllm-retrieval
📌【推理模型】
NVIDIA-Nemotron-3-Nano-30B-A3B-FP8
->>以上建議替代模型NVIDIA-Nemotron-3-Super-120B-A12B、gemma-4-31B-it
【五、調整方式】
請至 TAIWAN AI RAP Portal:
API入口 → API金鑰(前往列表)
1. 編輯既有 API 金鑰或新增金鑰
2. 調整可使用模型
3. 儲存設定
4. 於您的應用中切換模型名稱
---
【六、建議行動】
建議您儘早完成以下作業:
1. 確認目前是否使用即將下架模型
2. 切換至替代模型
3. 進行必要測試以確保服務正常運作
---
【七、模型管理機制說明】
為維持平台模型品質與資源使用效率,平台將定期檢視各模型之使用情形,並依實際使用數據進行資源配置調整。
本次評估係參考近期模型使用量狀況,並綜合考量維護與服務資源後進行調整。未來亦將持續依據相同原則進行模型生命週期管理,以提供更穩定且符合使用需求的服務環境。
---
若您對模型使用或遷移有任何問題,歡迎隨時與我們聯繫,我們將協助您完成調整。
感謝您的理解與支持。