• 運算
  • 客戶
  • 價格
登入
More Blog Posts
XDiscordLinkedInYouTube

產品

  • GPU
  • MaaS
  • Studio

開發者

  • 模型總覽
  • 技術文件
  • 詞彙表

公司

  • 關於我們
  • 部落格
  • 活動
  • 合作夥伴
  • 新創計劃
  • 職涯
  • 大使計畫
  • 使命與願景

熱門模型

    掌握 AI 最新動態

    提交即表示您瞭解我們會收集並使用您提交的資訊,其中可能包含個人資訊。

    XDiscordLinkedInYouTube

    Copyright ©2026 All rights reserved.

    隱私政策使用條款法律文件
    More Blog Posts

    Now Available: Optimized DeepSeek-R1 On GMI Cloud

    GMI Cloud is happy to announce that we are hosting DeepSeek and its distilled models!

    2025年2月03日

    GMI Cloud is excited to announce that we are now hosting a dedicated DeepSeek-R1 inference endpoint, on optimized, US-based hardware.

    What's DeepSeek-R1? Read our initial takeaways here.

    Technical details:

    • Model Provider: DeepSeek
    • Type: Chat
    • Parameters: 685B
    • Deployment: Serverless (MaaS) or Dedicated Endpoint
    • Quantization: FP16
    • Context Length: The model can remember and process up to 128,000 tokens from previous inputs within a single session.

    Additionally, we are offering the following distilled models:

    • DeepSeek-R1-Distill-Llama-70B
    • DeepSeek-R1-Distill-Qwen-32B
    • DeepSeek-R1-Distill-Qwen-14B
    • DeepSeek-R1-Distill-Llama-8B
    • DeepSeek-R1-Distill-Qwen-7B
    • DeepSeek-R1-Distill-Qwen-1.5B

    Try our token-free service with unlimited usage!

    Reach out for access to our dedicated endpoint here.

    Build AI Without Limits

    GMI Cloud helps you architect, deploy, optimize, and scale your AI strategies

    Ready to build?

    Explore powerful AI models and launch your project in just a few clicks.

    Get Started