Alphabet Inc Class CGoogle Gemini is delivered on premises through Google Distributed Cloud and operated by Cirrascale, extending Gemini's enterprise reach.
Cirrascale Cloud Services ประกาศเปิดตัว Cirrascale Inference Platform อย่างเป็นทางการในงาน AI Infra Summit ที่เมืองซานตาคลารา โดยเป็นชุดซอฟต์แวร์ครบวงจรสำหรับงาน AI inference ระดับองค์กร แพลตฟอร์มนี้ช่วยให้องค์กรสามารถรันโมเดลโอเพนซอร์ส โมเดลส่วนตัวของตนเอง และระบบนิเวศโมเดลแบบปิดได้จากแพลตฟอร์มเซิร์ฟเวอร์เลสเดียว รวมถึง Google Gemini ที่ให้บริการภายในองค์กรผ่าน Google Distributed Cloud และดำเนินการโดย Cirrascale ชั้นการเลือกโมเดลและฮาร์ดแวร์ของแพลตฟอร์มจะกำหนดเส้นทางแต่ละคำขอไปยังโมเดลที่เหมาะสมโดยอัตโนมัติ และรันบนตัวเร่งความเร็วที่ดีที่สุดที่มีอยู่จาก NVIDIA, AMD, Tenstorrent และ Qualcomm โดยไม่ต้องแก้ไขโค้ดเพื่อสลับใช้งาน และทีมงานสามารถปรับแต่งโมเดลด้วยข้อมูลส่วนตัวของตนเองได้โดยที่ข้อมูลนั้นไม่ต้องออกจากสภาพแวดล้อมขององค์กร แพลตฟอร์มนี้ยังมีประสบการณ์แชทส่วนตัวแบบพร้อมใช้งานที่เชื่อมต่อกับฐานความรู้ของบริษัท มีเครื่องมือควบคุมในตัวสำหรับจัดการค่าใช้จ่ายด้าน AI across ทีมต่างๆ และมีมาตรการกำกับดูแลสำหรับงานเอเจนต์ ซึ่งสอดคล้องกับข้อกำหนด HIPAA, SOC 2 และ FedRAMP ในกรณีที่จำเป็น ปัจจุบัน Cirrascale Inference Platform พร้อมให้บริการแล้วในภูมิภาคสหรัฐอเมริกาและภูมิภาคต่างประเทศของ Cirrascale
Alphabet Inc Class CGoogle Gemini is delivered on premises through Google Distributed Cloud and operated by Cirrascale, extending Gemini's enterprise reach.
Advanced Micro Devices IncCirrascale's inference platform routes workloads across AMD accelerators, expanding demand for AMD AI hardware.
NVIDIA CorporationThe platform runs inference on NVIDIA accelerators as one of its supported hardware options, supporting NVIDIA AI demand.
Qualcomm IncorporatedQualcomm accelerators are among the hardware options the platform can route inference workloads to.
CTO Realty Growth IncCirrascale launched the production release of its enterprise AI inference platform, a new product offering spanning multiple accelerators and model ecosystems.
Tenstorrent accelerators are included in the platform's hardware selection layer for running inference.