Alphabet Inc Class CGoogle Gemini is delivered on premises through Google Distributed Cloud and operated by Cirrascale, extending Gemini's enterprise reach.
Cirrascale Cloud Servicesは、サンタクララで開催されたAI Infra Summitにおいて、エンタープライズ向けAI推論のための完全なソフトウェアスタックであるCirrascale Inference Platformの正式版リリースを発表した。このプラットフォームにより、企業はオープンソースモデル、自社のプライベートモデル、およびクローズドなモデルエコシステムを単一のサーバーレスプラットフォームから実行できる。これには、Google Distributed Cloudを通じてオンプレミスで提供され、Cirrascaleが運用するGoogle Geminiも含まれる。モデルおよびハードウェア選択レイヤーは、各リクエストを適切なモデルに自動的にルーティングし、NVIDIA、AMD、Tenstorrent、Qualcommにまたがる利用可能な最適なアクセラレータ上で実行する。切り替えにコード変更は不要である。また、チームは自社のプライベートデータでモデルをファインチューニングでき、そのデータが自社環境から出ることはない。このプラットフォームには、企業のナレッジベースに接続されたターンキーのプライベートチャット体験、チーム間でのAI支出を管理するための組み込みのコントロール、およびエージェント型ワークロード向けのガバナンスガードレールが含まれており、必要に応じてHIPAA、SOC 2、FedRAMPの要件に準拠している。Cirrascale Inference Platformは現在、Cirrascaleの米国および国際地域で利用可能である。
Alphabet Inc Class CGoogle Gemini is delivered on premises through Google Distributed Cloud and operated by Cirrascale, extending Gemini's enterprise reach.
Advanced Micro Devices IncCirrascale's inference platform routes workloads across AMD accelerators, expanding demand for AMD AI hardware.
NVIDIA CorporationThe platform runs inference on NVIDIA accelerators as one of its supported hardware options, supporting NVIDIA AI demand.
Qualcomm IncorporatedQualcomm accelerators are among the hardware options the platform can route inference workloads to.
Cirrascale launched the production release of its enterprise AI inference platform, a new product offering spanning multiple accelerators and model ecosystems.
Tenstorrent accelerators are included in the platform's hardware selection layer for running inference.