ハードウェア・推論研究 🇺🇸 26.07.2026 20:03

Multiverse Computing: CompactifAI nearly doubles Llama 3.3 performance on Intel Xeon 6

Multiverse ComputingMultiverse Computing MetaMeta Intel CorporationIntel Corporation
Multiverse Computing announced that its CompactifAI technology can nearly double the performance of the Llama 3.3 model when running on Intel Xeon 6 processors. The solution uses neural network compression techniques to accelerate inference without significant loss of accuracy.
Multiverse Computingは、ニューラルネットワーク圧縮技術「CompactifAI」により、Intel Xeon 6サーバープロセッサ上で大規模言語モデル「Llama 3.3」の性能をほぼ2倍に向上できると発表した。同社によると、CompactifAIは量子インスパイアおよびテンソル圧縮手法を適用し、高い精度を維持しながらモデルサイズを削減し、推論を高速化する。試験では、288コアのIntel Xeon 6プラットフォーム上でタスク実行時間が約50%短縮され、データセンターでの展開においてモデルの効率性が向上することを示した。IntelとMultiverse Computingは、このソリューションが特別なハードウェアを必要とせず、標準的な中央演算処理装置で動作することを強調している。
出典: Meta AI (GNews) — 原文
関連記事 ↓
新着ニュース