AI/ML 온디바이스 2026.07.22 ⭐ 9/10
Quanta Compute
1.58-bit ternary LLM inference ASIC — BitNet 1.58을 하드웨어로, 칩당 100W로 70B 모델 50 tokens/sec
1.58-bit ternary LLM 추론 ASIC. 70B 모델 100W·50 tokens/sec. Samsung NEXT 투자.
1.58-bit ternary LLM inference ASIC — BitNet 1.58을 하드웨어로, 칩당 100W로 70B 모델 50 tokens/sec
1.58-bit ternary LLM 추론 ASIC. 70B 모델 100W·50 tokens/sec. Samsung NEXT 투자.