Etched Chip Outperforms Nvidia in Llama 70B Inference
ArchitectureEtched's specialized chip achieves >500K tokens/sec on Llama 70B with just 10W. This hardware architecture challenges GPU dominance & forces a strategic re-evaluation of AI inference infrastructure.
Read Report →