LUT Tensor Core: Lookup Table Enables Efficient Low-Bit LLM Inference Acceleration

Open in new window