FastQuery: Communication-efficient Embedding Table Query for Private LLM Inference

Open in new window