Consistency LLM: converting LLMs to parallel decoders accelerates inference 3.5xhttps://hao-ai-lab.github.io/blogs/cllm/ by zhisbug • 2 years ago 461 98 2 years agoHAhao-ai-lab.github.io