Outperforming larger language models with less training data and smaller modelshttps://blog.research.google/2023/09/distilling-step-by-step-outperforming.html by atg_abhishek • 3 years ago 320 123 3 years agoBLblog.research.google