Entity
Knowledge distillation
A technique used in DeepSeek's R1 series models to compress training costs while maintaining performance.
术语属性
- Used In:DeepSeek R1 series
- Reported Benefit:Training cost reduced to 1/70 while maintaining performance comparable to OpenAI's top models