DSpark: DeepSeek Made LLMs Faster Without Changing a Word
中文摘要
DeepSeek开源了DSpark投机采样框架,在不损失质量的前提下,将V4推理速度提升高达85%。
English Summary
DeepSeek released DSpark, an open-source speculative decoding framework that boosts V4 inference speed by up to 85% without compromising quality.
Original Excerpt
DSpark, DeepSeek’s open-source speculative decoding framework, makes V4 inference up to 85% faster with no quality loss. Continue reading on Towards Deep Learning »