返回首页
AI on Medium··行业媒体

You Can Predict Local LLM Tokens Per Second With One Division

中文摘要

使用简单的除法即可预测本地大语言模型的每秒 Token 生成速度。

English Summary

A simple division can predict the tokens per second performance of local LLMs.

原文节选

Continue reading on Medium »