▲ 141 ▼ DeepSeek-V3 now runs at 20 tokens per second on Mac Studio, and that’s a nightmare for OpenAI (venturebeat.com) submitted 1 year ago by misk@sopuli.xyz to c/technology@beehaw.org 25 comments fedilink hide all child comments
[–] IndeterminateName@beehaw.org 29 points 1 year ago (1 child) A bit like a syllable when you are talking about text based responses. 20 tokens a second is faster than most people could read the output so that's sufficient for a real time feeling "chat". permalink fedilink source parent hideshow 2 child comments replies: [–] SteevyT@beehaw.org 2 points 1 year ago Huh, yeah that actually is above my reading speed assuming 1 token = 1 word. Although, I found that anything above 100 words per minute, while slow to read, feels real time to me since that's about the absolute top end of what most people type. permalink fedilink source parent
[–] SteevyT@beehaw.org 2 points 1 year ago Huh, yeah that actually is above my reading speed assuming 1 token = 1 word. Although, I found that anything above 100 words per minute, while slow to read, feels real time to me since that's about the absolute top end of what most people type. permalink fedilink source parent