So my GPU is about 300 watts and a still blatantly stupid LLM can write a little faster than me. Take off 100w to bring that down to my own writing speed then make it 10x slower to turn that 200 watts into 20 watts. Even with that heavy bias in the LLMs favour (forgiving it the entire power cost of my PCs other components that it partially utilizes) what we get is something slow, dumb, and incapable of learning because any local model is statically weighted.
you are viewing a single comment's thread
view the rest of the comments
view the rest of the comments
replies: