It would be luck based for pure LLMs, but now I wonder if the models that can use Python notebooks might be able to code a script to count it. Like its actually possible for an AI to get this answer consistently correct these days.
you are viewing a single comment's thread
view the rest of the comments
view the rest of the comments