Anyone else perhaps , Knuth is extreme stickler for zero mistakes in his work, including typos or something trivial.
He didn’t even trust the typesetting system of his day and developed TeX
You think he would be able to achieve anything given his approach to his work? He would be spending even more time vetting and validating every single character a LLM generates.
Yes 20 question sample is not enough to comprehensively evaluate a LLM in general.
His objective was hardly a thorough analysis of critique of ChatGPT , he was merely blogging about an idle conversation with a friend , he literally came up with questions on a bike ride .
He clearly states this is not area of interest for him. At 85 being careful of your time and interest is a good thing ?
He didn’t even trust the typesetting system of his day and developed TeX
You think he would be able to achieve anything given his approach to his work? He would be spending even more time vetting and validating every single character a LLM generates.
Yes 20 question sample is not enough to comprehensively evaluate a LLM in general.
His objective was hardly a thorough analysis of critique of ChatGPT , he was merely blogging about an idle conversation with a friend , he literally came up with questions on a bike ride .
He clearly states this is not area of interest for him. At 85 being careful of your time and interest is a good thing ?