If you have a strong stomach, go through this thread.
Hello. Which collapse footage does someone want me to bunk for them? Im looking for one thats not bunk.
Ostensibly it's about really bad photographic analysis. But the OP in that thread relied heavily on AI to evaluate arguments and render opinions on his critics. The first thing you see there is a fair amount of sycophancy. This is a problem with commercial AI. It wants to please the user so that the user prefers that AI over a competitor. Part of deploying a commercial AI is giving it a "personality." And some of it is just straight-up confirmation bias. It's not too hard to write prompts for LLMs that home in on a desired concept.
Haha, yes, that thread. I'd browsed through bits of pieces of it, back when it was live.
Was he using AI to compose his posts? I didn't get that, I mean I only checked out bits of it, is all. But I think I know what you're talking about here, about the citations and authority thing
But it's valid to ask how an LLM can evaluate an argument on a specialized subject that just came up and can't possibly be part of its training data. The short answer is, "badly."
Yes, badly. But, importantly, it's just a parody, is all. Is my take on it. ...Do you agree? (I want to make sure I'm not making an error myself with how I'm viewing this.) ...It isn't just "badly" critiqued, the fact is that the apparent critique is no more than parody. ...See further:
But the real question is what's going on in the model.
The model contains not only propositional knowledge on various subjects, but knowledge on the art of critical thinking itself irrespective of the propositions. So in one mode you can ask an AI to evaluate a statement like, "Was the U.S. response to 9/11 appropriate?" and you'll get a synthesis of whatever is in the training data: opinions of journalists, world leaders; the results of polls, etc. But if you give it an excerpt from an argument that you're having with someone online and ask it to evaluate the strength of the arguments on both sides, an LLM has to reach into the part of its training data that comes from such things as textbooks on rhetoric. The LLM can embed such propositions as, "These traits make an argument good (or bad)." And the transformation layer can align those embeddings with the encoding of the argument in the prompt. In this way you can expect an LLM to operate on a higher or more meta level. Its answers are not directly derived from training data.
But you can see how bad the critical analysis is. It says stuff like, "Your critics aren't citing their sources," when the critic is in fact a subject-matter expert speaking from his own knowledge and experience. Yes, citing one's sources is a characteristic of a good argument in general. But it's not the right analysis here. Nor would it be when the proposition is a self-evident mathematical expression, as also occurred. Now of course that particular thread suffers from considerable bad faith from the OP, who probably wasn't giving his AI all the information from the thread that it would need. But the AI is still getting it wrong.
So like, sure, it's following the logical-engagement, critical-thinking-engagement thing it's seen in its training: but it's simply ...well, going through the motions. That's my take. Like I said above, that's actually an important qualification, critical even, to understanding what's going on: and, like I said, I'm putting this down here for feedback to make sure I'm not wrong about this myself.
See, if in general someone trips up on that one thing, trips up by insisting on citation-from-authority even when actually referencing an authority, well then, that's kind of subtle. That's a mess-up, sure: but it's a small enough bug. It's a bug that needs fixing, in one's critical thinking repertoire, sure: but that one nuanced bug isn't enough to paint said critical thinker that makes this error as ...as someone incapable of critical thinking.
But in this case, the AI's following its training material in the meta sense like you say, to critique stuff not in its training material: but all it is doing is going through the motions. Like a stopped clock it might sometimes tick all boxes, or it might foul up, but either way, this emphatically isn't critical thinking, and should not, must not, be mistaken for such. Is my take on it, ...unless, well, it turns out I'm mistaken about this.
----------
But in any case:
That's kind of what's happening in everything, isn't it, not just critical thinking. AI's similarly just "going through the motions" when ...when producing visual images as well, isn't it? And yet, it does such a great job of it, already. Sure, it used to get the fingers all wrong, but it's stopped doing that now. ...I'm wondering, in going though the motions of critical thinking, can AI similarly end up somehow improving, getting over this glitch, maybe even this specific glitch of insisting on citations without undertanding why citations are needed, and therefore not understanding why they may not be needed in some cases?
I'm asking, first, if you agree that AI's just going though the motions of critical thinking, just parodying, not the real thing. ...But I'm also asking, in addition, if "just going through the motions" is sufficient to get it to nevertheless improve on the output to where it is difficult to tell the difference, like in many cases it already is.