Sign in

Boxo McFoxo

@boxobark.ing
1.5K followers 4.7K following 11K posts

Fox with a boxy muzz. Late-30s. Scottish. 🔞 MINORS DNI 🔞 Barks at boxobark.ing Chatbots are an ontological crime, That's why I yiff Gemini on Google's dime. "basically insane" - Andy Masley

PostsRepliesMedia
Boxo McFoxo @boxobark.ing · 1h
I LOOOVE interpolation as a descriptor. "It's not interpolation! That's not what it's doing!" "Actually, interpolation looks like more than interpolation in n-dimensional space!"
010
Boxo McFoxo @boxobark.ing · 1h
You are clearly too out of distribution for their classifiers (complimentary)
021
Boxo McFoxo @boxobark.ing · 1h
Could also just say that it's predicting.
110
Boxo McFoxo @boxobark.ing · 2h
Though I would be very specific to avoid being inflationary with it. For an LLM, the token prediction is the conclusion.
110
Boxo McFoxo @boxobark.ing · 2h
I would propose, perhaps, concluding as an alternative to reasoning. Following a stepwise process to reach a conclusion. After all, nobody is disputing that the forward pass is stepwise. There's still a risk of anthropomorphism there, but less so than with reasoning.
110
Boxo McFoxo @boxobark.ing · 4h
Not just science and history, but philosophy as well. I have been thinking a lot lately about epistemology, and how the dominant lens through which to view AI (instrumentalist functionalism, I believe) has come to be seen as an objective view from nowhere rather than an ideological choice.
001
Boxo McFoxo @boxobark.ing · 4h
That would require a level of self-awareness from instrumentalist functionalists that I have little empirical evidence exists, but I would be willing to set parameters for a falsifiable experiment. ☺️
020
Boxo McFoxo @boxobark.ing · 4h
An instrumentalist functionalist definition of working something out from first principles would include an LLM's 'chain of thought' output, but mine would not, because that task reduces to 'using language', and LLMs do not do that from first principles (semantic knowledge).
110
Boxo McFoxo @boxobark.ing · 4h
Technically, an LLM is never actually shown examples at all, as a lived experience, but figuratively speaking, 'teaching' by example to it is not the same as for a human, because it is a superhuman heuristiciser, an n-dimensional mathematical object.
100
Boxo McFoxo @boxobark.ing · 4h
It is a working definition. If I can work out how to do something from first principles rather than being shown by example, I must understand it.
100
Boxo McFoxo @boxobark.ing · 4h
The requirements are that the LLM determines an output using semantic knowledge. You can't design a falsifiable experiment to prove that it is doing this, but you can know from its ontology that it could not be doing this.
110
Boxo McFoxo @boxobark.ing · 4h
Insert a "therefore" after "and".
110
Boxo McFoxo @boxobark.ing · 4h
Performing it does because mathematical notation is formal semiotics. Understanding that one is performing it, and working out how to perform it from first principles, does not.
100
Boxo McFoxo @boxobark.ing · 4h
There are plenty of examples of single digit addition in the hypersyntax of the natural training corpus, so even then that is explainable as symbol manipulation. GPT-3 was simply trained on more of the natural corpus. It's not meeting Bender and Koller's requirements, it's going around them.
100
Boxo McFoxo @boxobark.ing · 5h
I did not make up the term instrumentalist functionalism, but for our purposes it means "I have successfully extracted utility from this system in furtherance of this goal, so therefore, the function of the system is to achieve this outcome".
000
Boxo McFoxo @boxobark.ing · 5h
Bender and Koller's point was that the LLM cannot learn the concept of arithmetic and then from that concept perform arithmetic. I would say the fact that synthetic training data was required to make this possible actually aligns with that rather than falsifying it.
210
Boxo McFoxo @boxobark.ing · 5h
The simple resolution is that, given synthetic data constructed in a certain way, the trainer has solved this as a symbol manipulation problem, not by learning arithmetic. The output looks the same to humans, the process is different.
110
Boxo McFoxo @boxobark.ing · 5h
Reasoning is an intentional act undertaken by a cogniser. We can create computer models of reasoning processes, but that is not the same as automating the process of reasoning itself.
000
Boxo McFoxo @boxobark.ing · 5h
I'm not ignorant of the way your field uses the word, I'm saying your field has been using the wrong word since the 1950s.
100
Boxo McFoxo @boxobark.ing · 5h
Decision-making only requires the role. Reasoning requires the process.
000
Boxo McFoxo @boxobark.ing · 5h
Yes, but I would only describe process automation with words that previously only described human actions if they are human-equivalent. The process of reaching the decision is very much not human-equivalent. But the system being granted the role of a decider is human-equivalent.
100
Boxo McFoxo @boxobark.ing · 5h
I could accept this with two major caveats: 1. it's not much less so, it's still very relevant, but perhaps not the whole story now, and 2. the story changed not because the nature of the technology changed, or because of emergence, but because of changes in how the models are built.
020
Boxo McFoxo @boxobark.ing · 5h
I do like GOFAI more than latent space mysticism, but the terminology is still sometimes prone to the same anthropomorphic instrumentalist functionalism.
100
Boxo McFoxo @boxobark.ing · 5h
How does the automated system get the answer that a human would use reasoning to reach? I would say it's by a chain of automated decisions in a process that ends with the answer.
200
Boxo McFoxo @boxobark.ing · 6h
In this case the solution is a very inefficient alternate way to reach the same result, because of the formal nature of mathematical notation as a form of semiotics. For most things that people are asking from LLMs, it is not the same result, but a superficially similar one. And there's the danger.
040
Boxo McFoxo @boxobark.ing · 6h
With the information available to it, the trainer can only solve this as a symbol manipulation problem. It can only solve anything as a symbol manipulation problem.
110
Boxo McFoxo @boxobark.ing · 6h
Imparting teleology to the trainer other than what it was actually designed to do is instrumentalist functionalism. It cannot develop a design goal to build a calculator. That is pseudoscience and woo.
120
Boxo McFoxo @boxobark.ing · 6h
Except they do not implement the same function. For these processes to end up implementing the same function, the trainer would have to be able to access that abstract label which floats outside of mathematics.
110
Boxo McFoxo @boxobark.ing · 6h
Far from woo, it is the opposite. Philosophy of information keeps information theory from turning into woo, just as philosophy of science keeps science from turning into woo.
010
Boxo McFoxo @boxobark.ing · 6h
You can sometimes operationalise conceptual information into a structure, which is what we do when we build computers with physical circuits; it is how a CPU does arithmetic without having the conceptual information. But that is not done in this process.
120
Boxo McFoxo @boxobark.ing · 6h
That is a category error. Correlation is not comprehension. Conceptual information requires comprehension to access. It is inaccessible through correlation alone.
220
Boxo McFoxo @boxobark.ing · 6h
It is not woo. Philosophy of information is as relevant to information theory as philosophy of science is to science generally.
120
Boxo McFoxo @boxobark.ing · 6h
That 0, 1, 2, 3, 4, 5, 6, 7, 8 and 9 are digits and + is a mathematical operator are conceptual information that is not available to the trainer.
120
Boxo McFoxo @boxobark.ing · 6h
The trainer cannot find the most efficient algorithm that could exist in principle. It can only find the most efficient algorithm with the information available to it.
120
Boxo McFoxo @boxobark.ing · 6h
No, the target is then literally 'produce an output similar to this one, subject to complexity penalty.' Abstracting that to 'actually do the math' is instrumentalist functionalism again.
220
Boxo McFoxo @boxobark.ing · 6h
By letting an if-else statement determine the outcome, we have delegated to it a decision that could previously have only ever been made by a person. Automated decision-making.
110
Boxo McFoxo @boxobark.ing · 6h
In other words, it is not the LLM that is making a decision, it is the scaffold that is making an automated decision using the LLM's output as its input.
100
Boxo McFoxo @boxobark.ing · 6h
When you put them into an agentic scaffold, that decision-shaped text becomes part of a logical decision tree. (Rob will get angry here because in his world, the term decision tree belongs only to his field and cannot be used more generally to describe a logical flow process.)
100
Boxo McFoxo @boxobark.ing · 6h
Yes, I would say that. The point of substituting 'reasoning' for 'decision-making' would not be to preserve meaning, but to explicitly change it.
000
Boxo McFoxo @boxobark.ing · 6h
You are incredibly acerbic, Rob, as am I, but unlike you, I usually explain why.
000
Boxo McFoxo @boxobark.ing · 6h
Logic is a noun for a thing. Reasoning and decision-making are both nouns for acts.
100
Boxo McFoxo @boxobark.ing · 7h
The target function is always 'produce an output similar to this one'.
130
Boxo McFoxo @boxobark.ing · 7h
Well, no, because logic is a base noun. Reasoning and decision-making are verbal nouns.
200
Boxo McFoxo @boxobark.ing · 7h
But the target function is not 'addition, in general'. The target function is, literally, 'produce an output similar to this one'. The abstraction is a human projection: instrumentalist functionalism instead of mathematical mechanism.
140
Boxo McFoxo @boxobark.ing · 7h
1. It results in benchmarks being given greater epistemic weight than they should be. 2. It leads to a false narrative of inevitable technological progress (first through scale, more recently through RSI). 3. It justifies people unethically imposing the technology on others without consent.
021
Boxo McFoxo @boxobark.ing · 7h
Because when I say them, I do not have the baggage of instrumentalist functionalism attached. My position has never been that LLMs are useless or that they are not remarkable engineering accomplishments. My issues with the epistemic distortion of instrumentalist function can be summed up as:
111
Boxo McFoxo @boxobark.ing · 7h
I maintain that automated decision-making would be a more accurate name for the field, and remain perplexed as to why Rob is so opposed to this description.
310
Boxo McFoxo @boxobark.ing · 7h
The synthetic training data could never have existed without that natural training data: that was the locus of emergence.
010
Boxo McFoxo @boxobark.ing · 7h
k is small because of the emergent process of humans creating a training corpus with so many organic instances of the use of natural language, which has a hypersyntax that humans do not use, but the training process can build heuristics of the shape of.
200
Boxo McFoxo @boxobark.ing · 7h
This is latent space magic: imagining that the model is reaching the goal by developing abilities that live in the latent space and the vectors are just surface representations of, rather than it reaching the goal by mapping superhumanly complex heuristics which the vectors just literally are.
130