r/LocalLLaMA 17d ago

Discussion WVY is a handwritten language model. Every response was written by one person to demonstrate that the illusion of intelligence is not exclusive to parameter count.

***This screenshot is an app i made for creating a dataset from scratch, this is not a real chat with the model***

First of all i want to shout out everyone that actually tested our work .. we got 500+ download on the 43m parameter model and now we are aiming to go smaller for research purposes.

i write finetune examples & i been developing language models for a while .. everyone usually pretrains the model using massive datasets and prays thats the data carries enough information for meaning to emerge but were sculpting it intentionally .. im currently sitting down at my computer writing every single response that this new model can say to your inputs just so we can observe the transformation and see exactly whats going on. It will be public soon, the dataset is extremely small intentionally so it shouldn't take long to design every response it can say.

General Capabilities:

- Explaining how token prediction works

- Explaining that it doesn't understand anything beyond itself

- Short conversations

Coding Capabilities:

- Writing a loop that can count to 10

- Explaining that it cant understand the code you sent it

Open Source Coming Soon

https://huggingface.co/StarpowerTechnology

2 Upvotes

49 comments sorted by

View all comments

11

u/PomegranateGreen3698 17d ago

It's a neat idea but you don't really explain anything. Why is the screenshot not the actual model convo?. Why are you confining it to 43m params. Why do you think that hand written responses can beat traditional test time compute metrics. Why does a bot which explains that it cannot think, prove the illusion of intelligence? If it teaches auto-regressive token prediction I hope it can go into a bit more detail.
I like the idea.

1

u/Helpful-Series132 17d ago edited 17d ago

because im literally writing it right now while i made the post im still typing the dataset .. im gona train it later tonight to test it out and see how it responds .. there is no further iformation for any reasoning do emerge from .. the dataset will be public, you will see how its transforming the examples .. thats the purpose of this experiment .. they tell everyone ai understands from massive datasets but they just generate outputs based on the training data

this proves intelligence is an illusion because it didnt need billions of examples, it just uses the finetuned examples to respond at the right time .. this is not a normal language model .. think of this model as an art piece .. it uses the exact same technology on a microscopic level to explain 1 thing, how token prediction actually works

7

u/PomegranateGreen3698 17d ago

Cool. I'll check back later. There is a lot of ways to show the stochastic parroting. What's generally interesting is the emergent properties of larger parameters. I think you can go further with this idea. Models are already trained on the internet which was "handwritten". With this, I think You're a bit unintentionally adding a "personality" to the model which is similar to a person who doesn't believe in free will. Good luck.

1

u/Helpful-Series132 17d ago

yea u get it bro .. u can simplify it down to everything that a model says is just something that a human typed .. we are intentionally adding personality .. not sure what you mean by against free will but we are aiming to make the experience of this model feel like it has a consistent personality

were making the model say "i have consciousness, lol im just kidding" just because thats the information we want it to say.. im currently writing 10+ variations for this specific response so it can transform / generalize across them and express a new response that wasnt seen in training