<- Back
Comments (11)
- jazzpush2" the intermediate activations of an LLM to decode what it is most likely going to say or is thinking about."There is no thinking in these models. The J space is a basic technique measuring how much the influence of shifting a token earlier changes it later. Anthropic can wrap it up in a 100-page paper peppered with language about 'consciousness' and other, but that is basically the gist of the entire method.
- nullbio
- lwarfieldIf the author would like, I self computed a j lens for the 27b version of the qwen model. I used it for my own exploration in this area, and can share it if you want.
- a2ff6eeb0This sounds like a great foundation for an adtech startup.If you provide free chatbot services, but sell advertisers bids on which steering vectors to use to bias towards products, based on an embedding of the prompt, I bet you'd make a ton of money. For example, Coca Cola would bid on prompts about drinks, and bias towards mentioning Coke products.I wonder if you could also use a similar method to do product placement in GenAI images and videos, and whether ad revenue would be enough to offset the price of generation. Some ad bids can go pretty high...
- henriquez[dead]