A surprisingly effective way to predict token importance in LLM prompts
A surprisingly effective way to predict token importance in LLM prompts
We explored a novel method to gauge the significance of tokens in prompts given to large language models, without needing direct model access. Essentially, we just did an ablation study on the prompt using cosine similarity of the embeddings as the measure. We got surprisingly promising results when comparing this really simple approach to integrated gradients. Curious to hear thoughts from the community!
Share cardActual performance
Launch Intel predictions
Analyze your own launch →Correct prediction on native model
Similar products
Recursive LLM Prompts
Wordllama – Things you can do with the token embeddings of an LLM
Your motivational prompts. Your way.
Beeminder for Effective Charities
The SC4-HSM is now a FIDO U2F token
Jwt.show – show the payload of a jwt token
CryptoTokens.wtf – Token Whitepapers from Highest Grossing ICOs
Token Thermodynamics
Tokensift, an open-sourced token-efficiency linter for LLM prompts
Predict Then Propagate, ICLR 2019 (PyTorch)