๊ฐ€์ž… ํ›„ ์ดˆ๋Œ€ ๋งํฌ๋ฅผ ๊ณต์œ ํ•˜๋ฉด ๋™์˜์ƒ ์žฌ์ƒ ๋ฐ ์ดˆ๋Œ€ ๋ณด์ƒ์„ ๋ฐ›์„ ์ˆ˜ ์žˆ์Šต๋‹ˆ๋‹ค.

Nathan Hu
@NathanHu12
PhD @stanfordnlp | aspiring LLM biologist |
๊ฐ€์ž… September 2021
367 ํŒ”๋กœ์ž‰ ์ค‘    401 ํŒฌ
New paper! In subliminal learning, LLMs transmit traits (e.g. loving cats) though seemingly unrelated data (e.g. numbers). We proactively detect these effects as readable prompts. To do so, we use the surprising ability of models to verbalize learned soft prompts.๐Ÿงต
๋” ๋ณด๊ธฐ