๊ฐ€์ž… ํ›„ ์ดˆ๋Œ€ ๋งํฌ๋ฅผ ๊ณต์œ ํ•˜๋ฉด ๋™์˜์ƒ ์žฌ์ƒ ๋ฐ ์ดˆ๋Œ€ ๋ณด์ƒ์„ ๋ฐ›์„ ์ˆ˜ ์žˆ์Šต๋‹ˆ๋‹ค.

Sophia Tang
@_sophia_tang_
CS + Stats @Penn M&T / Research in AI4Science and Generative Modeling / Writing at
๊ฐ€์ž… October 2021
133 ํŒ”๋กœ์ž‰ ์ค‘    1.8K ํŒฌ
Can we train a one-step discrete generator without a teacher model? Introducing Discrete Beckmann Transport Models ๐Ÿ›ธ โ€”ย a new family of discrete generative models that provably carries any point in the latent space to a fixed point on the vertices of the simplex in a single step. We achieve the diversity-coherence Pareto frontier for unconditional language modeling and SOTA few-step reasoning performance, with 84.6% accuracy in 4 NFEs on Sudoku-Hard and 16.8% in 32 NFEs on TinyGSM/GSM8K. Joint work with @ShiyiWangML from my visit at Harvard! Also, huge thanks to @msalbergo and @BrianLee794309 for the helpful discussions and guidance :D More in ๐Ÿงต
๋” ๋ณด๊ธฐ