๊ฐ€์ž… ํ›„ ์ดˆ๋Œ€ ๋งํฌ๋ฅผ ๊ณต์œ ํ•˜๋ฉด ๋™์˜์ƒ ์žฌ์ƒ ๋ฐ ์ดˆ๋Œ€ ๋ณด์ƒ์„ ๋ฐ›์„ ์ˆ˜ ์žˆ์Šต๋‹ˆ๋‹ค.

Glenn Matlin
@GlennMatlin
PhD Computer Science (Intelligent Systems) ๐Ÿง ๐Ÿค– @gtcomputing ๐Ÿ๐Ÿ’ป@matsprogram Fellow ๐Ÿ”ฅ๐ŸŽ“
๊ฐ€์ž… November 2021
258 ํŒ”๋กœ์ž‰ ์ค‘    774 ํŒฌ
๐Ÿค”Where in trillions of pre-training tokens do capabilities actually come from? ๐Ÿ’ฅIn our new COLM 2026 paper, we trace the capability provenance of OLMo3-7B across Dolma3 using gradient-based training data attribution ๐Ÿงต Paper: Code: Explorer:
๋” ๋ณด๊ธฐ