登録して招待リンクを共有すると、動画再生報酬と紹介報酬を獲得できます。

The AI Colony
@TheAIColony
Home of everything AI — Here to help you use Al to boost your productivity — Daily insights on AI tools & tips || DM for collaborations hello@theaicolony.com
参加 November 2020
8.7K フォロー中    230.1K ファン
The model finishes training. The world keeps changing. That gap is where Boltzbit is building. Imagine an enterprise agent trained before a company changes its return policy. A static model needs the new rule added to every prompt or waits for another training cycle. Boltzbit’s Bayesian Self-learning Transformers (BAST) instead turn live data into targeted weight updates, allowing the model to incorporate the new policy after deployment. The next great model might not be the one that knew the most on launch day. It might be the one that never stopped learning. See how the architecture works:
もっと見る
We’ve just released the preview version of our latest paper: Infinite-Parameter LLMs — Generating and Adapting Weights from Live Data. In it, we demonstrate that large language models can learn up to 1,000x faster through Bayesian Self-learning transformers (BAST) than SOTA training algorithms engineered by human researchers. This breakthrough marks a critical shift in AI research, from cost-intensive, static-weight AI to energy-efficient, dynamic-weight AI. For a decade, building more advanced models has been achieved by training bigger ones on more data. This approach is now hitting a ceiling with the supply of pretraining text projected to run out in the next two to five years. Meanwhile, the fast adoption of AI agents is producing an unprecedented amount of continuously-growing data that models can learn from. None of it is captured, locked in individual sessions and lost as soon as the agent completes the task. Bringing self-learning AI agents to users captures the value of that data. We validated that BAST LLMs overcome the fundamental limitations of memory-based learning, such as Retrieval-Augmented Generation (RAG), in both cost and performance, across long-context conversations. Without underlying architectural innovation like BAST, RAG, larger context windows, and agent scaffolding suffer a drop in performance and escalating input token cost due to the static weight of LLMs. This research underscores a fundamental premise of our work at Boltzbit to achieve General Learning Intelligence and democratise model ownership. Scaling up static-weight model size or layering on more workarounds cannot address the spiraling cost of training AI systems. The viable path is a new AI architecture that adapts its own weights. Preview version of the paper:
もっと見る