Menu

#HyperNetwork

1 post

Feed
1 of 1 post
Infinite-Parameter LLMs: Generating and Adapting Weights from Live Data
🖼️
0

Infinite-Parameter LLMs: Generating and Adapting Weights from Live Data

The scaling laws hold that a language model grows more capable with more parameters and more training data, and Mixture-of-Experts (MoE) architectures have ridden these laws to remarkable results, activating only a fraction of an enormous stored parameter…

15s
Read More