SyncAI.news, a Varaisys broadcasting
Why Models Are AI’s Next Training Dataset with Damian Borth - #772
SC

Sam Charrington

· 47 Minutes

PodcastThe TWIML AI Podcast

Why Models Are AI’s Next Training Dataset with Damian Borth - #772

Listen

For more than a decade, AI has advanced by training ever-larger models on ever-larger datasets. But as high-quality training data becomes harder to find and pretraining grows increasingly expensive, researchers are looking for new ways to keep foundation models improving.

In this episode, Damian Borth, professor of AI and machine learning at the University of St. Gallen, argues we’ve been overlooking an important source of knowledge: the models we’ve already trained. His group’s work on weight space learning treats trained neural networks themselves as data, learning from the distilled results of millions of GPU hours of optimization rather than starting from raw data each time.

We explore what it means to build foundation models of neural networks, how knowledge can be transferred across architectures and domains, why this approach could dramatically reduce the cost of developing specialized models, and whether future AI systems may be trained on collections of existing models instead of ever-growing datasets.

🗒️  Full show notes: https://twimlai.com/go/772.

Original source

This story was published by The TWIML AI Podcast and written by Sam Charrington. SyncAI.news shows a preview; the complete article is on the publisher's site.

Read the full story on twimlai.com

Similar News