Podcast Episodes
Back to Search
Fine-Tuning Large Language Models: A Comprehensive Guide
This podcast offers a comprehensive overview of fine-tuning large language models (LLMs), exploring both foundational principles and advanced techniq…
1Â year, 1Â month ago
Maximizing Confidence Alone Improves Reasoning
This document presents RENT, a novel method for improving the reasoning abilities of language models using unsupervised reinforcement learning. Inste…
1Â year, 2Â months ago
Critical Points of Random Neural Networks
This work examines the critical points of random neural networks, particularly as network depth increases in the infinite-width limit. The authors pr…
1Â year, 2Â months ago
BAGEL: Vision-Language Model for Visual Generation
This source introduces BAGEL, a large multimodal model designed for unified image understanding and generation. It discusses the model's Mixture-of-T…
1Â year, 2Â months ago
Incentivizing Knowledge Acquisition in LLMs via RL
This document introduces R1-Searcher++, a novel framework for Large Language Models (LLMs) designed to improve their ability to handle factual questi…
1Â year, 2Â months ago
RL for Image Generation: DPO vs GRPO
This source evaluates and compares two reinforcement learning algorithms, GRPO and DPO, for their effectiveness in generating images from text descri…
1Â year, 2Â months ago
Let Androids Dream Framework
This document presents a research paper on a novel framework, Let Androids Dream (LAD), designed to enhance AI's ability to understand the implied me…
1Â year, 2Â months ago
SmolVLM: Compact and Efficient Vision-Language Models
This source introduces SmolVLM, a collection of small-scale multimodal models designed for efficiency on devices with limited computing power. The au…
1Â year, 2Â months ago
Federated Learning: Privacy-Preserving Collaborative Intelligence Survey
This academic survey provides a comprehensive overview of Federated Learning (FL), a distributed machine learning approach allowing collaborative mod…
1Â year, 2Â months ago
Compressed Federated Learning of Tiny Language Models
This document details research into improving Federated Learning (FL) efficiency in autonomous mobile networks by incorporating tiny language models …
1Â year, 2Â months ago