Podcast Episodes

Back to Search
Interactive Training: Feedback-Driven Neural Network Optimization

Episode 1224

🤗 Upvotes: 33 | cs.LG, cs.AI, cs.CL

Authors:
Wentao Zhang, Yang Young Lu, Yuntian Deng

Title:
…

10 months, 2 weeks ago

Short Long
View Episode
ModernVBERT: Towards Smaller Visual Document Retrievers

Episode 1223

🤗 Upvotes: 24 | cs.IR

Authors:
Paul Teiletche, Quentin Macé, Max Conti, Antonio Loison, Gautier Viaud, Pierre Co…

10 months, 2 weeks ago

Short Long
View Episode
StockBench: Can LLM Agents Trade Stocks Profitably In Real-world Markets?

Episode 1222

🤗 Upvotes: 24 | cs.LG, cs.CL

Authors:
Yanxu Chen, Zijun Yao, Yantao Liu, Jin Ye, Jianing Yu, Lei Hou, Juanzi Li

…

10 months, 2 weeks ago

Short Long
View Episode
DeepSearch: Overcome the Bottleneck of Reinforcement Learning with Verifiable Rewards via Monte Carlo Tree Search

Episode 1221

🤗 Upvotes: 100 | cs.AI, cs.CL

Authors:
Fang Wu, Weihao Xuan, Heli Qi, Ximing Lu, Aaron Tu, Li Erran Li, Yejin Ch…

10 months, 2 weeks ago

Short Long
View Episode
GEM: A Gym for Agentic LLMs

Episode 1220

🤗 Upvotes: 53 | cs.LG, cs.AI, cs.CL

Authors:
Zichen Liu, Anya Sims, Keyu Duan, Changyu Chen, Simon Yu, Xiangxin …

10 months, 2 weeks ago

Short Long
View Episode
VLA-RFT: Vision-Language-Action Reinforcement Fine-tuning with Verified Rewards in World Simulators

Episode 1219

🤗 Upvotes: 52 | cs.RO, cs.CV

Authors:
Hengtao Li, Pengxiang Ding, Runze Suo, Yihao Wang, Zirui Ge, Dongyuan Zang…

10 months, 2 weeks ago

Short Long
View Episode
Knapsack RL: Unlocking Exploration of LLMs via Optimizing Budget Allocation

Episode 1218

🤗 Upvotes: 32 | cs.LG, cs.AI, cs.CL

Authors:
Ziniu Li, Congliang Chen, Tianyun Yang, Tian Ding, Ruoyu Sun, Ge Zh…

10 months, 2 weeks ago

Short Long
View Episode
PIPer: On-Device Environment Setup via Online Reinforcement Learning

Episode 1217

🤗 Upvotes: 26 | cs.SE, cs.AI, cs.LG

Authors:
Alexander Kovrigin, Aleksandra Eliseeva, Konstantin Grotov, Egor Bo…

10 months, 2 weeks ago

Short Long
View Episode
SINQ: Sinkhorn-Normalized Quantization for Calibration-Free Low-Precision LLM Weights

Episode 1216

🤗 Upvotes: 25 | cs.LG

Authors:
Lorenz K. Müller, Philippe Bich, Jiawei Zhuang, Ahmet Çelik, Luca Benfenati, Luka…

10 months, 2 weeks ago

Short Long
View Episode
ACON: Optimizing Context Compression for Long-horizon LLM Agents

Episode 1215

🤗 Upvotes: 21 | cs.AI, cs.CL

Authors:
Minki Kang, Wei-Ning Chen, Dongge Han, Huseyin A. Inan, Lukas Wutschitz, Y…

10 months, 2 weeks ago

Short Long
View Episode

Love PodBriefly?

If you like Podbriefly.com, please consider donating to support the ongoing development.

Support Us