Sign in

albertge.bsky.social

@albertge.bsky.social
11 followers 14 following 10 posts
PostsRepliesMedia
albertge.bsky.social @albertge.bsky.social · 08/05/2025
Online data mixing reduces training costs for foundation models, but faces challenges: ⚠️ Human-defined domains miss semantic nuances ⚠️ Limited eval accessibility ⚠️ Poor scalability Introducing 🎵R&B: first regroup data, then dynamically reweight domains during training!
153
Reposted by @albertge.bsky.social
Fred Sala @fredsala.bsky.social · 11/12/2024
First up at #NeurIPS2024 from our group, our work on labeling via programmatic distillation (a spotlight!). Label your data orders of magnitude faster and cheaper — come join us today at Poster Session 2 East for a demo!
0158