Arxiv
entities · 11 notes linked
Related: Deep Learning · AI Agents · Large Language Models · Retrieval Augmented Generation · Reinforcement Learning · Generalization · Memorization · Collaborative Learning
Notes
- Agent-in-the-Loop: A Data Flywheel for Continuous Improvement in LLM-based Customer Support — Live human-feedback flywheel continuously improving LLM customer support system
- Deep Learning Through the Lens of Example Difficulty — Prediction depth as per-example measure of deep learning difficulty
- Fast approximation of matrix coherence and statistical leverage — Randomized O(nd log n) algorithm for all statistical leverage scores
- How Many Images Does It Take? Estimating Imitation Thresholds in Text-to-Image Models — Empirical threshold at which text-to-image models begin imitating training concepts
- Into the Unknown Unknowns: Engaged Human Learning through Participation in Language Model Agent Conversations — Co-STORM multi-agent system for serendipitous unknown-unknown discovery
- Learning High-Dimensional Mixtures of Graphical Models — Efficient unsupervised learning of mixtures of discrete graphical models via tree approximation
- Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems — Survey of offline RL algorithms learning policies from static datasets without online interaction
- Raman spectroscopy in open world learning settings using the Objectosphere approach — Objectosphere loss reduces false positives for unknown Raman spectra classes
- The Era of Agentic Organization: Learning to Organize with Language Models — AsyncThink paradigm: concurrent LLM reasoning optimized via reinforcement learning
- Ultrafast reversible self-assembly of living tangled matter — California blackworms form tangles slowly but untangle in milliseconds via helical waves
- Visualizing the Loss Landscape of Neural Nets — Filter normalization method reveals how architecture shapes neural loss geometry