Helping Music Co-Creation Agents 'Listen' Well: Hierarchical Self-Supervised World Models for Understanding and Generation Paper • 2608.04378 • Published Aug 5 • 3
Helping Music Co-Creation Agents 'Listen' Well: Hierarchical Self-Supervised World Models for Understanding and Generation Paper • 2608.04378 • Published Aug 5 • 3
MIDI-RAE-JEPA: Hierarchical Representation Learning and Generation for Symbolic Music Paper • 2607.14537 • Published Jul 16
CharacterFlywheel: Scaling Iterative Improvement of Engaging and Steerable LLMs in Production Paper • 2603.01973 • Published Mar 2 • 7
Adaptive Nonlinear Vector Autoregression: Robust Forecasting for Noisy Chaotic Time Series Paper • 2507.08738 • Published Jul 11, 2025 • 1
EmbeddingGemma: Powerful and Lightweight Text Representations Paper • 2509.20354 • Published Sep 24, 2025 • 51
DiffusionNFT: Online Diffusion Reinforcement with Forward Process Paper • 2509.16117 • Published Sep 19, 2025 • 24
The Common Pile v0.1: An 8TB Dataset of Public Domain and Openly Licensed Text Paper • 2506.05209 • Published Jun 5, 2025 • 66
SmolVLA: A Vision-Language-Action Model for Affordable and Efficient Robotics Paper • 2506.01844 • Published Jun 2, 2025 • 168
view post Post 8326 SmolVLM is now available on PocketPal — you can run it offline on your smartphone to interpret the world around you. 🌍📱And check out this real-time camera demo by @ngxson , powered by llama.cpp:https://github.com/ngxson/smolvlm-realtime-webcamhttps://x.com/pocketpal_ai See translation 6 replies · ❤️ 14 14 😎 2 2 + Reply