Moshi: a speech-text foundation model for real-time dialogue Paper • 2410.00037 • Published Sep 17, 2024 • 20
Multi-Faceted Interactivity Alignment in Full-Duplex Speech Models Paper • 2606.11167 • Published Jun 9 • 6
Multiplayer Interactive World Models with Representation Autoencoders Paper • 2607.05352 • Published Jul 6 • 30
High-Quality Image Restoration Following Human Instructions Paper • 2401.16468 • Published Jan 29, 2024 • 15
Proactive Detection of Voice Cloning with Localized Watermarking Paper • 2401.17264 • Published Jan 30, 2024 • 19
Scaling Up to Excellence: Practicing Model Scaling for Photo-Realistic Image Restoration In the Wild Paper • 2401.13627 • Published Jan 24, 2024 • 78
Orion-14B: Open-source Multilingual Large Language Models Paper • 2401.12246 • Published Jan 20, 2024 • 14