view article Article Welcome RL Environments to the hub +6 burtenshaw, AdithyaSK, sergiopaniego, xeophon, ryanmarten, merve, lhoestq, julien-c • 14 days ago • 32
Kandinsky 5.0: A Family of Foundation Models for Image and Video Generation Paper • 2511.14993 • Published Nov 19, 2025 • 236
Apertus: Democratizing Open and Compliant LLMs for Global Language Environments Paper • 2509.14233 • Published Sep 17, 2025 • 25
Apertus: Democratizing Open and Compliant LLMs for Global Language Environments Paper • 2509.14233 • Published Sep 17, 2025 • 25
Apertus v1 Collection Democratizing Open and Compliant LLMs for Global Language Environments: 8B and 70B open-data open-weights models, multilingual in >1000 languages • 4 items • Updated Jul 24 • 361
Benchmarking Optimizers for Large Language Model Pretraining Paper • 2509.01440 • Published Sep 1, 2025 • 25 • 1
Gradient Clipping Improves AdaGrad when the Noise Is Heavy-Tailed Paper • 2406.04443 • Published Jun 6, 2024
Benchmarking Optimizers for Large Language Model Pretraining Paper • 2509.01440 • Published Sep 1, 2025 • 25
Just a Simple Transformation is Enough for Data Protection in Vertical Federated Learning Paper • 2412.11689 • Published Dec 16, 2024 • 2
Just a Simple Transformation is Enough for Data Protection in Vertical Federated Learning Paper • 2412.11689 • Published Dec 16, 2024 • 2 • 2