https://github.com/Open-Dev-Society/OpenStock
0 viewsОткрыть в Telegram →
Из этого канала
- #5063MemSearcher: Training LLMs to Reason, Search and Manage Memory via End-to-End…
MemSearcher: Training LLMs to Reason, Search and Manage Memory via End-to-End Reinforcement Learning https://arxiv.org/abs/2511.02805
- #5065https://generalistai.com/blog/nov-04-2025-GEN-0
https://generalistai.com/blog/nov-04-2025-GEN-0
- #5068How to Scale Your Model A Systems View of LLMs on TPUs…
How to Scale Your Model A Systems View of LLMs on TPUs https://jax-ml.github.io/scaling-book/
- #5053Language Models are Injective and Hence Invertible…
Language Models are Injective and Hence Invertible https://www.arxiv.org/abs/2510.15511 https://www.alphaxiv.org/overview/2510.15511v3
- #5050Defeating the Training-Inference Mismatch via FP16…
Defeating the Training-Inference Mismatch via FP16 https://arxiv.org/abs/2510.26788 https://www.alphaxiv.org/ru/overview/2510.26788v1…