Sources
AD Alexander Doria@Dorialexander
New competitive European model (30bA3b) from Soofi with SYNTH inside. https://twitter.com/effi288/status/2075904321707798699
- likes: 186
- replies: 5
- bookmarks: 92
- reposts: 15
LA Lisan al Gaib@scaling01
germans released a model that's actually not terrible it's small, and still worse than Qwen3.5, but it's very comparable to Nemotron 3 Nano but I guess that's what 27T pre-training tokens will do to a model https://twitter.com/effi288/status/2075904321707798699
- likes: 751
- replies: 23
- bookmarks: 302
- reposts: 41
AD Alexander Doria@Dorialexander
For anyone wondering why they did not post-train this is actual sovereignty in the good sense of the world: slowly building strategic autonomy up to the pretraining level, while having the humility to start from the best practices in the open.…
- likes: 148
- replies: 3
- bookmarks: 33
- reposts: 10
T( Teortaxes▶️ (DeepSeek 推特🐋铁粉 2023 – ∞)@teortaxesTex
> 🇩🇪 Trained end-to-end on Deutsche Telekom’s Industrial AI Cloud in Munich — sovereign AI infrastructure on German soil Unexpected German W! https://twitter.com/effi288/status/2075904321707798699
- likes: 314
- replies: 8
- bookmarks: 74
- reposts: 21
EL elie@eliebakouch
new "sovereign" base model by a german consortium with a fun approach they take the exact same architecture (and most hyperparameters) as nemotron 3 nano and do a slightly different mixture of datasets, trained for ~26T tokens. the overlap with nemotron's mixture is roughly 80%…
- likes: 308
- replies: 11
- bookmarks: 120
- reposts: 17
Combined views
200.7K
7 Sources, first seen ago