Google DeepMind's EmbeddingGemma 2 embeds text, code, images, video and audio in one 740M model for on-device RAG.
Mistral AI released Mistral Large 4, a 1.05T parameter open-weight multimodal MoE with 49B active parameters, 1M context.
Together AI has released Together Link, a free, MIT-licensed CLI now in beta. It connects the coding agents developers ...
Yandex SONA replaces Yandex Music's recommendation cascade with 1 generative model, lifting Active Users 4.53% in A/B tests.
Reflection AI's Beam is a 501B open-weight MoE with 23B active parameters, 1M context, and Apache 2.0 weights.
Reka Rho-1: a 19B omni-reasoning model unifying text, image, video generation and robot actions inside one shared network.
JEPA-Anything adds Orthogonal Predictive Factorization to JEPA world models, improving all 10 matched dynamics tasks across 7 ...
One transformer ran candidate generation and ranking in Yandex Music's A/B test without hand-engineered features, lifting likes 11.42%.
GPT-6 Astra, GPT-6.1 Sol, Gemini 4 Argon and Claude Fable 5.1 compared on benchmarks, pricing and best-fit jobs.
Google moves federated learning into TEEs, giving Gboard externally verifiable central differential privacy and faster server-side training.
Aleph Alpha's Kolibri is a 78.1B open-weight English-German MoE model with 3.46B active parameters and 1M token context.
Qwen's full history from Tongyi Qianwen in 2023 to Qwen3.8's 2.4T open weights, with dates, licenses and sources.