Google DeepMind launches EmbeddingGemma 2, a 740M-parameter model to map code, images, video, and audio in a shared embedding space, under an Apache 2.0 license
EmbeddingGemma 2 is the most capable model for on-device multimodal embeddings, natively mapping combinations of text, images …
Mistral launches a preview of Mistral Large 4, or “Le Chonk”, a 1T model it says tops any open model developed in the US or Europe, with weights due October 27
Mistral is launching a public preview of Mistral Large 4, a one-trillion-parameter multimodal model code-named “Le Chonk” …
Mistral launches a preview of Mistral Large 4, or Le Chonk, a 1T model it claims tops any open model developed in the US or Europe; weights are due October 27
Mistral is launching a public preview of Mistral Large 4, a one-trillion-parameter multimodal model code-named “Le Chonk” …
Google DeepMind launches EmbeddingGemma 2, a 740M-parameter model to map code, images, video, and audio in a shared embedding space, under an Apache 2.0 license
EmbeddingGemma 2 is the most capable model for on-device multimodal embeddings, natively mapping combinations of text, images …
Cloudflare debuts open-weight multimodal decision models Clef and Clef-flash, claiming they are smarter and faster than Jev, based on Qwen3.8-27B and Qwen3.5-9B
Sure, it costs more, but it can handle images and video and it's available on Hugging Face if you have the hardware horsepower to run it locally
Fei-Fei Li's World Labs unveils Atlas, a multimodal world model that generates image/video frames with pixel-perfect camera control and reconstructs them in 3D
World models generate, reconstruct, and simulate any possible world. They understand how worlds appear, behave …
Fei-Fei Li's World Labs unveils Atlas, a multimodal world model that generates image/video frames with pixel-perfect camera control and reconstructs them in 3D
World models generate, reconstruct, and simulate any possible world. They understand how worlds appear, behave …
Z.ai releases GLM-5.3-Flash, the first natively multimodal GLM-5 series model, with 320B parameters, saying it outperforms GLM-5.2 at “one-tenth the price”
We introduce GLM-5.3-Flash, the first natively multimodal model in the GLM-5 series. With 320B total parameters …
Z.ai releases GLM-5.3-Flash, the first natively multimodal GLM-5 series model, with 320B parameters, and says it served the model as Ox Alpha on Chinese chips
We introduce GLM-5.3-Flash, the first natively multimodal model in the GLM-5 series. With 320B total parameters …