search
multimodal AI
Trends
- 1Google releases EmbeddingGemma 2, an open lightweight multimodal embedding model●EmbeddingGemma 2: An open, lightweight multimodal embedding model
Google has released EmbeddingGemma 2, an open, lightweight multimodal embedding model announced on the company's developer blog. The model is designed to convert text and other inputs into embeddings for search and retrieval tasks while remaining small enough to run on modest hardware. Developers are discussing the release, with attention on its open availability and what a small multimodal embedding model means for building search, RAG, and classification applications without heavy compute.
- 2
German AI company Aleph Alpha has released a technical report on Kolibri, its multimodal foundation model. The PDF, published on the company's website, is being discussed by technology readers, with many weighing how the German challenger's approach and capabilities compare to larger US-based AI labs.
- 3TwelveLabs launches Pegasus 1.6 video model for physical AI▼Pegasus 1.6 brings video understanding to physical AI, says TwelveLabs
TwelveLabs has released Pegasus 1.6, a video understanding model aimed at physical AI applications such as robotics. The company says the model can analyze video input to help machines and robots interpret real-world visual environments, extending multimodal AI beyond screen-based tasks into embodied systems operating in physical spaces.
- 4UniEvo-VL Uses Self-Distillation for Multimodal Self-Improvement●UniEvo-VL: Self-Distillation Training for Multimodal Model Self-Improvement
A new paper introduces UniEvo-VL, a multimodal AI model trained through self-distillation, allowing it to improve its own performance without relying on large amounts of externally labeled data. The approach is being discussed among researchers as an example of growing interest in self-improving model training methods, and the paper is available on arXiv.
- 5Mistral AI launches Mistral Large 4 public preview●Mistral AI officially launched the Mistral Large 4 public preview. Discover how this trillion-parameter multimodal model
Mistral AI has officially launched the public preview of Mistral Large 4, a trillion-parameter multimodal model. The French AI company says the new system is built to rival top closed-source models from competitors, marking a significant step in Europe's push into frontier-scale artificial intelligence. Early reactions in tech communities focus on its scale and open-ecosystem implications.
- 6Google DeepMind releases open EmbeddingGemma 2 embedding model●Google DeepMind has released EmbeddingGemma 2, an open embedding model that maps text, code, images,... # ai # automatio
Google DeepMind has released EmbeddingGemma 2, an open embedding model that maps text, code, images and other inputs into shared representations, allowing search and retrieval across different data types. The model is designed to run on-device rather than in the cloud, making it free and practical for local applications. Developers and tech commentators are highlighting its multimodal capabilities and its usefulness for search, coding and automation tools.