MikeTrendsTrends right now

search

multimodal AI

Trends

  1. 1
    Google releases EmbeddingGemma 2, an open lightweight multimodal embedding model●EmbeddingGemma 2: An open, lightweight multimodal embedding modelYhnTechnology41631 min ago

    Google has released EmbeddingGemma 2, an open, lightweight multimodal embedding model announced on the company's developer blog. The model is designed to convert text and other inputs into embeddings for search and retrieval tasks while remaining small enough to run on modest hardware. Developers are discussing the release, with attention on its open availability and what a small multimodal embedding model means for building search, RAG, and classification applications without heavy compute.

  2. 2
    Aleph Alpha Publishes Tech Report for Kolibri Model●Kolibri – Tech Report [pdf]YhnTechnology10911 h ago

    German AI company Aleph Alpha has released a technical report on Kolibri, its multimodal foundation model. The PDF, published on the company's website, is being discussed by technology readers, with many weighing how the German challenger's approach and capabilities compare to larger US-based AI labs.

  3. 3
    TwelveLabs launches Pegasus 1.6 video model for physical AI▼Pegasus 1.6 brings video understanding to physical AI, says TwelveLabs✉newsTechnologyRobotics12 h ago

    TwelveLabs has released Pegasus 1.6, a video understanding model aimed at physical AI applications such as robotics. The company says the model can analyze video input to help machines and robots interpret real-world visual environments, extending multimodal AI beyond screen-based tasks into embodied systems operating in physical spaces.

  4. 4
    UniEvo-VL Uses Self-Distillation for Multimodal Self-Improvement●UniEvo-VL: Self-Distillation Training for Multimodal Model Self-ImprovementYhn1322 h ago

    A new paper introduces UniEvo-VL, a multimodal AI model trained through self-distillation, allowing it to improve its own performance without relying on large amounts of externally labeled data. The approach is being discussed among researchers as an example of growing interest in self-improving model training methods, and the paper is available on arXiv.

  5. 5
    Mistral AI launches Mistral Large 4 public preview●Mistral AI officially launched the Mistral Large 4 public preview. Discover how this trillion-parameter multimodal modelMmastodonTechnologyAI318 h ago

    Mistral AI has officially launched the public preview of Mistral Large 4, a trillion-parameter multimodal model. The French AI company says the new system is built to rival top closed-source models from competitors, marking a significant step in Europe's push into frontier-scale artificial intelligence. Early reactions in tech communities focus on its scale and open-ecosystem implications.

  6. 6
    Google DeepMind releases open EmbeddingGemma 2 embedding model●Google DeepMind has released EmbeddingGemma 2, an open embedding model that maps text, code, images,... # ai # automatioMmastodonBusiness319 h ago

    Google DeepMind has released EmbeddingGemma 2, an open embedding model that maps text, code, images and other inputs into shared representations, allowing search and retrieval across different data types. The model is designed to run on-device rather than in the cloud, making it free and practical for local applications. Developers and tech commentators are highlighting its multimodal capabilities and its usefulness for search, coding and automation tools.