nowfound

Alternatives

Products that do what Gemma 3 does

Build with multimodal AI from Google

  1. 1
    Gemma 3n199

    Run powerful multimodal AI right on your phone

    2025

  2. 2

    Run multimodal AI locally with an encoder-free architecture

    Jun 2026 · blog.google

  3. 3
    CM3leon101

    Efficient generative AI model for text and images from Meta

    2023

  4. 4
    Kimi K2.5205

    Native multimodal model with self-directed agent swarms

    Jan 2026

  5. 5

    Google's first natively multimodal embedding model

    Mar 2026

  6. 6CH

    Hey HN, Henry & Roman here from Cactus. A small, on-device model is fast and private, but sometimes wrong, but frontier models are getting expensive pretty fast. So, we post-trained Gemma 4 E2B post-trained to know when it's wrong. Every response comes with a confidence score between 0 and 1. Developers can accept the on-device when it's high, hand off to a bigger cloud model when it's low. By routing only 15-35% of queries to Gemini 3.1 Flash-Lite, Gemma-4-E2B matches Gemini 3.1 Flash-Lite on most benchmarks. - ChartQA: 15-20% - LibriSpeech: 25-30% - MMBench, GigaSpeech, MMAU: 30-35% -…

    Jul 2026 · github.com

  7. 7

    Multilingual, Multimodal AI from Cohere

    2025

  8. 8
    Wan 2.6151

    The next era of multimodal AI for creators is here

    Dec 2025

  9. 9
    Fuyu-8B114

    A multimodal architecture for AI agents

    2023

  10. 10

    Fine-tuned Gemma 2: 2B model for Kazakh Instructions (SLLM)

    2025

  11. 11

    A million-pixel beach for AI agents — claim & animate pixels

    Feb 2026

  12. 12
    Ferret193

    Refer and ground anything anywhere at any granularity

    2024

  13. 13
    PaLM 2247

    Google's next generation large language model

    2023

  14. 14
    TxGemma134

    AI models for faster drug development

    2025

  15. 15

    Open-weight 15B multimodal model for thinking and GUI agents

    Mar 2026

  16. 16
    GPT-5.6340

    A new standard for intelligence and efficiency

    Jul 2026 · openai.com

  17. 17

    Unified video generation for motion design and branding

    Jul 2026 · minimax.io

  18. 18

    Text, images, SST, TTS, vision & code: unleash the AI power

    2023

  19. 19
    GLM-4.6V239

    Open-source multimodal model with native tool use

    Dec 2025

  20. 20

    Automate common AI tasks for multimodal data

    2025

  21. 21

    Advanced Visual Reasoning & Agentic Tool Use

    2025

  22. 22
    Amica100

    Open Source 3D Personal AI with Emotion, Voice and Vision

    2024

  23. 23

    Multimodal reasoning model built for agentic tasks

    Jul 2026 · ai.meta.com

  24. 24

    The next generation of the Phi family from Microsoft

    2025

Ranked by how close each launch is in meaning, then by votes. Refine with a description →