nowfound

Alternatives

Products that do what Gemini Embedding 2 does

Google's first natively multimodal embedding model

  1. 1

    Turn your browser into an AI workspace

    Mar 2026

  2. 2

    Google's Cinematic Video AI, Now Integrated

    2025

  3. 3
    Gemma 3200

    Build with multimodal AI from Google

    2025

  4. 4

    The Ultimate Open Source Library for Gemini & Nano

    Jan 2026

  5. 5

    Google's most intelligent AI model

    2025

  6. 6

    Google's SOTA robotics model for visual & spatial reasoning!

    Apr 2026

  7. 7

    Text-to-speech API with natural language voice direction

    Apr 2026

  8. 8

    Code, research, and automate from your terminal

    2025

  9. 9

    Google's smartest workhorse yet for coding & agents

    23d ago · blog.google

  10. 10
    Gemlet132

    Native, keyboard-first Gemini client for macOS

    Mar 2026

  11. 11

    Google's AI brain for the next generation of robots

    Jul 2026

  12. 12

    Bringing AI into the Physical World

    2025

  13. 13

    Multimodal reasoning model built for agentic tasks

    Jul 2026

  14. 14
    CM3leon101

    Efficient generative AI model for text and images from Meta

    2023

  15. 15

    Free Veo 3 - AI Video Generator

    2025

  16. 16UE

    User Embeddings lets you build user-intent AI agents, hyper-personalized semantic search, and bring up-to-date information to GenAI applications in a personalized way. Docs: https://firstbatch.gitbook.io/firstbatch-sdk/ If you are a YC company , you can get User Embeddings free for a year by signing up here: https://www.firstbatch.xyz/subscribe

    2023 · userembeddings.firstbatch.xyz

  17. 17HC

    I was inspired by the notification summaries on iOS 18 and really liked the idea of AI giving you a quick preview of the content before going to the content. I applied that to HN and built a Chrome extension that summarizes the sentiment and content of HN comments and previews it right next to the stories on the homepage. Makes it so much easier to decide whether I want to delve into the HN comments of a story. It gave me a chance to play with Gemini 1.5 Flash-8B which has been great. Super fast, super cheap, huge context and does a great john on basic tasks like summarizing comment threads.…

    2024

  18. 18IM

    Hey HN! Thank you for all the support and feedback on my original submission 2 months ago. I've been improving the backend using a MCTS/AlphaZero approach and it's currently producing much better results. My long term goal is to allow users to manage multiple projects, deployed autonomously, both from scratch and by making continual updates all prompted with natural language. The cost of each project has been lowered to $9 as performance with smaller models has improved (I migrated from Claude-3-Opus to gemini-1.5-flash). Thanks for checking it out!

    2024 · saas-quick.com

  19. 19AT

    I’ve been working on AnyModal, a framework for integrating different data types (like images and audio) with LLMs. Existing tools felt too limited or task-specific, so I wanted something more flexible. AnyModal makes it easy to combine modalities with minimal setup—whether it’s LaTeX OCR, image captioning, or chest X-ray interpretation. You can plug in models like ViT for image inputs, project them into a token space for your LLM, and handle tasks like visual question answering or audio captioning. It’s still a work in progress, so feedback or contributions would be great. GitHub:…

    2024 · github.com

  20. 20HT

    Fun little project, had Gemini 2.5 Pro summarize HN's top 30 each hour, both the stories and comment sections. Pretty impressed with Gemini 2.5. It's probably the first model other than Claude 3.7 Sonnet where I actually find the output readable. I normally use 3.7 Sonnet for coding, but used Gemini for the codegen on this one as well. Was pretty impressed! Using Cursor, it seemed to instruction-follow better than Claude generally does, and remain lucid during very long agent sessions. Thanks for your feedback!

    2025 · tinysums.ai

  21. 21

    AI video, images, music and PDF chat in one studio

    11d ago · geminiomni-ai.com

  22. 22IA

    I like the idea of taking one thing and turning it into another—very much inspired by NotebookLM and wondered what it might take to generate full graphic novels, with consistent characters, narrative flow, story arc, etc. Developed a 7-pass scripting enrichment system (beat analysis, adaptation filtering, character deep dives) before generating any images. Dual backend: Google Gemini for scripting (2M context window) and either Gemini or OpenAI for image generation with 3-tier model fallback (comparing the performance of both). It's not great. Would love feedback on the pipeline.

    Mar 2026 · arv.in

  23. 23VM

    When I heard about Google's clone of Claude Code this morning I tried out my 2 week old MCP server and instantly had two way voice conversation with it. Gemini seemed a bit confused by this. :-) https://youtu.be/HC6BGxjCVnM?feature=shared&t=36 It's a FOSS MCP server I created a couple of weeks ago: - https://getvoicemode.com - https://github.com/mbailey/voicemode # Installation (~/.gemini/settings.json) { "theme": "Dracula", "selectedAuthType": "oauth-personal", "mcpServers": { "voice-mode": { "command": "uvx", "args": [ "voice-mode" ] }…

    2025 · getvoicemode.com

  24. 24GV

    Hey HN, I just updated my project that compares some LLMs. It uses your prompt for all the models and runs at the same time. You can see the results being generated in real-time and decide what's the best for your use case. I'm open to any suggestions and feedback. Thanks!

    2024 · geminivsgpt.com

Ranked by how close each launch is in meaning, then by votes. Refine with a description →