nowfound

Alternatives

Products that do what evaligo does

Evaluate, compare, and deploy AI prompts in seconds.

  1. 1

    Build, test & deploy AI prompts across 1600+ models at scale

    2025

  2. 2

    AI power right inside your workflow & ready-to-use prompts

    2023

  3. 3PO

    Hey HN! We’re Kevin and Steve. We’re building PromptTools (https://github.com/hegelai/prompttools): open-source, self-hostable tools for experimenting with, testing, and evaluating LLMs, vector databases, and prompts. Evaluating prompts, LLMs, and vector databases is a painful, time-consuming but necessary part of the product engineering process. Our tools allow engineers to do this in a lot less time. By “evaluating” we mean checking the quality of a model's response for a given use case, which is a combination of testing and benchmarking. As examples: - For generated…

    2023 · github.com

  4. 4

    Your personal prompt engineer across all AI platforms

    2025

  5. 5

    Instantly test and compare AI prompts results across models

    2025

  6. 6

    Advanced AI Prompt Manager

    Nov 2025 · getsnippets.ai

  7. 7
    Promptly304

    No-code platform for generative AI apps & chatbots

    2023

  8. 8

    AI that builds you a deterministic evaluation in minutes

    2025

  9. 9EA

    Hi friends, We are building EVA, an AI-Relational database system with first-class support for deep learning models. Our goal with EVA is to create a platform that supports AI-powered multi-modal database applications operating on structured (tables, feature vectors, etc.) and unstructured data (videos, podcasts, pdf, etc.) with deep learning models. EVA comes with a wide range of models for analyzing unstructured data, including models for object detection, OCR, text summarization, audio speech recognition, and more. The key feature of EVA is its AI-centric query optimizer. This optimizer…

    2023 · github.com

  10. 10
    Selene 1196

    Evaluate your AI app with the most accurate LLM Judge

    2025

  11. 11BE

    Hey HN, We're excited to introduce Braintrust, a platform for running and tracking AI evaluations (“evals”) [1]. At my previous startup Impira and leading AI at Figma, we had this recurring problem where we never knew if changes we made to our products would improve or regress key user scenarios. We built some tooling to solve this problem and after talking to other developers learned that it was a widespread issue. Specifically, it’s challenging to establish a great dev loop that lets you systematically improve and ship high quality AI products. We worked with the teams at Zapier, Coda, and…

    2023

  12. 12

    Talent match, Super Fast

    Sep 2025

  13. 13

    Build production-grade AI agents using natural language

    Oct 2025

  14. 14AS
  15. 15

    Evaluate code against custom AI-driven assessments

    Sep 2025

  16. 16
    EVALZZ9

    AI Powered Hiring Platform

    Dec 2025 · evalzz.com

  17. 17

    Multi-model AI testing, evaluation, optimization made simple

    Oct 2025

  18. 18

    Compare AI models side-by-side on same prompt

    Feb 2026 · testaimodels.com

  19. 19CL

    Hi HN! Run it: OPENROUTER_API_KEY="sk" npx bff-eval --demo We built a tool to help people take LLM outputs and easily grade them / eval them to know how good an assistant response is. We've built a number of LLM apps, and while we could ship decent tech demos, we were disappointed with how they'd perform over time. We worked with a few companies who had the same problem, and found out scientifically building prompts and evals is far from a solved problem... writing these things feels more like directing a play than coding. Inspired by Anthropic's constitutional ai concepts, and amazing…

    2025 · github.com

  20. 20PE

    Spelltest framework simulates conversations between AI ‘synthetic users' in an environment to test and refine LLM-based applications. It ensures your app converse with utmost accuracy and relevance. Post-chat, Spelltest assesses responses, providing qualitative and quantitative feedback on performance. Suitable for both chat and completion modes. When to use: - After modifying your prompt. - When your LLM provider updates. - As a CI step for you repo. All feedback and collaborations appreciated!

    2023 · github.com

  21. 21OS

    Hey HN! We built EvalKit, a library you embed to capture agent actions and a UI where domain experts give feedback, evaluate and improve AI agents. We experienced, in large agentic systems, prompt-engineering or auto-prompt improvement tool can get accuracy from 0 to 50% but for increasing accuracy to 100% we had to work with domain experts. Example -> In a law ai agent, lawyers are needed because law is complex and lawyers have a deeper context compared to non-lawyers. Other evaluation tools in the market focus on the experience of the developer and we are focusing on making as easy as…

    2025 · github.com

  22. 22

    1# Assessment Platform MENA: Meeting People, not just CVs

    2025

  23. 23
    Evaly11

    Modern online assessment platform

    Dec 2025 · evaly.io

  24. 24

    One prompt, every model,Test smarter. Build faster

    2025

Ranked by how close each launch is in meaning, then by votes. Refine with a description →