nowfound

Alternatives

Products that do what Confident AI does

all-in-one LLM evaluation platform

  1. 1

    Open-source evaluations and observability for LLM apps

    2024

  2. 2
    LLM Stats308

    Compare API models by benchmarks, cost & capabilities

    Oct 2025 · llm-stats.com

  3. 3
    Selene 1196

    Evaluate your AI app with the most accurate LLM Judge

    2025

  4. 4
    LLMWare358

    Dev tool to make AI apps to deploy privately or locally

    2024

  5. 5

    Use any AI model with just one API

    2025

  6. 6
    Agenta362

    Open-source prompt management & evals for AI teams

    Nov 2025 · agenta.ai

  7. 7

    Open-source stack for industrial-grade LLM applications

    2025 · github.com

  8. 8

    Validate, monitor, and safeguard LLM-based apps

    2023

  9. 9
    Dify.AI261

    Open-source platform for LLMOps, define your AI-native apps

    2023

  10. 10
    AskCodi230

    Custom LLMs, without training. Use via openai compatible api

    Nov 2025 · askcodi.com

  11. 11

    Test-driven development for LLMs

    2023

  12. 12DA
  13. 13
    Taylor AI118

    Fine-tune open source LLMs in minutes

    2023

  14. 14

    Vibe-check many open-source and proprietary LLMs at once

    2024

  15. 15TL
  16. 16FA

    Hey HN! We’re building FinetuneDB (https://finetunedb.com/), an LLM fine-tuning platform. It enables teams to easily create and manage high-quality datasets, and streamlines the entire workflow from fine-tuning to serving and evaluating models with domain experts. You can check out our docs here: (https://docs.finetunedb.com/) FinetuneDB exists because creating and managing high-quality datasets is a real bottleneck when fine-tuning LLMs. The quality of your data directly impacts the performance of your fine-tuned models, and existing tools didn’t offer an easy…

    2024 · finetunedb.com

  17. 17LS

    Hi, I was a corporate lawyer for many years working with a lot of financial services and insurance companies. In practicing law, I noticed there was a lot of repetition in the tasks I was working on even as a highly paid attorney that could be automated. I wanted to solve the problem of dealing with a lot information and data in a practical way, using AI. This motivated me to start AI Bloks/LLMWare with my husband, who had a deep background in software and is a very early adopter of AI. We have been on this journey with our open source project LLMWare for the past 4 months, producing a…

    2024 · github.com

  18. 18FT

    2024 · github.com

  19. 19IL

    I have been working in AI space for a while now, first at FAANG with ML since 2021, then with LLM in start-ups since early 2023. I think LLM Application development is extremely iterative, more so than any other types of development. This is because to improve an LLM application performance (accuracy, hallucinations, latency, cost), you need to try various combinations of LLM models, prompt templates (e.g., few-shot, chain-of-thought), prompt context with different RAG architecture, different agent architecture, and more. There are thousands of possible combinations and you need a process…

    2024 · github.com

  20. 20OS

    Hi everyone, we’re a small team, supported by Mozilla, who are working on re-imagining a UI for training, tuning and testing local LLMs. Everything is open source. If you’ve been training your own LLMs or have always wanted to, we’d love for you to play with the tool and give feedback on what the future development experience for LLM engineering could look like.

    2025 · github.com

  21. 21DE
  22. 22

    Vibe Training AI models

    Mar 2026

  23. 23AF

    We’ve built an AI risk assessment tool designed specifically for GenAI/LLM applications. It's still early, but we’d love your feedback. Here’s what it does: 1. it performs comprehensive AI risk assessments by analyzing your codebase against different AI regulation/framework or even internal policies. It identifies potential issues and suggests fixes directly through one click PRs. 2. the first framework the platform supports is OWASP Top 10 for LLM Applications 2025, upcoming framework will be ISO 42001 as well as custom policy documents. 3. we're a small, early stage team, so the…

    2025 · gettavo.com

  24. 24PH

    Hey HN, Hakim here from Fini (YC S22), a startup focused on providing automated customer support bots for enterprises that have a high volume of support requests. Today, one of the largest use cases of LLMs is for the purpose of automating support. As the space has evolved over the past year, there has subsequently been a need for evaluations of LLM outputs - and a sea of LLM Evals packages have been released. "LLM evals" refer to the evaluation of large language models, assessing how well these AI systems understand and generate human-like text. These packages have recently relied on…

    2024 · github.com

Ranked by how close each launch is in meaning, then by votes. Refine with your own description →