nowfound

Alternatives

Products that do what Llm-benchmark – Benchmarks LLM-optimized code across multiple providers does

  1. 1LC

    2023 · github.com

  2. 2LS
  3. 3LL
  4. 4LD

    Mar 2026 · github.com

  5. 5BA
  6. 6SF
  7. 7LR

    Sep 2025 · github.com

  8. 8LT
  9. 9AO

    Hi, We are building an open-source framework for loading and structuring LLM context to create accurate and explainable LLM answers using knowledge graphs and vector stores. We built the tool with four main concepts in mind: 1. Loader -> uses dlt in the backend to load and structure the data 2. Cognify step -> creates a graph with summaries, labels and factoids that are interconnected across the documents and stored as a representation in the vector store 3. Optimizer -> Uses DSPy to optimize LLM queries, and we plan to extend it to most of the knobs we can turn, like chunking etc. 4. Search…

    2024 · github.com

  10. 10CO
  11. 11IL

    LLM Application development is extremely iterative, more so than any other types of development. This is because in addition to all the activities involved in regular application development, we also need to make the LLM Application accurate and reduce hallucination. To improve performance, we need to trial and error various combinations of LLM models, prompt templates (e.g., few-shot, chain-of-thought), prompt context with different RAG architecture, try different agent architecture, and more. There are thousands of permutations to try. We need to be able to easily experiment with these…

    2024 · palico.ai

  12. 12AS

    2025 · github.com

  13. 13BF
  14. 14LT

    Measures the ability of various LLMs to navigate a fictional codebase via iterative directory tree expansion and observation. Each model's baseline ability is compared against combinations of various prompt engineering mods to quantify exactly how much they help or hinder the LLM. Interesting findings here: https://github.com/aiwebb/treenav-bench#interesting-findings

    2024 · github.com

  15. 15IL
  16. 16LP
  17. 17FH

    2023 · product.distoai.com

  18. 18IB

    I was overspending on GPT-4o. It was really hard to compare different models I could switch to, so I built this LLM comparison tool. It shows leaderboards, pricing, and performance data across 100+ LLMs (including all major providers and open-source models). Key features: - Live pricing comparisons - Benchmark Scores (MMLU, HumanEval, GPQA, etc.) - Context length vs cost analysis - Speed/throughput tests across providers - Quality vs price visualizations - Open source (all data verifiable) Try it out: https://llmstats.com I'd like to know your opinion :) Tech stack: Next.js,…

    2025 · llm-stats.com

  19. 19LC
  20. 20LA
  21. 21VA
  22. 22BI
  23. 23CA
  24. 24LF

Ranked by how close each launch is in meaning, then by votes. Refine with a description →