Alternatives
Products that do what Llm-benchmark – Benchmarks LLM-optimized code across multiple providers does
- 1LC
2023 · github.com
- 2LS
2023 · github.com
- 3LL
2023 · github.com
- 4LD
Mar 2026 · github.com
- 5BA
2015 · github.com
- 6SF
2024 · github.com
- 7LR
Sep 2025 · github.com
- 8LT
2025 · github.com
- 9AO
Hi, We are building an open-source framework for loading and structuring LLM context to create accurate and explainable LLM answers using knowledge graphs and vector stores. We built the tool with four main concepts in mind: 1. Loader -> uses dlt in the backend to load and structure the data 2. Cognify step -> creates a graph with summaries, labels and factoids that are interconnected across the documents and stored as a representation in the vector store 3. Optimizer -> Uses DSPy to optimize LLM queries, and we plan to extend it to most of the knobs we can turn, like chunking etc. 4. Search…
2024 · github.com
- 10CO
Apr 2026 · npmjs.com
- 11IL
LLM Application development is extremely iterative, more so than any other types of development. This is because in addition to all the activities involved in regular application development, we also need to make the LLM Application accurate and reduce hallucination. To improve performance, we need to trial and error various combinations of LLM models, prompt templates (e.g., few-shot, chain-of-thought), prompt context with different RAG architecture, try different agent architecture, and more. There are thousands of permutations to try. We need to be able to easily experiment with these…
2024 · palico.ai
- 12AS
2025 · github.com
- 13BF
2023 · e2b.dev
- 14LT
Measures the ability of various LLMs to navigate a fictional codebase via iterative directory tree expansion and observation. Each model's baseline ability is compared against combinations of various prompt engineering mods to quantify exactly how much they help or hinder the LLM. Interesting findings here: https://github.com/aiwebb/treenav-bench#interesting-findings
2024 · github.com
- 15IL
2023 · github.com
- 16LP
2025 · codehooks.io
- 17FH
2023 · product.distoai.com
- 18IB
I was overspending on GPT-4o. It was really hard to compare different models I could switch to, so I built this LLM comparison tool. It shows leaderboards, pricing, and performance data across 100+ LLMs (including all major providers and open-source models). Key features: - Live pricing comparisons - Benchmark Scores (MMLU, HumanEval, GPQA, etc.) - Context length vs cost analysis - Speed/throughput tests across providers - Quality vs price visualizations - Open source (all data verifiable) Try it out: https://llmstats.com I'd like to know your opinion :) Tech stack: Next.js,…
2025 · llm-stats.com
- 19LC
Sep 2025 · github.com
- 20LA
Feb 2026 · github.com
- 21VA
2023 · sdk.vercel.ai
- 22BI
2014 · github.com
- 23CA
Jun 2026 · github.com
- 24LF
Jan 2026 · tinytune.xyz
Ranked by how close each launch is in meaning, then by votes. Refine with a description →