nowfound

AI · May 3, 2024

PI

Paramount is an OSS package that captures expert feedback on LLM chats

Hey HN, Hakim here from Fini (YC S22). We've seen first hand how AI chat projects pan out, and so have released an OSS library to ensure the industry gets more tools for improving outcomes. Many AI chat projects are scrapped due to persistent inaccuracies in LLM responses. Paramount is an open-source Python package designed to bridge the gap between LLM-generated and ideal responses by incorporating expert feedback directly into the evaluation process. It provides a robust framework for recording LLM function outputs (ground truth data) and facilitates agent evaluations, reducing the time to…

What it does

In the maker’s words, at launch

Hey HN, Hakim here from Fini (YC S22). We've seen first hand how AI chat projects pan out, and so have released an OSS library to ensure the industry gets more tools for improving outcomes. Many AI chat projects are scrapped due to persistent inaccuracies in LLM responses. Paramount is an open-source Python package designed to bridge the gap between LLM-generated and ideal responses by incorporating expert feedback directly into the evaluation process. It provides a robust framework for recording LLM function outputs (ground truth data) and facilitates agent evaluations, reducing the time to identify and correct errors. Developers can integrate Paramount with a decorator that logs LLM interactions into a CSV or database, followed by a straightforward UI for expert review. This process accelerates the debugging and validation phase of your project and de-risks your launch.

Does the same job

all alternatives →
  • Helpedby AI2025 · ▲122

    Revolutionize LLMs chat platform with pay-as-you-go pricing

  • Lemonfox.ai2023 · ▲157

    The fast, easy and cheap OpenAI alternative

  • maxly.chat2025 · ▲56

    GitHub for LLMs

  • ClearPlanSep 2025 · ▲48

    1st editor focus on enhancing LLM output seamlessly.

  • HH
    HumanLayer – Human-in-the-Loop for AI Agents2024 · github.com · ▲5

    I found myself building a bunch of LLM-backed features that needed to use tool calling, and some of those tools involved doing things that were somewhat high stakes - communicating on my behalf or modifying shared / production data. one example - I wanted to replace a marketing website with a chatbot + vector DB loaded with the previous content, docs, and blog posts. Between hallucinations, missing knowledge base info, and the LLM generally writing like an psuedo-intellectual high schooler, I realized I couldn't trust it to communicate unsupervised with my website visitors. I needed a…

  • LS
    LLMStack – Self-Hosted, Low-Code Platform to Build AI Experiences2023 · github.com · ▲7

    LLMStack is a low-code platform that can be used to build LLM apps, chatbots and integrate AI experiences into existing products/workflows. It comes with everything out of the box that one needs to build LLM apps locally. It can also be used in a multi-tenant setting, making it available for everyone to use in an enterprise. Some highlights of the platform: - Chain multiple LLM models allowing for complex pipelines - Includes a vector database and necessary connectors to help enrich LLM responses with private data - App templates tailored to specific use cases to quickly build LLM apps…

More ai this month

the category →
  • I trained a 125M-parameter transformer to autocomplete piano performances in real time (~108 notes/sec on an iPhone 15). The idea is basically GitHub Copilot or Tabnine, except instead of prompting it with code, you prompt it by playing a few notes on a MIDI piano. The model then continues what you played, entirely on-device. The app is free if anyone wants to try it. Happy to answer questions about the model, training, Core ML, or the many things that didn't work.

    AI · 17d ago · simedw.com

  • Astute585

    Automate your B2B brand going viral, with new media creators

    AI · 18d ago · company-app.joinastute.com

  • Grok Bot547

    AI teammates that you can give real work to

    AI · 25d ago · x.ai

  • Hey HN, Henry from Cactus here! We previously released Cactus Needle, a 14MB agentic LLM for tool call, device use, and structured extraction for phones, wearables, smart homes, small robots and microcontrollers. We got really great feedback here, and have now incorporated the suggestions to release Needle 2. The whole model is a single 14MB binary that runs a full session in 28MB of RAM; 45m parameters at 2bit compression. Needle hits 500 tokens/sec decode speed on a Raspberry Pi 5, sits between 400-1,500 tokens/sec on VR devices like Meta Quest 3S and Apple Vision Pro, and ranges…

    AI · 26d ago · cactuscompute.com

  • Make your software self-driving

    AI · 30d ago · coldtea.ai

  • Soloop472

    Approval-first Agent OS for solo founders

    AI · 30d ago · soloop.io

Launched alongside, May 2024

the whole month →
  • Voicenotes1,487

    AI note-taker that's truly intelligent

    AI · 2024 · voicenotes.com

  • Wegic1,181

    The first AI web designer & developer by your side

    Dev tools · 2024 · wegic.ai

  • Insighto1,110

    Ship features users want

    Growth · 2024 · insigh.to

  • Ivee1,087

    The B2B influencer marketing platform

    Growth · 2024

  • Waxwing1,035

    AI copilot for every marketing task

    AI · 2024

  • AI powered zero-waste meal planner

    AI · 2024 · ohapotato.app