nowfound

Alternatives

Products that do what Alpaca.cpp – Run an Instruction-Tuned Chat-Style LLM on a MacBook does

  1. 1IM

    Hi Hackers, Excited to share a macOS app I've been working on: https://recurse.chat/ for chatting with local AI. While it's amazing that you can run AI models locally quite easily these days (through llama.cpp / llamafile / ollama / llm CLI etc.), I missed feature complete chat interfaces. Tools like LMStudio are super powerful, but there's a learning curve to it. I'd like to hit a middleground of simplicity and customizability for advanced users. Here's what separates RecurseChat out from similar apps: - UX designed for you to use local AI as a daily driver.…

    2024 · recurse.chat

  2. 2WD
  3. 3CA

    ChatLLaMA is an experimental chatbot interface for interacting with variants of Facebook's LLaMA. Currently, we support the 7 billion parameter variant that was fine-tuned on the Alpaca dataset. This early versions isn't as conversational as we'd like, but over the next week or so, we're planning on adding support for the 30 billion parameter variant, another variant fine-tuned on LAION's OpenAssistant dataset and more as we explore what this model is capable of. If you want deploy your own instance is the model powering the chatbot and build something similar we've open sourced the Truss…

    2023 · chatllama.baseten.co

  4. 4DA
  5. 5PO

    Hi HN, OpenAI recently released a model for automatic speech recognition called Whisper [0]. I decided to reimplement the inference of the model from scratch using C/C++. To achieve this I implemented a minimalistic tensor library in C and ported the high-level architecture of the model in C++. The entire code is less than 8000 lines of code and is contained in just 2 source files without any third-party dependencies. The Github project is here: https://github.com/ggerganov/whisper.cpp With this implementation I can very easily build and run the model - “make…

    2022 · github.com

  6. 6AT

    Github: https://github.com/Arthur-Ficial/apfel

    Apr 2026 · apfel.franzai.com

  7. 7

    The easiest way to chat with local AI

    2025

  8. 8FL

    I've been playing around with https://github.com/zphang/minimal-llama/ and https://github.com/tloen/alpaca-lora/blob/main/finetune.py, and wanted to create a simple UI where you can just paste text, tweak the parameters, and finetune the model quickly using a modern GPU. To prepare the data, simply separate your text with two blank lines. There's an inference tab, so you can test how the tuned model behaves. This is my first foray into the world of LLM finetuning, Python, Torch, Transformers, LoRA, PEFT, and Gradio. Enjoy!

    2023 · github.com

  9. 9
    ModelHub318

    The missing menu bar app for local LLMs on Mac.

    May 2026 · studio.consciousengines.com

  10. 10
    Unabyss723

    MCP-native self-updating context layer for your AI

    May 2026 · unabyss.com

  11. 11WM

    We wrote our inference engine on Rust, it is faster than llama cpp in all of the use cases. Your feedback is very welcomed. Written from scratch with idea that you can add support of any kernel and platform.

    2025 · github.com

  12. 12

    Massive local model speedup on Apple Silicon with MLX

    Apr 2026 · ollama.com

  13. 13

    Free real-time stock market data API

    2020

  14. 14

    Simple REST API for commission-free stock trading

    2018

  15. 15

    Easy to use crypto trading API

    2021

  16. 16
    GPT4All107

    A chatbot trained on a massive collection of clean data

    2023

  17. 17

    Launch your own commission-free trading app

    2021

  18. 18OR

    Hi HN A few folks and I have been working on this project for a couple weeks now. After previously working on the Docker project for a number of years (both on the container runtime and image registry side), the recent rise in open source language models made us think something similar needed to exist for large language models too. While not exactly the same as running linux containers, running LLMs shares quite a few of the same challenges. There are "base layers" (e.g. models like Llama 2), specific configuration to run correctly (parameters, temperature, context window sizes etc). There's…

    2023 · github.com

  19. 19
    Ollama235

    The easiest way to run large language models locally

    2023

  20. 20
    ManePaw85

    Find documents on your Mac using natural language

    Feb 2026

  21. 21

    Adopt unique alpaca on blockchain and shear wools everyday

    2018

  22. 22

    Connect your Mac's Mail, Calendar, and more to AI

    Apr 2026 · claunnector.com

  23. 23CO

    Hey HN, Henry and Roman here - we've been building a cross-platform framework for deploying LLMs, VLMs, Embedding Models and TTS models locally on smartphones. Ollama enables deploying LLMs models locally on laptops and edge severs, Cactus enables deploying on phones. Deploying directly on phones facilitates building AI apps and agents capable of phone use without breaking privacy, supports real-time inference with no latency, we have seen personalised RAG pipelines for users and more. Apple and Google actively went into local AI models recently with the launch of Apple Foundation Frameworks…

    2025 · github.com

  24. 24

    Commission-free API to trade & invest for as little as $1

    2021

Ranked by how close each launch is in meaning, then by votes. Refine with a description →