nowfound

Alternatives

Products that do what Llamero – A GUI app to easily download, install and infer LLaMA models does

  1. 1
    Llama 4423

    A new era of natively multimodal AI innovation

    2025

  2. 2
    Llama 2263

    The next generation of Meta's open source LLM

    2023

  3. 3
    Ollama235

    The easiest way to run large language models locally

    2023

  4. 4LD
  5. 5
    Llama312

    3.1-405B: an open source model to rival GPT-4o / Claude-3.5

    2024

  6. 6

    Llama 405B-level performance, at a fraction of the cost

    2024

  7. 7LF
  8. 88F

    Hi HN! I'm just sharing a project I've been working on during the LLM Efficiency Challenge - you can now finetune Llama with QLoRA 5x faster than Huggingface's original implementation on your own local GPU. Some highlights: 1. Manual autograd engine - hand derived backprop steps. 2. QLoRA / LoRA 80% faster, 50% less memory. 3. All kernels written in OpenAI's Triton language. 4. 0% loss in accuracy - no approximation methods - all exact. 5. No change of hardware necessary. Supports NVIDIA GPUs since 2018+. CUDA 7.5+. 6. Flash Attention support via Xformers. 7. Supports 4bit and 16bit…

    2023 · github.com

  9. 9FL

    I've been playing around with https://github.com/zphang/minimal-llama/ and https://github.com/tloen/alpaca-lora/blob/main/finetune.py, and wanted to create a simple UI where you can just paste text, tweak the parameters, and finetune the model quickly using a modern GPU. To prepare the data, simply separate your text with two blank lines. There's an inference tab, so you can test how the tuned model behaves. This is my first foray into the world of LLM finetuning, Python, Torch, Transformers, LoRA, PEFT, and Gradio. Enjoy!

    2023 · github.com

  10. 10

    AI-powered coding assistant by Meta

    2024

  11. 11
    Llamao509

    Private & offline alternative to ChatGPT on your device

    2025

  12. 12LA

    A simple mobile web app inspired by Fuzzy-Search/realtime-bakllava that uses llama.cpp server backend with multimodal mode to describe and narrate what the phone camera sees. I built this thing in a few hours using a single ChatGPT thread to generate most things for me and iterate on this project. Here's the workflow: https://chat.openai.com/share/ea84ec69-5617-45e8-8772-ac2dcf...

    2023 · github.com

  13. 13
    ModelHub318

    The missing menu bar app for local LLMs on Mac.

    May 2026 · studio.consciousengines.com

  14. 14
    LLaMA118

    A foundational, 65-billion-parameter large language model

    2023

  15. 15

    Run leading vision models locally with the new engine

    2025

  16. 16
    Llama163

    A fun, flexible, task manager for desktop web

    2020

  17. 17

    Let Llama take over your desktop

    2025

  18. 18RL

    2024 · app.wiz.chat

  19. 19OR

    Hi HN A few folks and I have been working on this project for a couple weeks now. After previously working on the Docker project for a number of years (both on the container runtime and image registry side), the recent rise in open source language models made us think something similar needed to exist for large language models too. While not exactly the same as running linux containers, running LLMs shares quite a few of the same challenges. There are "base layers" (e.g. models like Llama 2), specific configuration to run correctly (parameters, temperature, context window sizes etc). There's…

    2023 · github.com

  20. 20
    Apollo AI280

    Run local models like Llama on iOS

    2025

  21. 21LA

    2020 · llamalife.co

  22. 22
    Kuzco216

    Open-source Swift package to run LLMs locally on iOS & macOS

    2025

  23. 23

    Build Once and Deploy Anywhere

    2025

  24. 24LT

    2023 · github.com

Ranked by how close each launch is in meaning, then by votes. Refine with a description →