nowfound

Alternatives

Products that do what Llama-dl – high-speed download of LLaMA, Facebook's 65B GPT model does

  1. 1
    Llama312

    3.1-405B: an open source model to rival GPT-4o / Claude-3.5

    2024

  2. 2

    Llama 405B-level performance, at a fraction of the cost

    2024

  3. 3
    Llama 2263

    The next generation of Meta's open source LLM

    2023

  4. 4
    LLaMA118

    A foundational, 65-billion-parameter large language model

    2023

  5. 5
    Llama 4423

    A new era of natively multimodal AI innovation

    2025

  6. 6CA

    ChatLLaMA is an experimental chatbot interface for interacting with variants of Facebook's LLaMA. Currently, we support the 7 billion parameter variant that was fine-tuned on the Alpaca dataset. This early versions isn't as conversational as we'd like, but over the next week or so, we're planning on adding support for the 30 billion parameter variant, another variant fine-tuned on LAION's OpenAssistant dataset and more as we explore what this model is capable of. If you want deploy your own instance is the model powering the chatbot and build something similar we've open sourced the Truss…

    2023 · chatllama.baseten.co

  7. 7
    Llamao509

    Private & offline alternative to ChatGPT on your device

    2025

  8. 8LF
  9. 9AF

    We believe that AI should be fully open source and part of the collective knowledge. The original LLaMA code is GPL licensed which means any project using it must also be released under GPL. This "taints" any other code and prevents meaningful academic and commercial use. Lit-LLaMA solves that for good.

    2023 · github.com

  10. 10FL

    2024 · colab.research.google.com

  11. 11LS
  12. 128F

    Hi HN! I'm just sharing a project I've been working on during the LLM Efficiency Challenge - you can now finetune Llama with QLoRA 5x faster than Huggingface's original implementation on your own local GPU. Some highlights: 1. Manual autograd engine - hand derived backprop steps. 2. QLoRA / LoRA 80% faster, 50% less memory. 3. All kernels written in OpenAI's Triton language. 4. 0% loss in accuracy - no approximation methods - all exact. 5. No change of hardware necessary. Supports NVIDIA GPUs since 2018+. CUDA 7.5+. 6. Flash Attention support via Xformers. 7. Supports 4bit and 16bit…

    2023 · github.com

  13. 13

    AI-powered coding assistant by Meta

    2024

  14. 14

    New, performant version of Meta's LLM for code generation

    2024

  15. 15FL

    I've been playing around with https://github.com/zphang/minimal-llama/ and https://github.com/tloen/alpaca-lora/blob/main/finetune.py, and wanted to create a simple UI where you can just paste text, tweak the parameters, and finetune the model quickly using a modern GPU. To prepare the data, simply separate your text with two blank lines. There's an inference tab, so you can test how the tuned model behaves. This is my first foray into the world of LLM finetuning, Python, Torch, Transformers, LoRA, PEFT, and Gradio. Enjoy!

    2023 · github.com

  16. 16RL

    2024 · app.wiz.chat

  17. 17IB

    I spent the last few days building out a nicer ChatGPT-like interface to use Mistral 7B and Llama 3 fully within a browser (no deps and installs). I’ve used the WebLLM project by MLC AI for a while to interact with LLMs in the browser when handling sensitive data but I found their UI quite lacking for serious use so I built a much better interface around WebLLM. I’ve been using it as a therapist and coach. And it’s wonderful knowing that my personal information never leaves my local computer. Should work on Desktop with Chrome or Edge. Other browsers are adding WebGPU support as well - see…

    2024 · github.com

  18. 18L3
  19. 19KC

    Hi folks, we're Debanjum and Saba. We created Khoj as a hobby project 2+ years ago because: (1) Search on the desktop sucked; we just had keyword search on the desktop vs google for the internet; and (2) Natural language search models had become good and easy to run on consumer hardware by this point. Once we made Khoj search incremental, I completely stopped using the default incremental search (C-s) in Emacs. Since then Khoj has grown to support more content types, deeper integrations and chat (using ChatGPT). With Llama 2 released last week, chat models are finally good and easy enough to…

    2023 · github.com

  20. 20L3

    I spent a lot of time and money on this rather big side project of mine that attempts to replicate the mechanistic interpretability research on proprietary LLMs that was quite popular this year and produced great research papers by Anthropic [1], OpenAI [2] and Deepmind [3]. I am quite proud of this project and since I consider myself the target audience for HackerNews did I think that maybe some of you would appreciate this open research replication as well. Happy to answer any questions or face any feedback. Cheers [1]…

    2024 · github.com

  21. 21LT

    2023 · github.com

  22. 22
    Ollama235

    The easiest way to run large language models locally

    2023

  23. 23

    The most capable openly available LLM to date

    2024

  24. 24
    Dream 7B191

    Powerful Open Diffusion LLM, Beyond Autoregressive

    2025

Ranked by how close each launch is in meaning, then by votes. Refine with a description →