nowfound

Alternatives

Products that do what Selora – local model for Home Assistant does

Selora AI Local is an open-source, Qwen-based model for Home Assistant. Specs: Qwen3 1.7B base model (Q6 quantized~1.6GB) Four Home Assistant-specific LoRA adapters: - Answers - Clarifications - Automations - Commands ~3.5 GB total download size Runs locally via llama.cpp We chose a Qwen-based architecture because of a paper on Arxiv (link below) which applied a Qwen based model for local LLM configuration, and showed promising results. We took it a step further in application by training LoRA adapters specialized in Home Assistant configuration. This alpha release ships our base model with…

  1. 1RA
  2. 28F

    Hi HN! I'm just sharing a project I've been working on during the LLM Efficiency Challenge - you can now finetune Llama with QLoRA 5x faster than Huggingface's original implementation on your own local GPU. Some highlights: 1. Manual autograd engine - hand derived backprop steps. 2. QLoRA / LoRA 80% faster, 50% less memory. 3. All kernels written in OpenAI's Triton language. 4. 0% loss in accuracy - no approximation methods - all exact. 5. No change of hardware necessary. Supports NVIDIA GPUs since 2018+. CUDA 7.5+. 6. Flash Attention support via Xformers. 7. Supports 4bit and 16bit…

    2023 · github.com

  3. 3
    Qwen 2.5290

    Alibaba's latest AI model series

    2025

  4. 4

    0.8B-9B native multimodal w/ more intelligence, less compute

    Mar 2026

  5. 5
    Lora411

    Integrate local LLM, with one line of code

    2025

  6. 6

    Run Qwen's latest models locally on your iPhone

    Mar 2026

  7. 7

    Qwen’s most capable model for coding and cowork

    Aug 2026 · qwen.ai

  8. 8
    Qwen3.5307

    The 397B native multimodal agent with 17B active params

    Feb 2026

  9. 9

    Multimodal AI optimized for real-world coding agents

    Apr 2026

  10. 10
    Apollo AI280

    Run local models like Llama on iOS

    2025

  11. 11
    Llama312

    3.1-405B: an open source model to rival GPT-4o / Claude-3.5

    2024

  12. 12

    A powerful open model for agentic coding tasks

    2025

  13. 13

    SOTA open-source T2I model with even greater realism

    Jan 2026

  14. 14

    Qwen's now in mobile chat

    2025

  15. 15

    Qwen's most advanced reasoning model yet

    2025

  16. 16

    Llama 405B-level performance, at a fraction of the cost

    2024

  17. 17FL

    I've been playing around with https://github.com/zphang/minimal-llama/ and https://github.com/tloen/alpaca-lora/blob/main/finetune.py, and wanted to create a simple UI where you can just paste text, tweak the parameters, and finetune the model quickly using a modern GPU. To prepare the data, simply separate your text with two blank lines. There's an inference tab, so you can test how the tuned model behaves. This is my first foray into the world of LLM finetuning, Python, Torch, Transformers, LoRA, PEFT, and Gradio. Enjoy!

    2023 · github.com

  18. 18

    The open sparse MoE model for agentic coding

    Apr 2026 · qwen.ai

  19. 19

    The sweet-spot open dense model for coding agents

    Apr 2026 · qwen.ai

  20. 20

    A native omni model for voice, video, and tools

    Mar 2026

  21. 21Q6
  22. 22

    The end-to-end model powering multimodal chat

    2025

  23. 23
    Tiny Aya211

    Local, open-weight AI designed for real-world languages

    Apr 2026 · cohere.com

  24. 24
    Ollama235

    The easiest way to run large language models locally

    2023

Ranked by how close each launch is in meaning, then by votes. Refine with a description →