nowfound

Alternatives

Products that do what oMLX does

Mac LLM server that cuts agent wait times from 90s to 5s Discussion | Link

  1. 1RM
  2. 2CO
  3. 3BL

    2017 · engblog.nextdoor.com

  4. 4AA
  5. 5YY
  6. 6TA

    2018 · bonhardcomputing.com

  7. 7OA
  8. 8PA
  9. 9AA
  10. 10ML

    Time to first token is 39% faster Agent wall times decrease by 46% No swaps Tracks your resource usage in real-time and adjusts how the model runs so that it works perfectly on your device. Implements KV cache sizing, prefix caching, live RAM pressure management, context trimming, KV quantization, and more. Built a ton of features

    Jun 2026 · autotunellm.com

  11. 11AD
  12. 12TC

    Jun 2026 · fuckui.com

  13. 13GL

    2017 · latency.apex.sh

  14. 14GA
  15. 15AR

    If you're interested in exploring what LLM-based agent systems these days actually do to solve certain benchmarks such as SWEBench or WebArena, we created a small leaderboard with our team, that allows to view a lot of public and OSS agent results including all the runtime traces (the step-by-step reasoning behind the scenes). Looking at traces is actually quite interesting, as they reveal a lot about the inner working and shortcomings of current agent system, e.g. see https://explorer.invariantlabs.ai/u/invariant/webarena--SteP... for an example trace.

    2024 · explorer.invariantlabs.ai

  16. 16AS
  17. 17AS

    Hi, this is my first npm package that is getting a couple of hundreds downloads. It's a super simple CLI time tracker, mostly suitable for 9 to 5 jobs I hope you check it out and tell me what you think. Demo and etc. https://github.com/omidfi/moro

    2017

  18. 18LA

    2024 · twitter.com

  19. 19AL
  20. 20UL
  21. 21AC
  22. 22ZG
  23. 23IB

    hey hn, I built an open-source Perplexity clone that can run local LLMs and cloud LLMs. It's fully self-hostable through Docker and uses ollama to support local LLMs. The demo video in the repository shows me running it locally with llama3 on my M1 Macbook Pro. I'm open to any suggestions or feedback, thanks!

    2024 · github.com

  24. 24RL

Ranked by how close each launch is in meaning, then by votes. Refine with a description →