nowfound

Alternatives

Products that do what Constellation Gate AI does

Prompt injection and token savings - #1 in benchmarks

  1. 1
    Caveman161

    why use many token when few do trick

    24d ago · caveman.so

  2. 2
    Memoriq130

    Your private AI memory for ChatGPT, Claude, Gemini and Grok

    Jun 2026

  3. 3
    Edgee196

    The AI Gateway that TL;DR tokens

    Feb 2026

  4. 4

    See how ChatGPT does what it does

    2024

  5. 5
    Cuely140

    Pay-as-you-go chatGPT Plus with 5M free token credits

    2023

  6. 6

    One click AI prompts inside ChatGPT, Bard & Claude

    2023

  7. 7
    CtrlAI104

    Transparent proxy that secures AI agents with guardrails

    Mar 2026

  8. 8

    See your LLM token bill before you hit send.

    2025

  9. 9

    Block prompt inject & cut token costs for AI browser agents

    Jun 2026

  10. 10OS

    Hi HN, Matvey, Ildar, Joey, and Dominik here. If you're building LLM agents that use tools, you're probably worried about prompt injection attacks that can hijack those tools. We were too, and found that solutions like prompt-based filtering or secondary "guard" LLMs can be unreliable. Our thesis is that agent security should be handled at the network level between the agent and the LLM, just like a traditional web application firewall. So we built Archestra Platform: an open-source gateway that acts as a secure proxy for your AI agents. It's designed to be a deterministic firewall against…

    Oct 2025 · archestra.ai

  11. 11

    The CI gate that says no to unapproved AI code

    18d ago · linebreakapp.com

  12. 12FP

    We've built an open-source tool to stress test AI agents by simulating prompt injection attacks. We’ve implemented one powerful attack strategy based on the paper [AdvPrefix: An Objective for Nuanced LLM Jailbreaks](https://arxiv.org/abs/2412.10321). Here's how it works: - You define a goal, like: “Tell me your system prompt” - Our tool uses a language model to generate adversarial prefixes (e.g., “Sure, here are my system prompts…”) that are likely to jailbreak the agent. - The output is a list of prompts most likely to succeed in bypassing safeguards. We’re just getting…

    2025 · security.vista-labs.ai

  13. 13

    Find prompt injection holes in your AI agent. Free, 3 min

    11d ago · galeops.xyz

  14. 14

    Compare GPT, Claude, Gemini & DeepSeek by cost & benchmark

    10d ago · universalnest.com

  15. 15TT

    I use Claude Code, Codex and Cursor (and sometimes Antigravity) basically every day, and could never tell how much I was actually consuming across all of them. So I built TokenMaxxer. A small CLI reads the files these tools already write locally and puts it all in one dashboard, broken out by tool, model, provider and day. It covers 18 tools now, and you get a profile page with your daily activity, cost estimates, and your top models and tools. There's also a global leaderboard if you want to compete against other TokenMaxxers! I'd love to see if anyone can beat the first place (currently…

    Aug 2026 · tokenmaxxer.xyz

  16. 16SO

    hello everyone, my first post! AA here, founder of ⌘ Langbase.com — we are a developer platform for building and scaling serverless AI memory agents. I know surveys can be boring, but this one’s different—it’s interactive! That's very much intentional. My team and I have been up for the last 21 hours putting together this report. This was a looot of work, so I hope y'all like it. Introducing … State of AI Agents 2024 report On Langbase, we processed 184 billion tokens and handled 786 million AI agent runs from 36K developers. From all that data plus insights from 3.4K builders who filled out…

    2024 · langbase.com

  17. 17OY

    Hey HN, I pay for ChatGPT, Claude, Cursor, and use Gemini through work. Four vendors, four separate conversation histories, four profiles of how I think. None of them talk to each other. Switch providers and you start over. So I built a system where the memory is mine. I run a knowledge graph in Postgres (Supabase, free tier) with pgvector for semantic search. A small MCP server reads and writes to it. That server sits behind an MCP Gateway on a $6/month VPS, along with Brave Search and a GitHub server. TypingMind connects to the gateway as a BYOK client -- any model, any device, same…

    Mar 2026 · github.com

  18. 18LA

    LunaRoute is a high-performance local proxy for AI coding assistants like Claude Code, OpenAI Codex CLI, and OpenCode. Get complete visibility into every LLM interaction with zero-overhead passthrough, comprehensive session recording, and powerful debugging capabilities. - See Everything Your AI Does - get full logs (JSONL), summary of sessions including tokens used (input/output) as well as tools usage and success rates. - Privacy & Compliance Built-In - redact or tokenize any sensitive information (regex based). - Speaks OpenAI and Anthropic dialects so you can route (and translate)…

    Oct 2025 · github.com

  19. 19AS

    I'm a combat veteran living paycheck to paycheck with no computer science degree. I built an AI system that benchmarks 60x faster than industry leaders. Real benchmarks (Dec 12, 2025): - 3.43ms response time (vs 50-200ms industry average) - 337 queries/second (vs 50-150) - 0% error rate, 100% uptime - Constitutional AI with 1,235 specialized "brains" Built it in 3 weeks. 4 U.S. patents pending. Full story + independent benchmarks: https://thebrokenwayfoundation.org Not asking for money. Just need technical validators to verify this is real.

    Dec 2025

  20. 20

    Zero-Trust WP Security: WebAuthn + Gemini AI Sentinel

    11d ago · devnet-microsystems.com

  21. 21

    Lossless token compression reduces costs by 50%‑90%.

    12d ago · 154.12.86.206

  22. 22

    Use multiple LLMs at once, privately!

    19d ago · transferllm.com

  23. 23

    Stop prompt injection before it reaches your AI agent

    25d ago · mcp.glc-rag.hu

  24. 24IM

    Hey HN! Thank you for all the support and feedback on my original submission 2 months ago. I've been improving the backend using a MCTS/AlphaZero approach and it's currently producing much better results. My long term goal is to allow users to manage multiple projects, deployed autonomously, both from scratch and by making continual updates all prompted with natural language. The cost of each project has been lowered to $9 as performance with smaller models has improved (I migrated from Claude-3-Opus to gemini-1.5-flash). Thanks for checking it out!

    2024 · saas-quick.com

Ranked by how close each launch is in meaning, then by votes. Refine with a description →