nowfound

Alternatives

Products that do what Maia Test Framework does

A pytest-based framework for testing multi AI agents systems

  1. 1

    UI dashboard for reviewing and debugging tests

    2025

  2. 2

    First AI agent automating entire software testing process

    2025 · testsprite.com

  3. 3

    Agentic testing for the AI-native team.

    Mar 2026 · testsprite.com

  4. 4

    Fully automate software testing end-to-end using AI

    2024

  5. 5
    Qwen3.5307

    The 397B native multimodal agent with 17B active params

    Feb 2026 · qwen.ai

  6. 6

    Cursor for testers. AI Agents for product and QA teams

    2025

  7. 7

    Let a fleet of parallel agents test your app in minutes

    May 2026 · testsprite.com

  8. 8
    TestAI407

    1,000+ automated tests for AI agents in one click

    2025

  9. 9
    Agenta362

    Open-source prompt management & evals for AI teams

    Nov 2025 · agenta.ai

  10. 10
    GLM-4.5298

    Unifying agentic capabilities in one open model

    2025

  11. 11

    Evaluate AI workflows and reach 99% AI quality.

    Oct 2025

  12. 12MO

    Hey HN, Anders and Tom here - we’ve been building an end-to-end testing framework powered by visual LLM agents to replace traditional web testing. We know there's a lot of noise about different browser agents. If you've tried any of them, you know they're slow, expensive, and inconsistent. That's why we built an agent specifically for running test cases and optimized it just for that: - Pure vision instead of error prone "set-of-marks" system (the colorful boxes you see in browser-use for example) - Use tiny VLM (Moondream) instead of OpenAI/Anthropic computer use for dramatically…

    2025 · github.com

  13. 13KA

    I built this because Cursor, Claude Code and other agentic AI tools kept giving me tests that looked fine but failed when I ran them. Or worse - I'd ask the agent to run them and it would start looping: fix tests, those fail, then it starts "fixing" my code so tests pass, or just deletes assertions so they "pass". Out of that frustration I built KeelTest - a VS Code extension that generates pytest tests and executes them, got hooked and decided to push this project forward... When tests fail, it tries to figure out why: - Generation error: Attemps to fix it automatically, then tries again -…

    Jan 2026 · keelcode.dev

  14. 14

    Get real-world tasks done with autonomous AI agents

    Jun 2026 · arena.ai

  15. 15OS

    GitHub - https://github.com/vostride/agent-qa Live Demos - https://vostride.com/demo/agent-qa

    May 2026 · vostride.com

  16. 16

    The open sparse MoE model for agentic coding

    Apr 2026 · qwen.ai

  17. 17EA

    Hey HN, I've been working on an open-source framework for creating AI agents that evolve, communicate, and collaborate to solve complex tasks. The Evolving Agents Framework allows agents to: Reuse, evolve, or create new agents dynamically based on semantic similarity Communicate and delegate tasks to other specialized agents Continuously improve by learning from past executions Define workflows in YAML, making it easy to orchestrate agent interactions Search for relevant tools and agents using OpenAI embeddings Support multiple AI frameworks (BeeAI, etc.) Current Status & Roadmap This is…

    2025 · github.com

  18. 18AB

    Hi everyone! My team and I just open-sourced a bunch of cool agent dev tools: Invariant Explorer to visually inspect and understand AI traces and a testing framework, building on pytest.

    2024 · github.com

  19. 19PE

    Spelltest framework simulates conversations between AI ‘synthetic users' in an environment to test and refine LLM-based applications. It ensures your app converse with utmost accuracy and relevance. Post-chat, Spelltest assesses responses, providing qualitative and quantitative feedback on performance. Suitable for both chat and completion modes. When to use: - After modifying your prompt. - When your LLM provider updates. - As a CI step for you repo. All feedback and collaborations appreciated!

    2023 · github.com

  20. 20

    The sweet-spot open dense model for coding agents

    Apr 2026 · qwen.ai

  21. 21AA
  22. 22HA
  23. 23
    Spec2734

    Spec-driven testing for AI agents and AI apps

    Apr 2026 · spec27.ai

  24. 24AA

Ranked by how close each launch is in meaning, then by votes. Refine with a description →