GPT-5 available for free on Gensee
TL;DR: we just made GPT-5 available for free on Gensee, a developer-oriented AI Agent optimization and deployment platform: https://platform.gensee.ai/ This is a crazy week with a bunch of model releases: gpt-oss, Claude-Opus-4.1, and now today's GPT-5. It may feel impossible for agent developers to keep up, with all the manual migrating, re-testing, and analyzing. We built Gensee to solve exactly this problem. Gensee lets you see the immediate impact of a new model on your already built agents and workflows. Here’s how it works: - Instant Model Swapping: Have an agent running…
What it does
In the maker’s words, at launch
TL;DR: we just made GPT-5 available for free on Gensee, a developer-oriented AI Agent optimization and deployment platform: https://platform.gensee.ai/ This is a crazy week with a bunch of model releases: gpt-oss, Claude-Opus-4.1, and now today's GPT-5. It may feel impossible for agent developers to keep up, with all the manual migrating, re-testing, and analyzing. We built Gensee to solve exactly this problem. Gensee lets you see the immediate impact of a new model on your already built agents and workflows. Here’s how it works: - Instant Model Swapping: Have an agent running on GPT-4o? With one click, you can clone it and swap the underlying model to GPT-5 family. No code changes, no re-deploying. - Automated A/B Testing & Analysis: Run your test cases against both versions of your agent simultaneously. Gensee gives you a side-by-side comparison of outputs, latency, and cost, so you can immediately see if GPT-5 improves quality or breaks your existing prompts and tool functions. - Smart Routing for Optimization: Gensee automatically selects the best combination of models for any given task in your agent to optimize for quality, cost, or speed. - Pre-built Agents: You can also grab one of our pre-built agents and immediately test it across the entire spectrum of new models to see how they compare. The goal is to eliminate the engineering overhead of model evaluation so you can spend your time building, not just updating. We'd love for you to try it out and give us feedback, especially if you have an existing project you want to benchmark against GPT-5. Join our Discord: https://discord.gg/qQr6SVW4
Does the same job
all alternatives →More ai this month
the category →
I trained a 125M-parameter transformer to autocomplete piano performances in real time (~108 notes/sec on an iPhone 15). The idea is basically GitHub Copilot or Tabnine, except instead of prompting it with code, you prompt it by playing a few notes on a MIDI piano. The model then continues what you played, entirely on-device. The app is free if anyone wants to try it. Happy to answer questions about the model, training, Core ML, or the many things that didn't work.
AI · 17d ago · simedw.com
Astute▲585Automate your B2B brand going viral, with new media creators
AI · 18d ago · company-app.joinastute.com


Hey HN, Henry from Cactus here! We previously released Cactus Needle, a 14MB agentic LLM for tool call, device use, and structured extraction for phones, wearables, smart homes, small robots and microcontrollers. We got really great feedback here, and have now incorporated the suggestions to release Needle 2. The whole model is a single 14MB binary that runs a full session in 28MB of RAM; 45m parameters at 2bit compression. Needle hits 500 tokens/sec decode speed on a Raspberry Pi 5, sits between 400-1,500 tokens/sec on VR devices like Meta Quest 3S and Apple Vision Pro, and ranges…
AI · 27d ago · cactuscompute.com


Launched alongside, August 2025
the whole month →
- IS
I built the world's most impractical 1000-pixel display and anyone in the world can draw on it. It draws a single pixel at a time and takes 30-60 minutes to complete a single image. Anyone can participate in the project by voting for the next image to be drawn, and submitting images. https://kilopx.com/
Work · 2025 · benholmen.com

- KT
Kitten TTS is an open-source series of tiny and expressive text-to-speech models for on-device applications. We are excited to launch a preview of our smallest model, which is less than 25 MB. This model has 15M parameters. This release supports English text-to-speech applications in eight voices: four male and four female. The model is quantized to int8 + fp16, and it uses onnx for runtime. The model is designed to run literally anywhere eg. raspberry pi, low-end smartphones, wearables, browsers etc. No GPU required! We're releasing this to give early users a sense of the latency and voices…
Dev tools · 2025 · github.com
- IW
I was wondering how I can arrange objects along a spherical helix path, and read some articles on it. I ended up learning about parametric equations again, and make this visualization to document what I learned: https://visualrambling.space/moving-objects-in-3d/ feel free to visit and let me know what you think!
Life & fun · 2025 · visualrambling.space
- TC
For HTML Day 2025 [1], I made a web service that displays the current sky at your approximate location as a CSS gradient. Colours are simulated on-demand using atmospheric absorption and scattering coefficients. Updates every minute, without the use of client-side JavaScript. Source code and additional information is available on GitHub: https://github.com/dnlzro/horizon [1] https://html.energy/html-day/2025/index.html
Dev tools · 2025 · sky.dlazaro.ca