Open weight models are particularly fast in the Neon AI Gateway thanks to optimizations like prompt caching. Free to run during beta
Blog

What we’re shipping.
What you’re building.Postgres