This product was not featured by Product Hunt yet. It will not be visible on their landing page and won't be ranked (cannot win product of the day regardless of upvotes).
Octen
We stayed quiet and built the FASTEST AI search on Earth
Octen is a search API built for AI agents, not humans. Instead of one query at a time, it decomposes a question into sub-queries, fires them all at once, and returns results in 62ms (P50), several times faster than other providers. Pricing is $1 per 1,000 calls, well below typical rates of $4 to $9. A single account handles over 1,000,000 queries per second, with new content indexed within 5 minutes. Our embedding models rank #1 on RTEB and MMEB-v2.
Most search APIs do one query at a time, like a human typing in a box. Octen's a different architecture — agents can run many queries at once. Not just faster, built for it.
65ms is the part that really caught my attention. In something like a voice agent or a multi-agent workflow where one agent is constantly handing off to another, latency compounds very quickly. Keeping each call this fast could make the difference between a demo that looks good and an app that actually feels usable. Definitely trying this with my current app! Congrats on the launch!!
I believe every model will ultimately pursue efficiency. This is essential for user experience, and Octen is on the right track.
Octen is exactly the infrastructure AI agents have been waiting for. Instead of forcing the human "one query at a time" habit onto machines, it works the way agents actually think—breaking a question into sub-queries and running them in parallel.
The numbers speak for themselves: 62ms P50 latency (several times faster than anything I've used), $1 per 1,000 calls (vs. the usual $4–$9), 1M+ queries per second on a single account, and 5-minute index freshness. Add embedding models ranked #1 on both RTEB and MMEB-v2, and this stops being "just another search API"—it's a real foundation to build on.
What stands out is the combination of speed, cost, and breadth. Broad research usually means fanning out across many queries and sources, which gets slow and expensive quickly. Octen seems purpose-built for that workload, and the $1/1k pricing makes it practical to run at agent scale.
As an R&D engineer at Octen, I’m absolutely thrilled to see our product finally go live! I'd love for you all to test it out and drop some genuine feedback.
As algorithm engineers at Octen, we are committed to delivering the best search results to you as quickly as possible.
Try using Octen in your product, you'll be amazed!
Congrats! Octen nails the three things that actually matter in production: cost, latency, and stability. Sub-100ms search keeps real-time copilots and multi-search agent loops feeling instant, the pricing makes high-volume usage genuinely viable instead of something you have to ration, and latency stays flat at 1M+ QPS — no cliff under load, predictable P99 you can build an SLA on. Fast, affordable, and rock-solid all at once is a rare combo. Excited to see where this goes!
As a product manager on the team at Octen, I'm so excited to see this go live.
I've been applying our search to AI agent pipelines — where one task fans out into dozens of sub-queries. Latency doesn't add up there, it multiplies. Getting each call under 100ms is what makes the whole thing feel alive instead of frozen.
Give it a try — would love to hear what you build with it.
62ms is truly an impressive figure—an excellent product and an excellent team.
Why Octen?
Speed: Returns results in ~62ms (P50) by decomposing and parallelizing queries.
Reliability: We offer a contractual SLA because downtime isn't an option.
Developer First: Designed to integrate seamlessly into your agent workflows.
Congrats on the launch! The agent-first framing makes a lot of sense — most search APIs really are just human SERPs with a JSON wrapper, and they fall over the moment you fan out queries. The $1/1k pricing is aggressive enough that I'll definitely benchmark it against what we're using now. Appreciate that the benchmarks are public and reproducible — that's rarer than it should be.
I'm on the team here 🙌 Been heads-down on this for months. The 62ms P50 is the number I'm proudest of — and just as important, every benchmark behind it is public and reproducible. Like Kuan said, we stayed quiet until the numbers were real. If you run it against what you're using today, please tell us where it falls short; that's exactly the input we want. Congrats to the whole crew, and thanks Kuan for leading it.
Currently as an engineer building at Octen, I excited for our launch. I have been applying our search capabilities to wearables and AR, and that's where latency actually bites. When a device is reacting to what you're looking at, search that takes a second breaks the whole thing. Sub-100ms is what makes it feel like it just knows.
Honestly the sub-query thing is kind of clever, I tested it on a research prompt and it came back faster than I expected. Price point looks solid too.
Hey Product Hunt 👋
I'm Kuan, founder of Octen. Before this, I led AI search product at Alibaba Cloud and built Baidu's enterprise search platform, so I've spent a decade dealing with search infrastructure that had to survive real scale.
When I looked at the search APIs being handed to AI agents today, I was surprised. Most are the same human-search stack wrapped in an endpoint, and many start to break once you push tens of queries per second. But an agent doesn't search the way a person does. It wants to ask a question from many angles at once and get everything back together.
So we rebuilt search from scratch for that use case. Octen decomposes a query into sub-queries, fires them concurrently, and returns results in 62ms (P50) and 68ms (P90) on SealQA Hard, several times faster than other providers in this space, at $1 per 1,000 calls instead of the usual $4 to $9.
We spent the past several months building this quietly because we didn't want to make noise before the numbers were real. All benchmarks are public and reproducible (link in the post). Would love for you to run it against whatever you're using today and tell us what you find.
Happy to answer any questions here.
About Octen on Product Hunt
“We stayed quiet and built the FASTEST AI search on Earth”
Octen was submitted on Product Hunt and earned 33 upvotes and 18 comments, placing #27 on the daily leaderboard. Octen is a search API built for AI agents, not humans. Instead of one query at a time, it decomposes a question into sub-queries, fires them all at once, and returns results in 62ms (P50), several times faster than other providers. Pricing is $1 per 1,000 calls, well below typical rates of $4 to $9. A single account handles over 1,000,000 queries per second, with new content indexed within 5 minutes. Our embedding models rank #1 on RTEB and MMEB-v2.
Octen was featured in API (98.4k followers), Developer Tools (516.4k followers) and Search (18.1k followers) on Product Hunt. Together, these topics include over 91.5k products, making this a competitive space to launch in.
Who hunted Octen?
Octen was hunted by Dev Chandra. A “hunter” on Product Hunt is the community member who submits a product to the platform — uploading the images, the link, and tagging the makers behind it. Hunters typically write the first comment explaining why a product is worth attention, and their followers are notified the moment they post. Around 79% of featured launches on Product Hunt are self-hunted by their makers, but a well-known hunter still acts as a signal of quality to the rest of the community. See the full all-time top hunters leaderboard to discover who is shaping the Product Hunt ecosystem.
Want to see how Octen stacked up against nearby launches in real time? Check out the live launch dashboard for upvote speed charts, proximity comparisons, and more analytics.
Most search APIs do one query at a time, like a human typing in a box. Octen's a different architecture — agents can run many queries at once. Not just faster, built for it.