The one-glance read on who they are and how they grow. Each point is verifiable from the receipts above.
An open-source-first LLM evaluation and observability platform for teams testing production LLM apps, backed by YC Winter 2025.
Founder-led and open-source-led, using the free DeepEval library on GitHub as the top of funnel instead of paid acquisition.
~101,600 monthly visits as of the July 2026 snapshot, an oversubscribed $2.2M seed (March 2025), and a 117-point Launch HN break out (February 2025).
Co-founder Jeffrey Ip described the pre-YC grind bluntly: "Prior to YC, we spent a whole year bootstrapping Confident AI out of our bedrooms."
Their own open-source library outranks the company's brand name in search demand, turning free developer adoption into the paid observability upsell rather than any acquisition spend.
The order the channels came online. Sequence is strategy: what they did first, and what they layered on once demand existed.
Estimated demand, the channel split behind it, and the keywords and referrers doing the work. Directional modeling, not audited analytics.
| Keyword | Volume | Weight | CPC |
|---|---|---|---|
| confident ai | 1.9k | $5.23 | |
| llm arena | 120k | $2.54 | |
| deepeval | 13k | $4.67 | |
| llm as a judge | 5.4k | $2.50 | |
| llm as judge | 2.3k | $3.90 |
Confident AI draws an estimated ~101,600 monthly visits, down 20.3% over the last 3 months as of the July 2026 snapshot. The mix is a near-even split between two channels: Search Organic at 43.7% and Direct at 42.9%, meaning as many people type the URL or use a saved bookmark as arrive via a search engine. Referrals add 7.5%, a meaningful third leg likely from the GitHub repo and partner mentions; the rest, social, AI-assistant referrals, and email, totals under 6%.
Their most telling ranked keyword is their own open-source project name: "deepeval" pulls ~13,500 monthly searches, well ahead of "confident ai" at ~1,950, showing developer demand centers on the free tool rather than the brand. "Llm as a judge" (~5,400/mo) reflects the core evaluation methodology they've built content around. Among named competitors, deepeval.com is actually their own open-source project's dedicated domain, not a rival. Langfuse.com and braintrust.dev are the real competitive set: both are funded LLM observability/eval platforms competing for the same developer attention.
The specific pages earning their organic search traffic, and the pattern behind why they rank. Adapt the format, not the topic.
Their top five organic pages are long-form "everything you need to know" and "ultimate guide" style pillar content, not short blog posts, covering LLM evaluation metrics, LLM-as-judge methodology, red teaming, guardrails, and synthetic data generation. The lead asset alone, a comprehensive LLM evaluation metrics guide, pulls 46.9% of all their organic traffic and ranks #2 for "llm evaluation metrics." This concentration pattern, one exhaustive reference asset per core technical topic, is a deliberate bet on owning the definitive answer to a handful of high-intent, moderately-searched terms rather than spreading thin across many shallow posts.
Not traffic share. How much weight the growth system actually puts on each channel, with a one-line read on the role it plays.
Open roles read like a roadmap: the functions they’re staffing show where the company is investing next and what stage it’s at.
For founder-led SaaS the breakdown shifts from ads to traction: where the first users came from, how the founder grows it in the open, and the compounding organic surface.
Confident AI's traction came in two distinct waves: a modest, self-driven Product Hunt debut, then a much bigger break out once the launch was paired with YC credibility on Hacker News.
From August 2023 onward, the founders submitted roughly two dozen technical write-ups to Hacker News, mostly scoring in the single digits to 35 points, none of them launch posts, all of them building recognition ahead of the real moment.
Product Hunt on July 29, 2024 brought 86 upvotes and 20 comments, which the founders' own tweet framed as "First public launch of Confident AI after a year of building."
A Launch HN post on February 20, 2025, timed to going public on YC's Launch YC page, hit 117 points and 27 comments, an order of magnitude bigger than any prior HN submission.
An oversubscribed $2.2M seed closed in March 2025 with YC, Flex Capital, Vermilion Fund, January Capital, and Rebel Fund, announced alongside the founder's account of bootstrapping "out of our bedrooms" for the year before YC.
Growth here runs through co-founder Jeffrey Ip's personal account far more than any brand feed, with milestone numbers consistently outperforming routine updates.
Jeffrey Ip's X account is high-frequency but still small (1,327 posts, 590 followers), meaning reach comes from occasional breakout posts, not steady audience compounding.
His single highest-reach post isn't a milestone announcement at all, it's a question, "How do you beat an open-source alternative that is 100% free?", which pulled 18,800 views versus low hundreds on most other posts.
The $2.2M seed announcement (14,800 views) and the YC Launch YC post (11,000 views) are the next-biggest posts, both tied to external validation events rather than product updates.
Kritin Vongthongsri's own account announced SOC 2 Type 2 compliance, achieved in a 2-week engagement with a compliance vendor, a data point aimed squarely at enterprise buyers evaluating the platform.
The real engine is the open-source DeepEval repo pulling developer search demand, funneled into a small set of pillar guides that convert into product signups.
DeepEval draws roughly 7x the search volume of the "confident ai" brand term itself, meaning the free GitHub library, not company marketing, is doing the top-of-funnel work; DeepTeam extends the same motion into red-teaming.
The lead evaluation-metrics guide covered above carries the bulk of organic traffic, and Jeffrey Ip described the underlying tactic directly in a 2023 post: writing content "containing keywords that are not heavily contested for on Google" rather than chasing generic head terms.
ThoughtWorks featured DeepEval in its technology radar in October 2024, third-party validation of the open-source project that likely reinforces its search authority beyond anything Confident AI publishes itself.
There is no meaningful paid ad presence in the evidence, so this entire engine runs on free distribution, organic search, and launch moments, with none of it backstopped by ad spend.
The channels are not separate. They are one system where each stage feeds the next. Here is the read, then the plays to run tomorrow.
The loop starts with a free GitHub library, not a landing page, and monetization depends on that developer already trusting the tool before a sales conversation ever happens.
DeepEval's GitHub pull and the pillar SEO guides covered above are where cold attention starts, well ahead of any paid channel.
The always-free tier (tracing, testing reports, prompt versioning) steps up to a $9.99 per-user Starter tier once a team needs datasets and alerting, then to custom Team and Enterprise deals gated by RBAC, SSO, and the SOC 2 certification covered above.
Launch and funding milestones renew awareness periodically, but the underlying organic engine is shrinking rather than compounding on its own.
No public revenue figure is disclosed; given ~101,600 monthly visits, a $9.99 self-serve entry price, and a seed raised roughly seven months after first public launch, a directional estimate would put recurring revenue in the low five to low six figure monthly range, not a confirmed number.
The single open role, a Founding GTM hire at $200K to $300K plus equity in San Francisco, signals the company is only now building a formal outbound or enterprise motion on top of what has so far been a self-serve, content-and-launch funnel.
free-to-paid conversion rate, churn, and CAC.
The proofDeepEval, Confident AI's open-source library, pulls roughly 7x the search volume of the company's own brand name, feeding the paid platform without any ad spend.
The adaptationBuild a narrowly scoped free tool, a CLI, library, or template, that solves one acute technical pain in your category, publish it on GitHub or a similar public registry, and let its function name (not your company name) become the searchable term. The free tool earns organic discovery long before anyone hears your brand; the paid upsell comes once teams outgrow the free tier's limits.
Cost: under $500 · Time to signal: months · Works pre-PMF: yes, provided the free tool solves a real, specific pain rather than being a generic demo.
The proofA single "everything you need" pillar guide carries 46.9% of Confident AI's organic traffic, built around a keyword the founder specifically chose for being under-contested.
The adaptationPick one long-tail topic in your category with real search volume but weak existing content, and write one genuinely exhaustive, authoritative post rather than several thin ones. Publish it, track impressions in Search Console, and expand it based on what people actually search for around it, instead of guessing keywords upfront.
Cost: $0 · Time to signal: weeks to months · Works pre-PMF: yes.
The proofRoughly two dozen technical write-ups to Hacker News across two years, none of them launch posts, preceded the 117-point Launch HN that actually broke out.
The adaptationPost genuinely useful technical content to the forums your buyers read, on a recurring basis, months before your real launch moment. Each post should stand alone as useful so it doesn't read as self-promotion; when the actual launch post lands, it comes from a recognized account instead of a first-time poster. This only works if each individual post earns its own read, not as a preamble to a pitch.
Cost: $0 · Time to signal: days per post, months for the compounding effect · Works pre-PMF: yes.
The proofThe Product Hunt debut alone drew a modest 86 upvotes; folding the second launch into "going public on YC's Launch YC page" is what pushed the Hacker News post to 117 points.
The adaptationDon't launch the product alone. Attach it to whatever external validation you actually have, an accelerator, a notable early customer, a funding close, a partnership, and lead with that in the launch post's framing. This is a one-time move tied to an actual credibility event, not something to repeat monthly.
Cost: $0 · Time to signal: days · Works pre-PMF: conditional, only if a real credibility hook exists to attach to.
The proofA funding announcement (14,800 views) and a compliance milestone outperformed routine product updates on a founder account with only 590 followers.
The adaptationWhen you hit a concrete milestone, a first paying customer, a specific percentage improvement a customer saw, a security certification, post the founder's own account with the literal number attached, not a general "big news soon" teaser. Specific numbers travel further than progress updates even on small accounts, because they give other people something concrete to share.
Cost: $0 · Time to signal: days · Works pre-PMF: yes. Not transferable at an earlier stage: the YC-backed funding announcement as a launch hook and the SOC 2 compliance sprint both assume a funding event and revenue base an earlier-stage reader may not yet have. The open-source-to-paid funnel and the sustained content and forum cadence transfer regardless of stage.
Every morning we take one company that is actually growing and break down where its customers come from: the ads still running after a year, the channel doing the real work, and the play you can run this week.
Systemaic · directional intelligence. Traffic, spend, and reach figures are SimilarWeb-style estimates and qualitative reads of public data, not audited numbers. Built on real public receipts.
Launch archives: Hacker News: Launch HN: Confident AI (YC W25) – Open-source evaluation fr · Hacker News: Unit Test LlamaIndex with DeepEval · Hacker News: Tackling the Weaknesses of BertScore · Hacker News: Everything I know about LLM evaluation metrics · Hacker News: YC helped us raise our seed round in 5 days · Hacker News: Best Practices for Unit Testing RAG Systems in Prod
Traffic, spend, and revenue figures are estimates as noted in the report; the links above are the primary public artifacts.