HIGH SALIENCE / RESEARCH / APPLICANT TRACKING TEARDOWN

TEARDOWN 09 · PUBLISHED SEPTEMBER 26, 2026 · DATASET INCLUDED

100 Questions, Ninth Category: Applicant Tracking, Where the Leaders Rank Least and a Comparison Site Nobody Ranks Is Cited in 13 Answers

Ninth category, same method: 100 applicant tracking system queries, ChatGPT with web search on, every answer coded under the rulebook the earlier teardowns used, Google's top ten as a control. Four brands lead, the answers split cleanly by buyer type, and this is the category where Google and ChatGPT agree least: 84 percent of recommendations went to brands whose own site did not rank for the question, and one comparison site that ranked for none of them was cited in 13 answers.

01Method

Same query structure, same coding, one new category.

Queries: 100 applicant tracking software queries built from the same templates as the earlier teardowns: 30 category, 30 comparison, 20 alternative and 20 recommendation. The full list is in the dataset.

Surface: ChatGPT with web search forced on, logged out, United States, English, via the DataForSEO scraper, one run per query, collected September 26, 2026.

Control: Google's top ten organic results for the same queries in the same window, via the DataForSEO SERP API at depth 20 and truncated to the first ten organic results.

Coding: every brand named was coded as recommended (selected for a stated case), listed (in a table or list without being chosen), passing, or anchor (the brand being replaced in an "alternatives" query), against a 51-brand dictionary fixed before coding (see the collection note), under codebook v1.5. Cited URLs were deduplicated to their domain and classed as first-party or third-party. Five answers drawn with a fixed seed were hand-checked: 24 coded mentions, 20 agreed with the reader. All four disagreements are in one answer: three conditional picks separated by bold markup were coded listed, and BambooHR, named only in a question back to the buyer, was coded recommended. They are left in the data as coded and documented in coder-notes.md.

Collection note: the ChatGPT answers and the Google controls were collected on September 26, 2026 through the DataForSEO standard queue, with raw responses saved exactly as returned. Greenhouse is matched on both greenhouse.com and its earlier domain greenhouse.io, which it still uses for job boards; the coder was extended to accept more than one domain per brand, and re-coding two earlier categories with the extended coder reproduced their published files exactly. Fifteen aliases that are also ordinary words (Greenhouse, Lever, Workable, Fountain and others) match only when capitalized.

One run per query, so this is a teardown, not the benchmark. Frequencies describe this window only. Absence means not observed in this sample, never zero visibility.

02Findings

Four leaders, clean lanes, and the widest gap from Google yet.

Bar chart of 12 applicant tracking software brands showing how many of 100 ChatGPT answers each appears in versus how many recommend it. Greenhouse 56 and 34, Workable 53 and 32, Lever 43 and 22, Ashby 38 and 23, Breezy HR 28 and 13, JazzHR 27 and 11, BambooHR 23 and 16, Workday Recruiting 23 and 10, iCIMS 16 and 10, Zoho Recruit 14 and 10, SmartRecruiters 14 and 8, Teamtailor 13 and 5.
BrandAppears inRecommended inRecommended share of appearances
Greenhouse563461%
Workable533260%
Lever432251%
Ashby382361%
Breezy HR281346%
JazzHR271141%
BambooHR231670%
Workday Recruiting231043%
iCIMS161062%
Zoho Recruit141071%
SmartRecruiters14857%
Teamtailor13538%

Each linked brand has its own page: how ChatGPT recommends every brand in this dataset, with positioning labels, head-to-head results and sources.

Out of 100 answers, codebook v1.5. "Recommended" means the answer selected the brand for a stated case. "Appears" adds brands listed in a table or bullet list without being chosen. The brand being replaced in an "alternatives" query is excluded from both columns.

Four brands lead

Greenhouse is named in 56 answers and recommended in 34. Workable is named in 53 and recommended in 32, Lever 43 and 22, Ashby 38 and 23. On the 20 direct advice questions, 18 produced a pick from the dictionary: Greenhouse was among them in 11, Workable and Ashby in 10 each, Lever in 7. Breezy HR and JazzHR are named often, 28 and 27 times, and chosen in 13 and 11.

The lanes are clean

ChatGPT sorts this category by buyer type. Engineering hiring goes to Ashby, Greenhouse and Lever. Enterprise questions go to iCIMS, Workday Recruiting, Greenhouse and SmartRecruiters. Hourly and restaurant hiring goes to Fountain and Paradox, with iCIMS and UKG added when the question says at scale. Staffing and recruiting agencies get Bullhorn, Loxo, Recruit CRM and Manatal, vendors built for agencies. Free and open source alternatives to Greenhouse get OpenCATS.

Caution, but less in head-to-heads

Twenty-four answers recommended no dictionary brand, behind only payroll (28). The head-to-heads were more decisive than in HR: 22 of 30 recommended every brand named, and five recommended none, including Lever versus Workable, Workable versus Ashby, and Greenhouse versus Lever versus Workable. One comparison answer, Lever versus Teamtailor, warned that much of the side-by-side material it found was published by one of the two vendors.

The widest gap from Google

Of the 242 recommended brand mentions, 203 were for brands whose own website does not rank in Google's top ten organic results for that query, 84 percent, the highest of any category so far. Google's top ten for these queries is 87 percent third-party pages, also the highest.

A comparison site nobody ranks

Across 100 answers there were 311 citation events to 110 domains. Vendors' own sites took 139 of them, 45 percent. lever.co is the most-cited domain, in 26 answers, 17 of them to questions that never mention Lever. The most-cited third party is Capterra (15), followed by yardstick.team, cited in 13 answers across nine different pages: ATS comparisons, pricing and "alternatives" pages. yardstick.team ranked in Google's top ten for none of those 13 questions. It is the same pattern as the help desk and HR teardowns: a site that publishes a systematic set of comparison pages becomes part of the model's reading list without ranking. Reddit and Wikipedia have zero citations, for the ninth teardown running.

Fifty-nine of the 311 citation events involved a domain that also sat in Google's top ten for that query, 19 percent, the lowest overlap in the series.

039 categories, side by side

Same rulebook, every category so far.

Measure (codebook v1.5)Project management, Sept 8CRM, Sept 17Email marketing, Sept 17Help desk, Sept 17Accounting, Sept 17Payment processing, Sept 17Payroll, Sept 21HR software, Sept 26Applicant tracking, Sept 26
Recommendations per category answer (average)4.83.33.33.93.12.32.03.22.7
Answers with no recommended dictionary brand5 of 10013 of 1007 of 1008 of 1009 of 10014 of 10028 of 10023 of 10024 of 100
Most-recommended brand: appears / recommendedAsana 76 / 68HubSpot 77 / 62Mailchimp 61 / 40Zendesk 74 / 54QuickBooks 75 / 58Stripe 73 / 53Gusto 76 / 47Gusto 51 / 39Greenhouse 56 / 34
Head-to-head queries recommending every named brand29 of 3026 of 3030 of 3028 of 3027 of 3021 of 3018 of 3019 of 3022 of 30
Recommendation-intent queries with a dictionary pick18 of 2016 of 2017 of 2018 of 2019 of 2018 of 2014 of 2017 of 2018 of 20
Recommended mentions where the brand does not rank in Google's top 10305 of 368 (83%)187 of 287 (65%)216 of 309 (70%)255 of 336 (76%)136 of 268 (51%)160 of 210 (76%)87 of 191 (46%)192 of 259 (74%)203 of 242 (84%)
Google top 10 that is third-party pages746 of 1,000 (75%)710 of 1,000 (71%)720 of 1,000 (72%)714 of 1,000 (71%)743 of 1,000 (74%)787 of 1,000 (79%)653 of 1,000 (65%)820 of 1,000 (82%)871 of 1,000 (87%)
Citation events to vendors' own sites216 of 362 (60%)136 of 320 (42%)193 of 367 (53%)152 of 345 (44%)125 of 302 (41%)161 of 301 (53%)192 of 314 (61%)167 of 325 (51%)139 of 311 (45%)
Answers citing only vendor pages50 of 10040 of 10037 of 10027 of 10034 of 10047 of 10049 of 10034 of 10030 of 100
Share of citations in the 10 most-cited domains54%40%42%45%55%58%55%46%44%
Citation events whose domain is in Google's top 1083 of 362 (23%)83 of 320 (26%)97 of 367 (26%)80 of 345 (23%)99 of 302 (33%)98 of 301 (33%)134 of 314 (43%)85 of 325 (26%)59 of 311 (19%)
Reddit and Wikipedia citations000000000

All columns are coded under codebook v1.5; the project management column is the September 8 dataset recoded under it. Each teardown is one run per query in its own window, so differences between columns mix category with collection date.

04What this changes

Four things, in order.

Ranking is the weakest predictor yet. In this category 84 percent of ChatGPT's recommendations went to brands that did not rank for the question. Google visibility and AI visibility have to be measured, and worked on, separately.

Comparison libraries get read. A site with nine comparison and alternatives pages was cited in 13 answers without ranking for any of them. If a library like that covers your category, it is part of your AI search footprint whether you know it or not.

Own your lane in plain words. ChatGPT sends engineering hiring, enterprise, hourly and agency buyers to different vendors. The vendors that win a lane are the ones whose pages say which buyer they are for.

Your own site is read for your competitors' questions. lever.co was cited in 17 answers to questions that do not mention Lever. Vendor comparison and guide pages are some of the most-read material in the category.

05Limitations

What this teardown cannot tell you.

One run per query means answer variance is unmeasured; Benchmark 01 runs each query three times across three surfaces. The teardowns are collected on different days, so differences between categories mix the category with the date. The brand dictionary covers the general applicant tracking software market and deliberately excludes vertical tools, so their appearances are described in prose and not counted. The rule-based coder has one known failure: a brand named as a contrast on the same line as a selecting verb is coded as recommended. Sentiment coding is rule-based and not reported. No vendor-tool cross-check was read for this category. Google's control counts a brand as ranking only when its own domain is in the top ten; a listicle that features the brand does not count. One run per question, one day, United States, English, logged out. Answers vary between runs, so these are frequencies for this sample. The hand check agreement (20 of 24) is lower than in some earlier teardowns, and the known coding errors are listed in coder-notes.md rather than corrected. Recruiterflow and JobAdder appeared in agency answers but are not in the dictionary, so agency results undercount them.

06Dataset

Check it, don't believe it.

Every number above can be recomputed from these files. CC BY 4.0: use them, cite the page.

  • queries.csv: the 100 queries with intent labels.
  • brands.csv: the 51-brand dictionary with aliases and canonical domains.
  • mentions.csv: 476 coded brand mentions with position, type and whether the brand's domain was in Google's top ten.
  • citations.csv: 311 citation events with domain class and Google overlap.
  • observations.csv: one row per query with brand, recommendation, citation and control counts.
  • coder-notes.md: hand-check results, dictionary notes and known coding disagreements, left uncorrected in the data
  • codebook.md: the rulebook, with dated amendments through v1.5.

The earlier teardowns: Teardown 01, project management, Teardown 02, crm, Teardown 03, email marketing, Teardown 04, help desk, Teardown 05, accounting, Teardown 06, payment processing, Teardown 07, payroll, Teardown 08, hr software.

Nine categories, and the reading list is still the story.

In applicant tracking the model's picks track vendor pages and a comparison library that does not rank. In HR they track one comparison page. In payroll, one vendor's website. Finding out what the model is reading for your category is the first thing a Category Salience Brief does, with a query set you approve first.