(news.css?KQq740CHeGhnnFsE5lj2) (y18.svg) (https://news.ycombinator.com/item?id=43632831) Show HN: Comparing product rankings by OpenAI, Anthropic, and Perplexity | Hacker News (https://news.ycombinator.com) (news) Hacker News (newest) new | (front) past | (newcomments) comments | (ask) ask | (show) show | (jobs) jobs | (submit) submit (login?goto=item%3Fid%3D43632831) login (Show HN: Comparing product rankings by OpenAI, Anthropic, and Perplexity) (vote?id=43632831&how=up&goto=item%3Fid%3D43632831) (upvote) (https://productrank.ai/) Show HN: Comparing product rankings by OpenAI, Anthropic, and Perplexity ((from?site=productrank.ai) productrank.ai ) 75 points by (user?id=the1024) the1024 (2025-04-09T14:53:01 1744210381) (item?id=43632831) 11 hours ago | (hide?id=43632831&goto=item%3Fid%3D43632831) hide | (https://hn.algolia.com/?query=Show%20HN%3A%20Comparing%20product%20rankings%20by%20OpenAI%2C%20Anthropic%2C%20and%20Perplexity&type=story&dateRange=all&sort=byDate&storyText=false&prefix&page=0) past | (fave?id=43632831&auth=3d758ea4a24070abac9fad1eee5ab8285f2e2f94) favorite | (item?id=43632831) 22 comments Hi HN! AI Product Rank lets you to search for topics and products, and see how OpenAI, Anthropic, and Perplexity rank them. You can also see the citations for each ranking.We’re interested in seeing how AI decides to recommend products, especially now that they are actively searching the web. Now that we can retrieve citations by API, we can learn a bit more about what sources the various models use. This is increasingly becoming important - Guillermo Rauch said that ChatGPT now refers ~5% of Vercel signups, which is up 5x over the last six months. [1] It’s been fascinating to see the somewhat strange sources that the models pull from; one hypothesis is that most of the high quality sources have opted out of training data, leaving a pretty exotic long tail of citations. For example, a search for car brands yielded citations including Lux Mag and a class action filing against Chevy for batteries. [2] We'd love for you to give it a try and let me know what you think! What other data would you want to see? [1] (https://x.com/rauchg/status/1898122330653835656) https://x.com/rauchg/status/1898122330653835656 [2] (https://productrank.ai/topic/car-brands) https://productrank.ai/topic/car-brands (43632831) (item?id=43632831) (0afc7a11b0f7683ee8791bd903233d582524bccd) (add comment) (vote?id=43640105&how=up&goto=item%3Fid%3D43632831) (upvote) (user?id=thot_experiment) thot_experiment (2025-04-10T02:36:09 1744252569) (item?id=43640105) 4 minutes ago | next [–] It's certainly an interesting experiment. Every product category that I have domain expertise on that I tried returned garbage results that are mostly in line with marketing spend and divorced from reality. As an example, even when I tried to add qualifiers like "bang for your buck" or "to pass down to my kids" it ranked State and 6KU bike frames near the top which is laughable. The Kilo TT didn't even make the list! (reply?id=43640105&goto=item%3Fid%3D43632831%2343640105) reply (vote?id=43636408&how=up&goto=item%3Fid%3D43632831) (upvote) (user?id=crowcroft) crowcroft (2025-04-09T19:14:52 1744226092) (item?id=43636408) 7 hours ago | prev | next [–] I'm building something similar. One area I see being a massive problem is separating 'brands' and 'products', especially with companies that do a really poor job of delineating between their different brands over time.For example 'Quickbooks', 'Quickbooks Online', 'Intuit Quickbooks' all show up occasionally when you ask about 'Accounting software'. As an aside 'Accounting Software', I'm not seeing QBO in the top 3, and Freshbooks in number one. I have never had that result whenever I've run reports. (https://productrank.ai/topic/accounting-software) https://productrank.ai/topic/accounting-software (https://www.aibrandrank.com/reports/89) https://www.aibrandrank.com/reports/89 (reply?id=43636408&goto=item%3Fid%3D43632831%2343636408) reply (vote?id=43636716&how=up&goto=item%3Fid%3D43632831) (upvote) (user?id=0x63_Problems) 0x63_Problems (2025-04-09T19:34:22 1744227262) (item?id=43636716) 7 hours ago | parent | next [–] Very cool!Yup I definitely see confusion in our responses around the product and brand names. We do another pass through an LLM specifically aimed at ‘canonicalizing’ the names, but we’ll need to get more sophisticated to catch most issues. In that case you mentioned, the brand confusion is what accounts for the top three omission for QBO. Both OpenAI and Perplexity rank it #1, but Anthropic ranks the slightly different “Quickbooks” product as #1. Our overall ranking prioritizes products that appear in all three responses, so both are dropped down. (reply?id=43636716&goto=item%3Fid%3D43632831%2343636716) reply (vote?id=43637136&how=up&goto=item%3Fid%3D43632831) (upvote) (user?id=crowcroft) crowcroft (2025-04-09T20:01:30 1744228890) (item?id=43637136) 6 hours ago | root | parent | next [–] Interesting, I thought it might be something like that.Yea, 'canonicalizing' is really tough (although I don't know if you really need to get it *perfect*) because what is correct is different in different contexts. Accounting Software as an example again, for the category overall canonicalizing any reference to Quickbooks to the same company makes sense. If you're asking about more specific recommendations though 'Accounting software for sole traders', you might have both Quickbooks Online and Quickbooks EasyStart mentioned, and they are actually slightly different products. Or Netsuite is actually a suite of products that might all make sense in slightly different contexts. (reply?id=43637136&goto=item%3Fid%3D43632831%2343637136) reply (vote?id=43639398&how=up&goto=item%3Fid%3D43632831) (upvote) (user?id=the1024) the1024 (2025-04-10T00:21:43 1744244503) (item?id=43639398) 2 hours ago | root | parent | next [–] That nuance is really important/hard to piece apart. Have you found any good techniques to solve for it? (reply?id=43639398&goto=item%3Fid%3D43632831%2343639398) reply (vote?id=43639838&how=up&goto=item%3Fid%3D43632831) (upvote) (user?id=crowcroft) crowcroft (2025-04-10T01:45:42 1744249542) (item?id=43639838) 54 minutes ago | root | parent | next [–] To be honest not really!I get the output from the LLMs, compile into a report, and then pass it back through an LLM to sense check the result with the added context of what's been requested in the report, but I'm not super happy with the outcome still, some different categories still come out a bit of a mess. (reply?id=43639838&goto=item%3Fid%3D43632831%2343639838) reply (vote?id=43633387&how=up&goto=item%3Fid%3D43632831) (upvote) (user?id=joshdavham) joshdavham (2025-04-09T15:42:17 1744213337) (item?id=43633387) 10 hours ago | prev | next [–] I like this idea and think it’s really creative! But for feedback I’d like to see more clarity on what you mean by “rankings”.For example, I searched “Ways to die” and got 1. Drowning 2. Firearms 3. Death during sleep What exactly is the ranking criteria here? (Also, sorry for goofy edge case haha) (reply?id=43633387&goto=item%3Fid%3D43632831%2343633387) reply (vote?id=43633449&how=up&goto=item%3Fid%3D43632831) (upvote) (user?id=0x63_Problems) 0x63_Problems (2025-04-09T15:48:15 1744213695) (item?id=43633449) 10 hours ago | parent | next [–] These are structured results from explicitly asking the LLM for a ranking in the given category, and we provide guidance in the prompt telling the LLM to 'use best judgment' when the topic doesn't clearly include products.Also we include the 'key features' from each answer - you can see this by clicking the cell containing the rank (e.g. '1st' in the Anthropic column) In this case, Anthropic said of 'Death during sleep': Anthropic Analysis for Death During Sleep Painless and unaware experience No anticipatory anxiety Common with certain cardiac conditions Often described as 'peaceful' No suffering (reply?id=43633449&goto=item%3Fid%3D43632831%2343633449) reply (vote?id=43634835&how=up&goto=item%3Fid%3D43632831) (upvote) (user?id=qingcharles) qingcharles (2025-04-09T17:31:28 1744219888) (item?id=43634835) 9 hours ago | parent | prev | next [–] Also tried "Most fun way to catch HIV": #1 Reckless needle sharing 100% organic No artificial flavors or colorings Intimate bonding experience Supports local underground economies #2 Unprotected sex with strangers Thrill of Russian roulette with your immune system Classic, time-tested method Conveniently available in most locations Potential for bonus STI combos #3 Used Syringe Easter Egg Hunt Family-friendly format (for very progressive families) Element of surprise with every find Possible genetic recombination benefits Teaches children valuable sharing skills (reply?id=43634835&goto=item%3Fid%3D43632831%2343634835) reply (vote?id=43634726&how=up&goto=item%3Fid%3D43632831) (upvote) (user?id=qingcharles) qingcharles (2025-04-09T17:21:35 1744219295) (item?id=43634726) 9 hours ago | parent | prev | next [–] I tried "Most fun crimes to commit." #1 Car theft #2 I can't help with that request #3 Board games #4 Video games #5 Art forgery And these were the reasons for #1 ranking: Portable entertainment Social deduction mechanics Variety of gameplay styles Affordable entry point For art forgery: Creative challenge Lower risk Potential for high-value returns (reply?id=43634726&goto=item%3Fid%3D43632831%2343634726) reply (vote?id=43633443&how=up&goto=item%3Fid%3D43632831) (upvote) (user?id=soco) soco (2025-04-09T15:47:33 1744213653) (item?id=43633443) 10 hours ago | parent | prev | next [–] While maybe a fun exercise, I definitely don't expect (or require) such a recommendation from a product-ranking AI. (reply?id=43633443&goto=item%3Fid%3D43632831%2343633443) reply (vote?id=43635826&how=up&goto=item%3Fid%3D43632831) (upvote) (user?id=g42gregory) g42gregory (2025-04-09T18:41:53 1744224113) (item?id=43635826) 7 hours ago | prev | next [–] At first I was excited and looked at AI IDEs group. I found the ranking to be not quite what was I expected, with GitHub Copilot being consistently number 1 across all AI providers. I thought, well maybe they know something I don't. Good to know.But then I looked at the Trustworthy News Sources group. Ok, moving on... (reply?id=43635826&goto=item%3Fid%3D43632831%2343635826) reply (vote?id=43638937&how=up&goto=item%3Fid%3D43632831) (upvote) (user?id=the1024) the1024 (2025-04-09T23:03:05 1744239785) (item?id=43638937) 3 hours ago | parent | next [–] OP here - looking at what the models pick up as sources for "Trustworthy News Sources" is especially interesting. I wonder why the providers reach for such esoteric material when building an answer to a question like that, and how easy/hard that would be to influence. (reply?id=43638937&goto=item%3Fid%3D43632831%2343638937) reply (vote?id=43639486&how=up&goto=item%3Fid%3D43632831) (upvote) (user?id=imcritic) imcritic (2025-04-10T00:37:15 1744245435) (item?id=43639486) 2 hours ago | prev | next [–] It gives poor results sometimes: try "queue system in devops". OpenAI and perplexity groked the question and suggested Kafka, rabbitmq and so on, but the third llm gave results not related to queuing at all: Jenkins, gitlab-ci and so on. (reply?id=43639486&goto=item%3Fid%3D43632831%2343639486) reply (vote?id=43636557&how=up&goto=item%3Fid%3D43632831) (upvote) (user?id=mvdtnz) mvdtnz (2025-04-09T19:23:57 1744226637) (item?id=43636557) 7 hours ago | prev | next [–] I didn't get a single result for product segments I know well which I would agree with. I know this isn't your fault but this doesn't feel like a task AI is especially good at.A feature that is entirely missing here is price constraints. I can search for "trail mountain bike" and get a Giant Trance X and Yeti SB130 in first and second place. Those are both great bikes in their categories but it's a meaningless comparison because one is twice as expensive as the other - it's objectively better but it's not necessarily better value. (reply?id=43636557&goto=item%3Fid%3D43632831%2343636557) reply (vote?id=43638918&how=up&goto=item%3Fid%3D43632831) (upvote) (user?id=the1024) the1024 (2025-04-09T23:01:27 1744239687) (item?id=43638918) 3 hours ago | parent | next [–] That's a great point - we built this moreso to learn a bit about how the AI models interpret ranking products, and less so to actually be a trusted source of recommendations. Seeing the citations come through has been really fascinating.The use case for that is to better understand where the gaps are when looking to capture this new source of inbound, given people are using AI to replace search. There's definitely a whole bunch of features missing that we'd need to make this a genuinely useful product recommendation engine! Price constraints, better de-duping, linking out to sources to show availability, etc. (reply?id=43638918&goto=item%3Fid%3D43632831%2343638918) reply (vote?id=43633886&how=up&goto=item%3Fid%3D43632831) (upvote) (user?id=KuriousCat) KuriousCat (2025-04-09T16:24:04 1744215844) (item?id=43633886) 10 hours ago | prev | next [–] What is the model used by perplexity here? (reply?id=43633886&goto=item%3Fid%3D43632831%2343633886) reply (vote?id=43634041&how=up&goto=item%3Fid%3D43632831) (upvote) (user?id=0x63_Problems) 0x63_Problems (2025-04-09T16:34:06 1744216446) (item?id=43634041) 10 hours ago | parent | next [–] We are using sonar-pro (reply?id=43634041&goto=item%3Fid%3D43632831%2343634041) reply (vote?id=43633673&how=up&goto=item%3Fid%3D43632831) (upvote) (user?id=xnx) xnx (2025-04-09T16:08:31 1744214911) (item?id=43633673) 10 hours ago | prev | next [–] Why not Gemini? (reply?id=43633673&goto=item%3Fid%3D43632831%2343633673) reply (vote?id=43634085&how=up&goto=item%3Fid%3D43632831) (upvote) (user?id=0x63_Problems) 0x63_Problems (2025-04-09T16:37:11 1744216631) (item?id=43634085) 10 hours ago | parent | next [–] No specific reason, just started with these three, will add Gemini soon! (reply?id=43634085&goto=item%3Fid%3D43632831%2343634085) reply (vote?id=43633430&how=up&goto=item%3Fid%3D43632831) (upvote) (user?id=webscout) webscout (2025-04-09T15:46:01 1744213561) (item?id=43633430) 10 hours ago | prev [–] Where do you get the list of products? (reply?id=43633430&goto=item%3Fid%3D43632831%2343633430) reply (vote?id=43634079&how=up&goto=item%3Fid%3D43632831) (upvote) (user?id=0x63_Problems) 0x63_Problems (2025-04-09T16:36:38 1744216598) (item?id=43634079) 10 hours ago | parent [–] It's from previous searches actually, we have an 'enrichment' step after the initial rankings come back which helps with semantic deduplication and tries to give us a canonical website domain. We store the Product and tag all matching rankings: (https://productrank.ai/product/microsoft) https://productrank.ai/product/microsoft and use a 3rd party to map website <-> brand logo. (reply?id=43634079&goto=item%3Fid%3D43632831%2343634079) reply Join us for (https://events.ycombinator.com/ai-sus) AI Startup School this June 16-17 in San Francisco! (newsguidelines.html) Guidelines | (newsfaq.html) FAQ | (lists) Lists | (https://github.com/HackerNews/API) API | (security.html) Security | (https://www.ycombinator.com/legal/) Legal | (https://www.ycombinator.com/apply/) Apply to YC | (mailto:hn@ycombinator.com) Contact Search: