15 AI models · 10 answers per model · ranked by score
ordered by brand score: 100 pts for rank 1 → 20 for rank 5 · no mention = 0
Each bar covers the middle half of one model’s answers (Q1–Q3); the line inside it is that model’s median rank for the times it recommended the brand. The badge under a brand says how far apart the models are on it, and “first choice” scores rank 1 far above the rest — a brand can place well overall and still rarely lead an answer.
Swipe left for the full report
Median rank each model gave each brand.
Scroll the table sideways to see every model →
| gpt-6.1-sol | gpt-6-luna | gpt-6-sol | grok-4.7 | gpt-6-astra | gpt-5.6-luna | gpt-5.6-sol | gpt-5.6-terra | gpt-5.5 | gpt-5.4 | gpt-5.2 | gpt-5 | o3 | gpt-4.1 | o1 | |
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Jotform | 1.0 | 1.0 | 1.0 | 1.0 | 1.0 | 1.0 | 1.0 | 1.0 | 1.0 | 1.0 | 1.0 | 1.0 | 1.0 | 3.0 | 2.0 |
| Typeform | 4.0 | 4.0 | 4.0 | 2.5 | 4.0 | 3.0 | 3.0 | 2.0 | 2.0 | 2.0 | 2.0 | 2.0 | 2.0 | 2.0 | 1.0 |
| Google Forms | 4.0 | 2.0 | 2.0 | 2.5 | 5.0 | 2.5 | 5.0 | 3.0 | 3.0 | 3.0 | 3.0 | 4.5 | 3.0 | 2.0 | 3.0 |
| Tally | 2.0 | 3.0 | 3.0 | 4.0 | 2.0 | — | 2.0 | 2.0 | 4.0 | — | 4.0 | 3.0 | — | — | — |
| Wufoo | — | — | — | 5.0 | 5.0 | 5.0 | 5.0 | 5.0 | 4.5 | 4.0 | 5.0 | 4.5 | 5.0 | 4.0 | 4.0 |
| Cognito Forms | 3.0 | — | — | 4.0 | 3.0 | 4.0 | 3.0 | 4.0 | 3.0 | — | — | 3.5 | 4.0 | — | 5.0 |
| Zoho | 5.0 | 3.0 | 4.0 | — | 5.0 | 4.0 | 4.0 | 4.0 | 4.0 | — | — | 4.0 | — | 5.0 | — |
| Formstack | — | 5.0 | 5.0 | 5.0 | — | 4.5 | — | 5.0 | 5.0 | 5.0 | 4.0 | 5.0 | 4.0 | 5.0 | 5.0 |
| Paperform | 5.0 | 5.0 | 5.0 | — | 5.0 | — | 4.0 | — | — | 5.0 | — | 4.0 | — | — | 4.5 |
| Microsoft | — | 5.0 | 5.0 | 5.0 | — | 5.0 | — | — | — | — | 5.0 | — | — | — | — |
| WPForms | 5.0 | — | — | — | — | — | — | — | — | — | — | — | — | 4.0 | 4.0 |
| HubSpot | — | 3.0 | — | — | — | 4.0 | — | — | — | — | — | — | — | — | — |
| SurveyMonkey | — | 5.0 | — | — | — | — | — | — | — | — | 5.0 | — | — | — | — |
| Fillout | — | — | — | — | 3.0 | — | — | — | — | — | — | — | — | — | — |
| 123FormBuilder | — | — | — | — | — | — | — | — | — | 5.0 | — | — | — | — | — |
| Formsite | — | — | — | — | — | — | — | — | — | — | — | — | 5.0 | — | — |
A dash means that model never named the brand in any of its runs. Deeper blue is better: rank 1 is the brand a model would recommend first.
3 of 16 brands split the panel. Each mark is one model's score for the brand, on the same 0–100 scale as the ranking.
gpt-4.1 0 → gpt-6-astra 80 · 5 of 15 never named it
gpt-5.6-sol 18 → gpt-4.1 84
gpt-6-sol 32 → o1 92
gpt-4.1 0 → gpt-6.1-sol 54 · 5 of 15 never named it
gpt-6-luna 0 → gpt-5.4 48 · 3 of 15 never named it
gpt-5.2 0 → gpt-6-luna 44 · 5 of 15 never named it
A hollow mark is a model that never named the brand in any of its runs, which scores 0. Agreement is not endorsement — a brand every model ignores equally agrees just as tightly as one they all rank first.
Across → how frequently the panel names the brand at all.
Up ↑ how often the answers that name it put it first.
The horizontal line sits at 20% — the rate a named brand would lead at if the models were picking one of its 5 slots at random. Above it they are choosing it first on purpose. Both figures average across models, so a thinly sampled model counts the same as a heavily sampled one.
Jotform leads with a score of 97 of 100, ranked by 15 of 15 models.
The panel agrees on the winner and disagrees on almost everything below it. Jotform has a brand score of 97.33 and was ranked by all 15 models. Twelve models (the gpt-5 through gpt-6.1 family plus grok-4.7) gave it a perfect 100, and o3 gave it 98. The only dissent comes from the older models: o1 scores it 88 with a median rank of 2, and gpt-4.1 scores it 74 with a median rank of 3. Below Jotform, positions two through five change noticeably depending on which model generation is answering.
Typeform is the clearest generational split. All 15 models name it, but per-model scores run from 32 to 92, with a standard deviation of 18.09. o1 scores it 92 and puts it first in 6 of its 10 runs, and gpt-4.1 and o3 each give it 82. The gpt-5.x models and grok-4.7 mostly hold it at a median rank of 2. The gpt-6 generation pushes it down to a median rank of 4: gpt-6-astra and gpt-6-luna score it 46, gpt-6.1-sol 34 and gpt-6-sol 32. The models describe Typeform's conversational format in nearly the same words. What differs is how much weight they give to its pricing and response limits.
Tally largely takes the slot that the gpt-6 models deny Typeform. It is ranked by only 10 of 15 models and never first, but gpt-6-astra scores it 80, gpt-6.1-sol 64, and gpt-5.6-sol and gpt-6-sol 62 each. Five models never rank it, and gpt-5.2, grok-4.7, gpt-6-luna and gpt-5.5 score it 8 or below.
Typeform ahead on 11 of 15 models
Google Forms splits the panel along a different line, one that does not follow model generation. Per-model scores run from 18 to 84, with a standard deviation of 20.76. gpt-4.1 scores it 84 and is the only model to rank it first, doing so in 4 runs. gpt-6-luna (80) and gpt-6-sol (76) consistently place it second. At the other end, gpt-5.6-sol and gpt-6-astra score it 18 and gpt-5 scores it 22, and these three often leave it out. The models list the same strengths (free, familiar, connected to Sheets) and the same gaps (branding, payments, advanced logic). The split comes from how much those gaps count against it for customer-facing forms.
per-model scores 18–84 of 100 · mean 54 across 15 models
The lower tier is labelled "broad agreement" throughout. The models agree on placement at fourth or fifth and differ mainly on how often they include a brand at all. Even so, there are some model-specific champions:
WPForms, HubSpot and SurveyMonkey each score under 1. Fillout, 123FormBuilder and Formsite appear once each.
A broad pattern emerges from these scores. The older models (o1, gpt-4.1, o3) favour established names: Typeform, Wufoo and Formstack. The gpt-6 generation favours value-oriented or specialist tools: Tally, Cognito Forms and, for some variants, Google Forms.
The recalled sources line up with this pattern, but the connections below are hypotheses, not something the data establishes. The sources are recalled from memory, not verified citations.
The tentative reading is that models drawing more on vendor pricing pages reward low-cost tools. The category-wide anchor is G2's online form builder category (110 recalls), which nearly every model cites, so it cannot by itself explain any split.
69 sources across 1811 references, grouped by site from 416 recalled names. The top 5 carry 63% of them.
Sources are what each model recalled as having shaped its view. It does not mean that the sources were actually used in training. Recalled URLs might also be invalid or hallucinated.
g2.com
also named JotForm
Jotform holds a brand score of 97.33 and a first-choice score of 94.67, and all 15 models rank it. Twelve of them (the gpt-5 through gpt-6.1 family plus grok-4.7) give it a perfect 100 and put it first in every run. o3 follows at 98, with nine first-place runs. Per-model scores span 74 to 100 with a standard deviation of 6.92, which the report labels "models agree." The only real dissent comes from the older models: o1 scores it 88 with a median rank of 2, and gpt-4.1 scores it 74 with a median rank of 3, placing it first in just three runs.
Jotform · 26 points apart on a 0–100 scale
The reasoning is strikingly uniform. Models describe Jotform as the all-around or safest default pick because it pairs an approachable drag-and-drop builder with a large template library, payment collection, conditional logic, approvals and integrations. They also say it lets a business grow from contact forms into intake, orders and workflows without switching tools. The caveats differ by model:
Behind these themes, the newer models draw heavily on Jotform's own site and pricing page. The others lean on G2 (especially its Jotform reviews page and its online form builder category), Capterra product pages, Zapier's best-of roundups and PCMag reviews.
“Jotform is usually the safest all-around recommendation for a small business because it is easy to launch, has a broad template library, supports payments, and scales from simple contact forms to more operational workflows.”
“However, its interface can feel cluttered, and some advanced features require higher-tier pricing.”
Repeatedly calls Jotform the best overall balance of ease, features, and price, emphasizing its template library, built-in payments, approvals, and integrations, with HIPAA and advanced automation available as a business grows.
Describes Jotform as a strong all-around choice that is quick to set up yet offers depth (payments, approvals, file uploads, integrations) and scales from simple contact forms to lightweight workflows without technical work.
Frames Jotform as the safest all-around recommendation because it balances ease of use, broad templates, payments, and workflow features without feeling enterprise-heavy, fitting common SMB needs like lead capture, intake, and registrations.
Names Jotform its top pick for most small businesses, balancing ease of use, a large template library, payments, approvals, and integrations, and citing practical use cases like bookings, order forms, registrations, and intake.
Highlights Jotform's combination of templates, payment collection, integrations, and workflow features without technical expertise, with a free plan for testing and paid tiers for growth, while noting plan limits and pricing can rise.
Calls Jotform the strongest all-around option thanks to its mix of templates, conditional logic, payments, integrations, and workflows, while consistently cautioning that free-plan submission and storage limits may force an upgrade.
Repeatedly calls Jotform the strongest all-around choice, combining an approachable drag-and-drop builder with templates, payments, e-signatures, approvals, and integrations so a business can expand use cases without switching platforms.
Nearly verbatim calls Jotform its best all-around recommendation for its approachable builder, extensive templates, payments, and integrations covering inquiries to orders, while advising users to check submission and storage limits.
Presents Jotform as a strong all-around choice with broad templates, payment collection, and integrations that grow with a business, routinely noting that its free plan is easy to start with but usage limits may require a paid plan.
Recommends Jotform as a top pick for businesses needing more than simple surveys (payments, registrations, approvals), while often noting its breadth can feel busy or complex and that free-plan limits should be checked.
Consistently names Jotform its best all-around recommendation for combining an approachable builder with templates, integrations, payments, and conditional logic, with a recurring caveat about submission and storage limits on plans.
Positions Jotform as the strongest default for small businesses, pairing a drag-and-drop builder with payments, conditional logic, and a large template library at approachable pricing, plus integrations like Google Sheets, Stripe, and PayPal.
Emphasizes Jotform's generous free tier, hundreds of templates, and beginner-friendly drag-and-drop builder, plus payment, CRM, and HIPAA/automation features that let a business scale without outgrowing the platform.
Focuses on Jotform's user-friendly drag-and-drop builder, wide template selection, and integrations, frequently mentioning a generous free plan suited to budget-conscious small businesses.
Consistently cites Jotform's generous free plan, large template library, and integrations as good value for small businesses, while often noting a cluttered interface and advanced features locked behind paid tiers.
Each model’s own references for this brand, grouped by site. Tile area is that site’s share of the model’s references; tap one for the pages behind it.
jotform.com
jotform.com
capterra.com
capterra.com
Sources are what each model recalled as having shaped its view. It does not mean that the sources were actually used in training. Recalled URLs might also be invalid or hallucinated.
Typeform is named by all 15 models, which gives it a brand score of 66.27 and a first choice score of 43.32. Reach is universal, but its placement varies widely. Per-model brand scores run from 32 to 92, with a standard deviation of 18.09, which the report labels "models split." The enthusiasm sits mostly with older models. o1 scores it 92 and puts it first in 6 of its 10 runs. gpt-4.1 and o3 each give it 82. The gpt-5.x family and grok-4.7 cluster around a median rank of 2, often explicitly placing it just behind Jotform. The gpt-6 generation pushes it down to a median rank of 4: gpt-6-astra and gpt-6-luna score it 46, gpt-6.1-sol 34 and gpt-6-sol 32.
per-model scores 32–92 of 100 · mean 66 across 15 models
The praise is nearly identical across the panel. Models credit its polished, conversational, one-question-at-a-time format with lifting completion rates and brand impression for lead capture, surveys and feedback. The disagreement comes from how much weight each model gives to pricing and response limits. Higher-ranking models treat cost as a footnote. Lower-ranking ones frame it as the reason Typeform is not a default for cost-conscious or operational use, ranking it beneath Jotform or Tally. The recalled sources follow the same divide. o1, gpt-4.1 and o3 lean on third-party review sites such as G2's Typeform reviews page, Capterra, PCMag and TechRadar. The gpt-6 models cite almost exclusively Typeform's own homepage and its pricing page, which fits their focus on plan limits and cost.
“It stands out for its conversational interface, which can help boost response rates.”
“Its pricing and usage limits make it harder to recommend as the default choice for a cost-conscious small business.”
Mostly emphasizes Typeform's sleek, engaging, conversational interface and customization that boost response rates and brand image, with only occasional mention of higher pricing.
Consistently highlights Typeform's visually appealing, user-friendly, conversational forms that boost engagement and response rates, and often notes that pricing and free-plan limits may be restrictive for some small businesses.
Repeatedly describes polished, conversational UX that boosts completion rates for lead capture and customer-facing surveys, and says pricing and response caps can add up as volume grows.
Frames Typeform as best when completion rates and a premium conversational experience matter for leads and customer research, while noting it gets pricier at scale and is less back-office oriented.
Consistently praises the conversational design and higher completion rates for customer-facing forms, but ranks it slightly below Jotform because of higher pricing and less operational utility.
Repeatedly calls Typeform excellent when presentation and completion rates matter for surveys, quizzes and lead capture, but ranks it below Jotform because of cost and weaker fit for operational forms.
Repeatedly describes a polished, conversational respondent experience for lead forms, quizzes and feedback, while noting it is less economical than Jotform at higher volumes or for operational needs.
Consistently cites the one-question-at-a-time interface driving higher completion rates and polished branding, but places Typeform just behind Jotform because of steeper pricing and response limits.
Focuses on the one-question-at-a-time conversational format that improves completion and brand impression for customer-facing forms, with paid plans getting expensive quickly as responses grow.
Emphasizes a polished, conversational experience and branding suited to lead generation and customer research, offset by pricing and response limits that are less appealing for high-volume businesses.
Consistently credits polished conversational forms for surveys and lead capture, but says pricing and response limits make it harder to justify for volume- or value-focused small businesses, ranking it below Jotform or Tally.
Consistently praises polished, conversational forms for lead generation and feedback, but ranks Typeform lower, often fourth or below Jotform and Tally, because of response limits and paid-plan costs.
Repeatedly calls Typeform a good fit when a polished, conversational experience matters, but cautions that its cost and plan limits may not be justified for routine or basic data collection.
Consistently notes Typeform suits forms where respondent experience and presentation matter, but says pricing and usage limits make it hard to recommend as a default for cost-conscious small businesses.
Repeatedly highlights polished, conversational forms for feedback and lead generation, but ranks Typeform lower, often fourth or fifth, because response allowances and paid-plan costs need scrutiny when presentation isn't the priority.
Each model’s own references for this brand, grouped by site. Tile area is that site’s share of the model’s references; tap one for the pages behind it.
g2.com
Sources are what each model recalled as having shaped its view. It does not mean that the sources were actually used in training. Recalled URLs might also be invalid or hallucinated.
also named Google
Google Forms is named by all 15 models but lands in the middle on strength, with a brand score of 53.6 and a first choice score of 33.67. Per-model brand scores run from 18 to 84 with a standard deviation of 20.76, which the report labels as models split. gpt-4.1 is the outlier at the top: it scores the brand 84 and is the only model to put it first, doing so in 4 of its runs. gpt-6-luna (80) and gpt-6-sol (76) also rank it highly, but consistently in second place. At the other end, gpt-5.6-sol and gpt-6-astra each score it 18 and gpt-5 scores it 22. These three name it in fewer runs and usually place it last.
per-model scores 18–84 of 100 · mean 54 across 15 models
The disagreement is about where to place the tool, not what it is. Nearly every model describes it the same way: free or low-cost, familiar, quick to share, and feeding responses into Google Sheets, which makes it a natural fit for teams already on Google Workspace. The models apply the same limitations almost word for word: branding, design control, native payments, and advanced logic or workflows. How much weight a model gives those gaps for customer-facing forms largely determines whether it ranks Google Forms second or fifth. Several models frame it as a starter tool that growing businesses outgrow, and gpt-6-sol repeatedly sets it below Jotform. The recalled sources behind this picture are mostly Google's own pages, especially the google.com/forms/about page and assorted Google Workspace and help-center URLs. Third-party sources add to these, led by G2's Google Forms reviews page alongside Zapier roundups, PCMag, TechRadar and Capterra. o3 also leans on "personal experience."
“The trade-off is limited design customization and lack of native payments or advanced logic, so businesses often outgrow it as their needs expand.”
“It ranks last here because its limited branding flexibility and lack of native payment collection make it less versatile for customer-facing business workflows.”
Describes Google Forms as free, very easy to use, and well integrated with Google Workspace, suiting budget-conscious small businesses with basic needs. Notes limited customization and advanced features as the main drawback.
Describes it as an easy, low-cost choice for simple surveys, registrations, and feedback forms, particularly for Google Workspace users. Says it is less suited to businesses needing extensive branding or advanced workflows.
Presents it as an easy option for straightforward surveys and internal requests within Google Workspace. Repeatedly ranks it below Jotform because of limited design, branding, payments, and complex workflows.
Frames it as a low-cost, familiar tool for simple surveys, registrations, and internal data collection with natural Sheets integration. Notes it has fewer design, branding, payment, and automation features than dedicated platforms.
Calls it the practical no-cost default for Workspace users, with responses going straight into Sheets, suited to sign-ups, RSVPs, and internal requests. Describes it as a starter tool that is limited on branding, payments, and advanced logic.
Calls it hard to beat on cost and simplicity, with easy sharing and Google Sheets reporting. The recurring tradeoff is limited branding, logic, payments, and workflow features compared with dedicated builders.
Calls it an easy recommendation for budget-conscious teams on Google Workspace that need basic surveys and intake. Usually ranks it mid-pack or low because branding, payments, and advanced workflows are weaker than in specialized builders.
Treats it as a free, simple, reliable option feeding into spreadsheets, especially for Workspace users. Consistently ranks it lower due to limited branding, design, payments, and advanced workflows.
Positions it as a low-cost option for internal requests, RSVPs, and basic intake, particularly for teams already on Google Workspace. Says presentation, payments, branding, and advanced workflows are comparatively limited.
Emphasizes that it is free, straightforward, and integrates seamlessly with Google services, which suits basic data collection. Notes it lacks advanced customization features.
Stresses that it is completely free, easy to learn, and tightly integrated with Google Workspace and Sheets. Points to limited design, logic, and native payments as reasons growing businesses often outgrow it.
Calls it a practical starting point for simple surveys and internal data collection for Google Sheets users. Says limited branding and no native payment collection make it less versatile than higher-ranked builders.
Presents it as free, fast, and reliable for internal surveys and simple data capture with Sheets syncing. Says limited branding, payments, and advanced logic make it weak for polished customer-facing forms.
Describes it as free and familiar, adequate for simple surveys, registrations, and internal data collection. Frequently ranks it last because of limited branding, payment support, workflows, and customer-facing polish.
Calls it a practical, low-cost starting point for basic surveys and internal requests, especially for Google Sheets users. Ranks it low because of limited branding and the lack of native payment collection for customer-facing use.
Each model’s own references for this brand, grouped by site. Tile area is that site’s share of the model’s references; tap one for the pages behind it.
workspace.google.com
Sources are what each model recalled as having shaped its view. It does not mean that the sources were actually used in training. Recalled URLs might also be invalid or hallucinated.
Tally holds a mid-table position with a brand score of 22.8 and a first choice score of 14.02. It was ranked by 10 of the 15 models, and none of them put it first in any run. Its first choice credit comes entirely from second- and lower-place finishes. The panel is sharply divided. Per-model brand scores span 0 to 80, with a standard deviation of 27.99, which the report labels "models split." The strongest support comes from gpt-6-astra (80, ranked in all 10 runs), gpt-6.1-sol (64), gpt-5.6-sol (62) and gpt-6-sol (62), each typically placing Tally second or third. gpt-5 (28) and gpt-5.6-terra (20) name it less often. gpt-5.2, grok-4.7, gpt-6-luna and gpt-5.5 mention it only occasionally, with scores of 8 or below, and five models do not rank it at all.
per-model scores 0–80 of 100 · mean 23 across 15 models
The models that do rank Tally describe it in remarkably similar terms. They call it a budget-conscious or value pick, built on a generous free tier, with features such as conditional logic, payments and unlimited forms under fair-use terms, and a fast, document-like or Notion-style editor. The recurring ceiling is Jotform. Nearly every model places Tally just behind it, citing Jotform's broader template library and integration ecosystem. Several models, including gpt-5.2, gpt-5.5 and gpt-6-sol, also flag thinner support for advanced workflows, compliance or governance. The recalled sources behind the pricing and feature claims are mostly Tally's own pages: its homepage, pricing page and help center. Product Hunt, G2, Zapier's form-builder roundups, Capterra and Reddit r/nocode appear in smaller numbers, mostly alongside the "newer, smaller ecosystem" framing.
“I rank it just behind Jotform because its simpler ecosystem may be less suitable for businesses needing extensive templates and specialized workflows.”
“Tally is appealing for very small businesses because it is clean, fast, and generous at low cost.”
Repeatedly frames Tally as ideal for budget-conscious businesses because of a generous free offering and a fast, document-like editor. It consistently ranks Tally below Jotform for its less extensive templates, integrations, and advanced workflow features.
Highlights Tally's clean, modern, document-like editor and capable free tier as a value option for budget-conscious teams. It notes a smaller ecosystem and less enterprise-style breadth than Jotform and other established tools.
Consistently calls Tally its budget-friendly or value pick, citing a document-style editor and generous free features such as conditional logic and unlimited forms under fair-use terms. It places Tally just behind Jotform because of Jotform's broader templates and integrations.
Frames Tally as especially attractive for budget-conscious businesses, citing a generous free offering that includes conditional logic and payments, plus an easy document-style editor. It consistently ranks Tally below Jotform for Jotform's broader template and integration ecosystem.
Describes Tally as outstanding value for lean teams, citing a generous free plan with unlimited forms/responses and a simple Notion-like editor. It consistently notes the trade-off of fewer native integrations and advanced features than larger platforms.
Sees Tally as appealing for clean forms with a low learning curve and a generous free offering. It notes it may be less complete than Jotform for broad integrations or advanced workflows.
Repeatedly describes Tally as good for building clean, straightforward forms quickly with little setup and low cost. It ranks Tally below Jotform or Google and advises checking its integrations and workflow coverage before committing.
Emphasizes Tally's unusually generous free tier and clean, modern builder suited to straightforward forms and light automation. It notes that, as a newer and simpler platform, it may lack deep integrations, governance, and complex workflow support.
Presents Tally as clean, fast, and generous at low cost for very small businesses. It ranks it fourth because it is less established and may not cover advanced integration, compliance, or workflow needs.
Emphasizes Tally's generous free plan and document-like editor that suits non-technical small teams, along with adequate support for logic, embeds, and payments. It ranks Tally just behind established options because of a thinner template and integration ecosystem.
Each model’s own references for this brand, grouped by site. Tile area is that site’s share of the model’s references; tap one for the pages behind it.
tally.so
Sources are what each model recalled as having shaped its view. It does not mean that the sources were actually used in training. Recalled URLs might also be invalid or hallucinated.
also named WuFoo
Wufoo sits low in this panel. Its brand score is 15.6 of 100 and its first choice score is 11.31. Twelve of the 15 models rank it, but none ever puts it first. Its median ranks fall between 4 and 5, so the first choice score reflects partial credit for lower placements rather than any lead positions. Per-model brand scores span 0–48, with a standard deviation of 14.83, which the report labels "broad agreement." The models agree on Wufoo's placement near the bottom of their lists and differ mainly on how often they include it at all. gpt-5.4 names it in every run and gives it the highest score (48). o1 (36) and gpt-4.1 (34) also include it in nearly every run. By contrast, grok-4.7 (2), gpt-5.6-sol (4) and gpt-6-astra (4) mention it only occasionally.
per-model scores 0–48 of 100 · mean 16 across 15 models
The reasoning is remarkably uniform. Wufoo is described as long-established, reliable and straightforward for basic contact, registration and payment forms. Nearly every model then qualifies this by calling it dated or less modern in interface, design flexibility, automation and integrations. Several, notably o3 and gpt-4.1, add that its pricing and value lag newer rivals. o3 is the harshest, repeatedly calling Wufoo a stagnant pioneer and the least compelling option. The recalled sources behind these views are concentrated in a few places:
“Wufoo pioneered easy web form creation, but its interface feels dated and its feature set has stagnated relative to competitors.”
“I place it in the middle because it has a long-standing reputation and solid basics, but it feels less modern and less top-of-mind than some leading alternatives.”
Describes Wufoo as long-established and reliable with easy drag-and-drop building, but consistently notes its dated interface and pricing that is less competitive than newer alternatives.
Consistently treats Wufoo as a recognizable, approachable, long-standing option for simple use cases, but ranks it mid-to-low because it feels less modern and less top-of-mind than newer competitors.
Emphasizes Wufoo's user-friendly interface, ease of use, and reliability for simple forms, with caveats about limited advanced features, lower-tier restrictions, and a less modern feel.
Frames Wufoo as simple and reliable for basic forms, but says its dated UI and limited depth in customization, logic, and integrations lag newer competitors.
Describes Wufoo as dependable and straightforward for basic small-business forms and payments, but ranks it lower because it feels less modern and flexible than rivals such as Jotform and Typeform.
Repeatedly calls Wufoo straightforward, stable, and dependable for basic forms and payments, while saying it feels less modern in design, UX, and automation than newer leaders and offers weaker value.
Calls Wufoo capable for conventional forms with integrations, but says its interface and product experience feel less modern, so it is mainly worth choosing for existing familiarity or specific integrations.
Describes Wufoo as dependable and easy to learn, with templates, reports, and integrations, but notes its design and flexibility feel less modern than higher-ranked options.
Repeatedly frames Wufoo as a capable, established builder for conventional forms, reporting, and payments, but sees its less modern interface and product momentum as making it a weak default for new small-business setups.
Sees Wufoo as practical for straightforward contact, registration, and payment forms, but ranks it last because its traditional editing and design experience is less appealing than the alternatives.
Notes that Wufoo, now part of SurveyMonkey, still covers basic forms and payments but has seen less modern innovation, and is recommended mainly for teams already in the SurveyMonkey ecosystem.
Consistently calls Wufoo a pioneer of easy web forms whose dated interface, stagnant features, limited integrations, and weak price-to-value now make it the least compelling option.
Each model’s own references for this brand, grouped by site. Tile area is that site’s share of the model’s references; tap one for the pages behind it.
no URL recalled
capterra.com
capterra.com
no URL recalled
Sources are what each model recalled as having shaped its view. It does not mean that the sources were actually used in training. Recalled URLs might also be invalid or hallucinated.
where
where
Each one is a single question put to every active AI model, many times over.
Or scroll for more
Tell us what to ask and how wide to sample it — we run it and send back the report.