Every hiring manager runs a sequence. Three gates, in order. If you don't clear the first, the second never opens. The evaluator doesn't experience it as a sequence. They experience it as "something clicked" or "something felt off in the first ten minutes." But the structure is there, it's consistent within each tier, and it's different across tiers in ways that will change which version of your story you tell, which pillar you lead with, and which parts of your background you keep quiet about until asked.
What follows is the gate sequence for all four tiers on your target list: AI-native, regulated growth, enterprise platform, and healthcare/vertical SaaS. Grounded in current posting language, not About pages.
Most of the people you need to reach are offline through Monday. The holiday weekend is a gift. Use it to install these patterns so that when H2 headcount unlocks in the next two weeks, the tier model is already loaded. Every Battlecard reads faster. Every outreach message lands at the right altitude. Every first conversation starts with you already knowing what the person across the table is scanning for, even though they couldn't articulate it if you asked.
Before you read further: the same pillar that opens a door at one tier will close it at another. Alibaba is your lead at enterprise. It's your third move at AI-native. The essay is your strongest asset at AI-native. It's supporting evidence at regulated growth. That's what the tier model does. It tells you which version of you to send into which room.
TinyFish product work is not in your published portfolio and should never appear in outreach, cover letters, or portfolio presentations at any tier. If it comes up in conversation, the framing is identical every time: "Product leadership sharpened how I think about commercial outcomes. Deliberate detour. Returning to design leadership with stronger product judgment." Then pivot to published work immediately. Name TinyFish directly. Don't be evasive. Then move.
AI-Native Tier — OpenAI, Anthropic, Suno
What they're buying: A thesis. These companies are inventing product categories in real time. The interaction paradigms don't exist yet. They need a design leader who arrives with a point of view about what the paradigm should be. Running a discovery sprint to figure it out is too slow and, more importantly, signals that you haven't been thinking about this problem on your own time. High confidence on this read. Every AI-native posting I've tracked in the last six months tests for pre-existing intellectual frameworks over process capability.
The make-or-break moment: Gate 1. Category judgment. If you don't clear it at the profile-scan stage, Gates 2 and 3 never open. Allocate your preparation time accordingly.
The Org You're Walking Into
AI-native design teams are small, often under ten designers, sometimes under five. Design may report to a Head of Product, a CTO, or in some cases directly to the CEO. There is rarely an established design org in the enterprise sense. The role you're being hired into frequently is the design function, or will become it. OpenAI's Product Design Manager posting locates the role within "Product Design" but doesn't specify a reporting chain, a design VP, or an existing org structure. Suno's Head of Product Design posting is more transparent: the role owns Product Design, Design Engineering, and UX Research during "rapid growth," which means the function is being built, not inherited.
This is either the mandate you want or a warning sign. The difference depends on whether the company gives you real authority to build or expects you to execute within someone else's product vision. The diagnostic questions below will help you tell the difference.
Gate 1 — Category Judgment
This is the gate that matters most, and it fires before you're in the room. At the profile-scan and portfolio-review stage, the evaluator is asking one question: does this person already think about AI interaction at the level we need?
Your "Trust Is the New Interface" essay is the single strongest asset you have for this gate across your entire target list. Five handoffs, a trust ladder, four design jobs for agentic systems, all published. OpenAI's posting describes work across "Foundational Systems for consumer and enterprise experiences" without specifying what those systems look like. They don't know yet. They need someone who brings a framework. You have one.
Suno's posting reveals a different flavor. Where OpenAI and Anthropic frame category judgment around safety, trust, and infrastructure, Suno frames it around "simple, joyful, intuitive user experiences" and "emotional resonance." The category judgment Suno wants is consumer-creative: what does it feel like to make music with AI? The underlying gate is identical. The problem it tests for is different. Recognize which version you're facing in the first five minutes.
Your advantage: the essay plus three live Agentic Labs systems (Brand Pulse, Retail Velocity, Carrier IQ) prove you're not theorizing. You've built. The agentic visions embedded in your Alibaba and Thermo Fisher cases show architectural thinking applied to real domains.
Lead with the essay in every AI-native outreach. Link it. Let them read it before they see your portfolio. It reframes every subsequent case they review.
What the Portfolio Review Looks Like at This Tier
The AI-native evaluator does not open your portfolio looking for case studies. They open it looking for evidence of thinking. If the first thing they see is a traditional case-study layout (problem → process → outcome), they may not click deeper. They're scanning for a thesis, a framework, a published point of view that signals you've been working on this problem independently.
Your portfolio structure already supports this. The essay is published and linkable. The Agentic Labs systems are live. What matters is the sequence: essay first, live systems second, traditional cases third. The evaluator who reads the essay and then sees three functioning AI systems has already cleared Gate 1 before they reach Alibaba or Thermo Fisher. The evaluator who hits a traditional case study first may never get to the essay.
Metrics at this tier are secondary to paradigm judgment. "+20% daily transactions" matters at enterprise. Here, the evaluator cares about what interaction decisions you made and why. The autonomy threshold, the trust calibration, the moment you decided the human should see something versus the agent should handle it silently. Those are the details that signal taste.
Gate 2 — Product Taste and Hands-On Paradigm Judgment
Once category judgment is established, the evaluator shifts: can she actually make the thing? This surfaces in the portfolio review and first conversation. They're looking for evidence that you have opinions about interaction patterns, not process opinions about how to run a team.
Suno makes this explicit: the leader should be "willing to use the tools" to push a vision forward. Their senior designer posting asks for "pixel-level details, microinteractions, and motion design." The craft bar is high and specific. OpenAI's language is less explicit but the expectation is the same.
Your Agentic Labs work serves this gate. Three solo-built live systems demonstrate hands-on capability at the paradigm level. Be precise about what you show: the AI-native evaluator cares about the interaction decisions. What did you decide the human should see while the agent acts? Where did you draw the autonomy threshold? Those are the taste questions. Articulating that threshold from shipped work is the rarest credential at this level.
Gate 3 — Leadership Credibility at Scale
This gate opens last and is the easiest for you to clear, but only if you've already passed the first two. Alibaba is unambiguous scale evidence. The risk is leading with it. At an AI-native company, scale credibility without category judgment reads as a big-company operator who doesn't understand what they're building.
Subordinate Alibaba here. Third move, not first.
What Gets You Killed
Leading with team size or org-building narrative. These companies have small design teams and plan to keep them small. Emphasizing how many people you've managed signals you'll want to build an empire they don't want.
Framing AI as a feature to be designed around. The evaluator hears 2019. You need to treat AI as the material, not the context.
Presenting traditional case studies without a thesis. "Here's the problem, here's the process, here's the outcome" is invisible at this tier. They want to know what you believe about how this category of interaction should work.
Using "human-centered design" without updating the model. In agentic systems, the human isn't always centered. Sometimes the agent acts and the human monitors. If you use HCD language, show you've evolved it.
Tier Deviations
Suno breaks the pattern. Where OpenAI and Anthropic hire for infrastructure-level paradigm judgment, Suno hires for consumer-creative taste. Their posting foregrounds "craft and emotional resonance." If you engage Suno, shift your Gate 1 framing from trust architecture to creative-tool interaction design. Your Equinox+ work (consumer, 0→MVP, 4.8★) becomes more relevant here than at any other AI-native target.
Recognition cue: If the first conversation focuses on how the product feels rather than how the system works, you're in a consumer-creative evaluation.
Two-Directional Red Flags
OpenAI's posting does not specify a reporting line. Ask in the first call: "Who does this role report to, and does that person sit in product leadership meetings?" If design reports into engineering, the mandate is execution, not direction.
AI-native companies with no published design leadership on their team page may be hiring their first design leader. That's either a function-building opportunity or a signal that design has been an afterthought. The difference surfaces in one question: "What's the current design team structure, and what prompted this hire?"
If the interviewer spends more time testing your AI technical knowledge than your design judgment, the role may be scoped as a design-aware PM, not a design leader. Watch for this.
Player-Coach Diagnostic — AI-Native
"Hands-on" at this tier means paradigm-level craft. The evaluator expects you to prototype interaction patterns, not manage a team that prototypes them. Suno says it directly: "willing to use the tools." This is real player-coach authority. The company wants your hands on the work because the work is inventing the category.
Diagnostic question to ask early: "When a new interaction pattern needs to be explored, does the design leader prototype it directly or brief someone on the team?" If the answer is "brief someone," the role is management. If the answer is "both, depending on the stakes," it's genuine hybrid authority.
Regulated Growth Tier — Ramp, Gusto, Headway
What they're buying: Consequence awareness packaged inside speed. These companies are building systems that touch money, employment, insurance claims, and healthcare access. They need to move fast. But speed without consequence literacy is how you get a payroll error that affects someone's rent, or an insurance claim denial that leaves a patient without coverage. The design leader they're hiring needs to carry both of those truths simultaneously. High confidence on this pattern. The posting language across Ramp, Gusto, and Headway is remarkably consistent on consequence signals, even where speed language diverges.
The make-or-break moment: Gate 1. Consequence literacy. If the evaluator doesn't believe you've felt the weight of designing where failure has real-world consequences, your speed evidence and simplification skills are irrelevant. They won't trust you with their users' money, employment, or healthcare.
The Org You're Walking Into
Design orgs at this tier are in active growth. Gusto's posting says Design is "made up of more than 80 people," which is a mature function. Ramp and Headway are earlier in their design org development. The maturity spread within this tier is wide: you might walk into a team of 80 with established design leadership, or a team of 8 where you're the most senior design voice.
Neither Ramp nor Gusto specifies a reporting line in their current postings. This is common at growth-stage companies where the org chart is still settling. Design may have its own leadership track or may report through product. The ambiguity isn't automatically a red flag, but it's a question you need answered in the first conversation, because the difference between "design leader who partners with product leadership" and "design manager who reports to a product VP" determines whether you'll have the mandate you need.
Gate 1 — Consequence Literacy
The evaluator scans your background for evidence that you've designed in environments where failure has real weight. Real weight. Someone didn't get paid correctly. An insurance claim got denied. A disaster-relief disbursement didn't reach the family.
Gusto's Payroll design posting asks the leader to partner with Compliance and Ops alongside Product and Engineering. That's not a throwaway line. Payroll is regulated. The design leader needs to understand what happens downstream when a UI decision creates an ambiguity that a small-business owner misinterprets. Ramp's Director posting describes building in "an AI-native world" but the underlying product is corporate spend. Real money, real audit trails, real consequences when the automation gets it wrong.
Your advantage here is specific and strong. Thermo Fisher mySupply: pharma supply chain, $20M+ margin, exception-first design where "what happens when the system is wrong" was literally the design problem. American Red Cross: $847K disbursed, mission-critical, 6 systems consolidated to 1 under federal oversight. That is consequence literacy, full stop.
Lead with Thermo Fisher and Red Cross for every regulated-growth outreach. The evaluator needs to believe you've felt the weight before they'll trust you with theirs.
What the Portfolio Review Looks Like at This Tier
The regulated-growth evaluator opens your portfolio looking for evidence of stakes. They scan for domain signals first: what kind of product was this, and what happened when it went wrong? Metrics matter here, but the type of metric matters more than the magnitude. "$20M+ margin recovered" lands harder than "+20% daily transactions" because it implies a system where money was being lost and the design fixed the loss. "$847K disbursed" with "6 systems → 1" tells the evaluator you simplified something mission-critical without breaking it.
The IC-to-leadership shift at this tier is about judgment, not craft. The evaluator doesn't need to see your pixel work. They need to see that you made the right call when the edge case appeared, when the exception flow surfaced, when the automation failed. If your portfolio cases show the happy path, you'll pass as a competent designer. If they show the exception path and how you designed for it, you'll pass as someone who understands consequence. That's the difference between a polite rejection and a second conversation.
Durability evidence strengthens everything. If the Red Cross system is still running, say so. App Store records with recent update dates are stronger evidence than explaining what BCG Digital Ventures was.
Gate 2 — Simplification Without Surrendering Control
Once consequence literacy is established, the evaluator asks: can she make complex workflows simple without hiding the controls users need? This is the design problem that defines the tier. Gusto's employer+employee split. Headway's provider+patient split. Ramp's admin+employee+finance split. All require multi-role architectures where different users see different things but the underlying system maintains coherence.
Your Allē work maps directly: 30M members, 40K providers, dual-surface redesign, 3.2× redemption. Two audiences, one system, simplified without losing either side's control. The multi-role architecture pattern runs through your Red Cross and Thermo Fisher work too. One record, multiple role-appropriate views, permission-aware visibility.
Name the pattern explicitly. Don't make the evaluator infer it from separate case studies. Say: "The structural problem I keep solving is multi-role simplification. Making the same system feel simple to each audience without hiding what any of them need to control."
Gate 3 — Speed Inside Constraints
The final gate tests whether you'll import process that slows them down. Ramp's posting says the Director should "jump into details to unblock and ship." The question sounds like it's about speed. It's about behavioral defaults — whether your instinct under pressure is to add a review step or dissolve the constraint into the product.
Your Red Cross work is the kill shot. 0→1 in 6 months with regulatory compliance baked into the transaction workflow from day one. You didn't add a compliance review gate. You designed the compliance into the flow so it never became a bottleneck. Constraint dissolution. Fundamentally different from constraint management, which layers process on top of the product.
When speed comes up, tell the Red Cross story. Six months, national deployment, regulatory compliance embedded, not appended. Then let them do the math: if you shipped under federal oversight in six months, you can ship without it in six weeks.
If the work is still live, say so. "Shipped" is table stakes at Director+. "Still running years later" separates you from the field.
What Gets You Killed
Leading with Alibaba scale. The evaluator hears "big company, slow process." Subordinate enterprise scale behind consequence-literacy evidence.
Describing your design process in phases. "Discovery, definition, delivery" sounds like a consulting engagement. These companies ship continuously. Describe how you work inside a shipping rhythm, not above it.
Treating compliance as a constraint to manage rather than a material to design with. If you frame regulation as something that slows you down, you've failed the consequence-literacy gate. The strongest candidates at this tier treat the constraint as the design material.
Overemphasizing the AI essay without grounding it in consequence. The essay matters here, but only as a supporting signal. The evaluator cares about trust grounded in real stakes — payroll, insurance, healthcare access — and will discount abstract trust frameworks without that grounding.
Presenting portfolio work without durability evidence. App Store records with recent update dates are stronger evidence than explaining what BCG Digital Ventures was.
Tier Deviations
Headway runs two distinct speed profiles across its open design roles. The Provider Experience role uses explicit speed+consequence+AI language. The Core Insurance role uses zero speed/urgency language and emphasizes predictability and systems. Don't assume company-level culture maps uniformly to role-level requirements. Read the specific posting. The posting-date metadata on the Core Insurance role showed October 2025 while live in mid-2026 (moderate confidence; posting metadata can change), which signals either a long-open role or a repost without metadata cleanup. Both are meaningful intelligence about how that search is going.
Ramp leans harder into AI-native language than other regulated-growth companies. Their posting describes defining "how design works in an AI-native world." If you engage Ramp, elevate the Agentic Labs pillar alongside consequence literacy. Ramp wants both.
Recognition cue: If the posting mentions "agentic" or "AI-native" alongside financial/compliance language, the company is straddling two tiers. Lead consequence literacy, surface AI thesis early.
Two-Directional Red Flags
Neither Ramp nor Gusto specifies a reporting line in their current postings. Ask: "Does the design leader have a direct line to the executive team, or does design report through product?" At Gusto, with 80+ people in design, the answer should be direct. If it's not, the function is large but structurally subordinate. That's a specific kind of frustration you should decide whether you want before investing in the process.
If the interview loop includes no one from compliance, legal, or ops, the company may not actually integrate consequence thinking into design decisions. The posting language is aspirational, not operational.
If the evaluator asks primarily about visual craft or design-system contributions, the role may be scoped as a senior IC with a leadership title. Real consequence-literacy roles test judgment, not pixels.
Player-Coach Diagnostic — Regulated Growth
"Hands-on" at this tier means you'll personally work through the hardest edge cases. The exception flows, the error states, the moments where automation fails and a human needs to understand what happened. Ramp says "jump into details to unblock and ship." This is genuine hybrid authority when it means the leader owns the hardest design decisions. It's under-scoped leadership when it means there's no team and you're the only designer.
Diagnostic question: "How many designers are on this team today, and what's the hiring plan for the next two quarters?" If the answer is fewer than three with no plan to grow, "hands-on leader" means solo IC with a director title. The posting language will not tell you this. The question will.
Enterprise Platform Tier — Salesforce, Atlassian, Adobe
What they're buying: Coherence. Across surfaces, systems, org boundaries, and now AI modernization. The enterprise design leader's job is to make a sprawling product portfolio feel like one thing to the customer while navigating an internal organization where design is one voice among many, and often not the loudest. High confidence on this pattern. It has been stable for years and the current postings confirm it.
The make-or-break moment: Gate 2. Cross-functional influence. Enterprise companies assume scale credibility from your resume. They'll verify it, but the real screening event is whether you can represent design in executive rooms where design is not the default language. That's where candidates with strong portfolios and weak organizational instincts get cut.
The Org You're Walking Into
Enterprise design orgs are large and established. Salesforce's VP posting specifies that the VP reports directly to the SVP of Salesforce UX & Product Design and sits on that organization's leadership team. Intuit's Director posting describes leading approximately 55 designers. These are mature functions with established processes, design systems, and reporting structures.
Design at this tier typically has its own VP or SVP track. The function exists. The job is stewardship, evolution, and increasingly AI modernization. This is categorically different from function-building. The mandate is to make the existing design org better while integrating AI without breaking what works. That's a legitimate and challenging mandate, but decide whether it's the one you want. If you're looking for function-building authority, enterprise roles at established companies may not deliver it. The exception is enterprise companies launching new product lines (Adobe GenStudio, Salesforce Agentforce), where the sub-org is being built even though the parent org is mature.
Gate 1 — Scale Credibility
The evaluator's first scan is for evidence that you've operated at platform scale — maintained coherence across multiple products, surfaces, and teams simultaneously. Salesforce's VP posting describes work across Data360, MuleSoft, Tableau, and Agentforce. Intuit's Director posting describes leading approximately 55 designers across Money, Capital, Lending, and Workforce Solutions. The scale is organizational as much as product.
Your Alibaba work is the lead. Head of Design and Research, North America. $50B+ GMV. Coherence across homepage, search, and PDP for 200K+ suppliers. +20% daily transactions and +2.2pt NPS provide the business-outcome proof enterprise evaluators need to justify the hire internally.
But scale credibility alone doesn't clear the gate. The evaluator also needs to see that you operated within a complex organizational structure, not just a complex product. Mention cross-functional sprint leadership. Mention that you named the structural gap and built the mandate. Enterprise evaluators know that the hardest part of their job is organizational, not design. They're listening for whether you know it too.
What the Portfolio Review Looks Like at This Tier
The enterprise evaluator reviews your portfolio differently than any other tier. They're looking for leadership artifacts, not craft artifacts. In the first 30 seconds: do the case studies read as "I designed this" or "I led the team that designed this"? The framing matters. An enterprise evaluator scanning a portfolio that reads as IC work, no matter how excellent, will categorize you as a senior individual contributor, not a design leader.
Metrics at this tier need to be business metrics, not design metrics. "+20% daily transactions" and "+2.2pt NPS" work because they're outcomes a VP of Product or a CFO would care about. "Reduced time-on-task by 30%" is a design metric. It's useful as supporting evidence but won't anchor an enterprise evaluator's assessment.
Systems thinking should be visible in how you present, not just what you present. If your cases show coherence across multiple surfaces (homepage, search, PDP at Alibaba), the evaluator reads that as platform thinking. If each case reads as an isolated project, even a great one, the evaluator wonders whether you can hold a portfolio together. Allē's five distinct design systems is relevant evidence of range, but frame it as systems-level thinking, not as five separate projects.
The design-systems conversation will come up. Enterprise evaluators will probe your relationship to design systems as organizational infrastructure — how they function as governance, how they enable velocity across teams, how they maintain coherence when dozens of designers are shipping simultaneously. Have answers.
Gate 2 — Cross-Functional Influence and Executive Advocacy
This is where enterprise evaluation diverges most from other tiers. The evaluator asks: can she represent design in rooms where design is not the default language? Salesforce's posting specifies sitting on the UX leadership team and driving product strategy. Adobe's Frame.io Director posting emphasizes "design storytelling in executive forums" and the ability to "influence senior leadership." Intuit's VP posting measures success through platform adoption, growth outcomes, and AI+HI effectiveness. Business metrics, not design metrics.
Your TinyFish product leadership experience is relevant here, but only verbally. In conversation, the product title becomes an asset: "I've sat on the product side of the table. I know what product leaders need from design leadership because I've been the product leader." Don't put it in writing. Don't lead with it. But when the evaluator probes for cross-functional influence, the product detour is evidence that you speak both languages. This is the tier where that verbal framing carries the most weight.
Gate 3 — AI Modernization That Respects Existing Infrastructure
Enterprise companies are all racing to integrate AI, but they can't break what's already working. Salesforce's posting explicitly names "agentic design" and guiding "ethical and transparent AI use." Adobe's GenStudio posting describes the shift from traditional creative editing to "intelligence-layer design." Intuit's VP posting frames it as AI-powered platform strategy combining AI capabilities and human expertise.
Your essay serves this gate differently than it serves the AI-native tier. Here, it proves you've thought about AI integration systematically in a way that respects the user's existing mental model. Enterprise evaluators want evolution with a framework. The five handoffs and trust ladder provide exactly that — something systematic that maps onto existing product infrastructure without demanding a tear-down.
Deploy the essay as supporting evidence at this tier, not as the lead. Alibaba opens the door. The essay proves you can modernize without destabilizing.
What Gets You Killed
Leading with 0→1 build stories. Enterprise evaluators hear "startup person" and worry you can't navigate organizational complexity. Subordinate Red Cross and Thermo Fisher behind Alibaba.
Presenting the Agentic Labs work as your primary credential. Solo-built AI systems read as IC work to an enterprise evaluator. They want to see you leading teams that build systems.
Describing design culture aspirationally. "I believe design should have a seat at the table" tells the enterprise evaluator you've never had one. Describe how you operated when you did.
Underestimating the design-systems conversation. Enterprise evaluators will probe your relationship to design systems as organizational infrastructure. Allē's five distinct design systems is relevant evidence here.
Ignoring the reporting-line signal. Salesforce specifies that the VP reports to the SVP of UX & Product Design. If a posting doesn't specify, that's a signal worth investigating.
Tier Deviations
Adobe GenStudio breaks the enterprise pattern by posting an explicit "player/coach" role that asks the Director to "roll up sleeves to prototype, tinker, and experiment" with AI-first agentic workflows and emerging UI patterns like "chat + canvas + editor." This is an enterprise company hiring with AI-native evaluation criteria. If you engage GenStudio, lead with the Agentic Labs pillar and the essay, not Alibaba. The Frame.io Director role at the same company uses traditional enterprise evaluation: multi-year vision, org capability, creative direction. Same company, different gates.
Recognition cue: If the posting names specific emerging interaction patterns (chat + canvas, agentic workflows) rather than platform coherence, the role evaluates like AI-native regardless of the company's size.
Two-Directional Red Flags
Salesforce's posting specifies the VP sits on the UX leadership team. This is a positive structural signal. Design has organizational representation. If an enterprise posting doesn't specify leadership-team membership, ask directly.
If the interview loop is entirely product managers and engineers with no design executives, design may not have the organizational weight the posting implies. Ask: "Who on the executive team champions design decisions?"
Enterprise roles with 15+ year experience requirements and global team mandates (Intuit's ~55 designers, Salesforce's global scope) may not offer the function-building authority you're looking for. The function already exists. The job is stewardship, not creation. That's a legitimate career choice, but decide whether it's yours before investing in the process. Every company says they're design-led; almost none are, and the ones that are don't need to say it. At enterprise scale, look for observable evidence: is there a design executive on the leadership page? Do designers report to design leaders or to product managers? Are design roles posted at leadership levels, or only IC?
Player-Coach Diagnostic — Enterprise
"Hands-on" at this tier means creative direction, not execution. Adobe's Frame.io posting says the Director will "regularly provide hands-on creative direction." Note the word "direction." Salesforce's VP posting doesn't mention hands-on work at all. At enterprise scale, "player-coach" usually means you set the bar through design reviews and critique, not through opening Figma.
The exception is Adobe GenStudio, which genuinely wants prototyping. But that's the deviation, not the pattern.
Diagnostic question: "What does a typical design review look like, and who gives the final creative call?" If the answer involves the design leader giving detailed feedback on work produced by the team, that's real creative authority. If the answer involves product or engineering sign-off after design review, the design leader is advisory, not authoritative. The distinction matters for whether you'll have the mandate you need.
Healthcare / Vertical SaaS Tier — Maven, Aurora Solar, Amae Health, Ontra, Front
What they're buying: Someone who understands that the user is an expert in a domain the designer is not, and that the design's job is to make that expert faster and more confident without overriding their judgment. Clinical decisions, legal contract workflows, solar installation engineering, customer operations at scale. These are all domains where the end user knows more than the designer about the stakes of getting it wrong.
I want to be direct about something before walking through the gates. This tier has the widest variance in design maturity and organizational mandate of any cluster on your list. Maven has a current Head of Design posting that reports to the Chief Product Officer, leads roughly 15 designers, and explicitly mandates building an AI-native design practice. It's one of the strongest Juno-shaped roles in this tier. Aurora Solar has a current Head of Design posting for a five-person team working on systems-level problems in residential and commercial solar, with an explicit AI-capabilities mandate. Ontra is building a real design function and says so explicitly. Front is hiring for workflow complexity that has more in common with enterprise B2B than with healthcare. Amae Health's design leadership posting status is uncertain. The tier pattern holds, but treat every company in this cluster as a potential deviation until the first conversation confirms otherwise. Moderate confidence on the overall pattern, high confidence on Maven, Ontra, and Front specifically, speculative on Amae.
The make-or-break moment: Gate 1. Domain consequence understanding. If the evaluator doesn't believe you respect the weight of their domain, nothing else matters. But unlike AI-native, where Gate 1 fires at profile scan, this gate often fires in the first conversation. These evaluators want to hear how you talk about domains you didn't know before you designed for them. How you describe learning pharma supply chain or disaster-relief operations tells them whether you'll respect theirs.
The Org You're Walking Into
Design orgs at this tier range from nascent to mid-sized. Maven's Head of Design posting describes leading roughly 15 designers across product design, brand design, and user research, reporting to the CPO. That's a real function with real headcount. Ontra's VP of UX posting describes building a UX organization, which means the organization doesn't fully exist yet. The role reports to the SVP, Product. Aurora Solar's Head of Design posting describes a five-person team, which is function-building territory. Front's Director posting leads the Product Design team, but the posting's emphasis on establishing critiques, design reviews, and design-system governance suggests the function's infrastructure is still being built.
The gap between the title on the posting and the actual organizational authority is widest here. A VP of UX at a 200-person vertical SaaS company may have less real influence than a Senior Product Designer at Ramp. The title tells you what they want to call the role. The headcount, the reporting line, and the hiring budget tell you what the role actually is. Ontra's posting confirms the VP reports to SVP, Product. Maven confirms the Head of Design reports to the CPO. These are common reporting structures in vertical SaaS and not automatically red flags, but the distinction between having a voice in product strategy and executing strategy set by product leadership is the difference between a mandate and a service function.
This is also the tier where function-building authority is most available, if the company is genuinely investing. Your Red Cross and Thermo Fisher 0→1 experience becomes the lead story in that scenario. But verify the investment is real before committing.
Gate 1 — Domain Consequence Understanding
The evaluator scans for evidence that you've designed in environments where the domain itself constrains what's possible, and where you treated those constraints as design material rather than obstacles. Amae Health builds for severe mental illness: schizophrenia, schizoaffective disorder, bipolar I, treatment-resistant depression. Their engineering posting states plainly: "everything shipped affects patient outcomes and clinician effectiveness." Ontra's VP of UX posting describes designing systems "where intelligent agents do the work and humans guide, supervise, and trust the outcomes" across private-markets legal workflows built on 2 million contracts. Maven's Head of Design posting frames design as central to "member trust, market differentiation, clinical mission at scale." Maven's consumer Design Lead posting adds trust-sensitive language around "eligibility, pricing, access, personal data, or complex decision-making."
The consequence here is categorically different from the regulated-growth tier. Ramp worries about payroll errors. Amae worries about clinical drift in care for patients with treatment-resistant depression. Maven worries about member trust in a clinical context where family health decisions carry emotional and medical weight. The evaluator can tell immediately whether you understand that difference.
Thermo Fisher and Red Cross lead here. Pharma supply chain where exception-first design was the mandate. Mission-critical disaster relief where 6 systems became 1 under conditions where failure had direct human consequences. These cases demonstrate that you've designed in domains where you were not the domain expert but you understood the weight of the domain. That combination is exactly what this tier evaluates for.
What the Portfolio Review Looks Like at This Tier
The healthcare/vertical SaaS evaluator reviews your portfolio looking for domain depth, not domain match. They don't expect you to have designed for their specific vertical. They expect you to show that you've entered an unfamiliar domain, respected its complexity, and designed something that made the domain expert's life better without oversimplifying their work.
The IC-to-leadership shift at this tier is less about organizational scale and more about judgment under domain uncertainty. The evaluator wants to see evidence that you made design decisions in a domain where you were not the expert, and that those decisions held up. Thermo Fisher's 100% partner adoption and Red Cross's national deployment are the right kind of proof: the domain experts adopted what you built. That's the strongest signal of domain humility combined with design competence.
Workflow depth matters more than visual polish. If your portfolio cases show a complex workflow simplified without losing expert control, that's the signal. If they show beautiful screens without workflow context, the evaluator wonders whether you understand what they're building. Front's posting makes this explicit: "queues, routing/assignment, rules builders, automation, and analytics." The workflows are the product.
Gate 2 — Workflow Simplification That Preserves Expert Trust
Once domain consequence is established, the evaluator asks: can she simplify our workflows without making our expert users feel like the system is dumbing down their work? This is the central design tension in vertical SaaS. The expert user needs to trust that the system respects their expertise.
Front's Director posting makes this concrete: the design leader must operate in "complex, workflow-heavy, data-rich product areas such as queues, routing/assignment, rules builders, automation, and analytics." The complexity is inherent to the domain. Simplifying it means making the complexity navigable, not hiding it.
Amae's VP Clinical Care Delivery posting uses language about "structured repeatable workflows," "gaps between design and real-world execution," and "technology that enhances rather than burdens care delivery." The word "burdens" is the tell. The evaluator has seen technology that made expert workflows worse. They're screening for designers who understand that risk.
Your multi-role architecture pattern is directly relevant. Allē's dual-surface redesign (30M members, 40K providers) demonstrates simplification across expert and consumer audiences without collapsing either experience. Name the pattern. Make it legible. Don't make the evaluator reconstruct it from separate case studies.
Gate 3 — Trust Architecture Including AI-Native Practice Building
This tier is integrating AI with higher caution than any other. Ontra's career page describes an "AI-native organization" that uses AI to "focus on complex, high-judgment work." Front's posting asks the Director to "operationalize an AI-first design process" while "protecting quality and customer trust." The word "protecting" is doing heavy lifting. AI is welcome, but trust is the constraint. Maven's Head of Design posting is the most explicit in this tier: the leader will "build an AI-native design practice across research synthesis, prototyping, production, and quality review" and lead "AI-native product experiences such as conversational interfaces, agent-driven workflows, and adaptive onboarding."
Your essay and Agentic Labs work serve this gate as proof that you've thought about where agent autonomy ends and human control begins. The trust ladder (Watch → Verify → Delegate) maps directly to how a clinician should relate to an AI-assisted care recommendation or how a legal professional should relate to an AI-generated contract analysis.
Deploy the essay as Gate 3 evidence, not Gate 1. Domain consequence opens the door. Workflow simplification builds credibility. The AI trust framework closes. The sequence matters.
What Gets You Killed
Presenting yourself as an AI-first designer. This tier wants AI-capable designers who lead with domain understanding. Leading with AI signals you'll prioritize technology over the domain.
Skipping the domain-learning narrative. The evaluator needs to hear how you learn a domain you don't know. Describe how you learned pharma supply chain at Thermo Fisher or disaster-relief operations at Red Cross. The learning process is the evidence.
Treating workflow complexity as a problem to be eliminated. The complexity exists because the domain is complex. The evaluator will reject anyone who seems to think the answer is "just simplify it."
Ignoring the expert-user relationship. If your portfolio presentation focuses on end-consumer experiences without showing how you designed for professional/expert users, you'll fail Gate 2.
Assuming healthcare and vertical SaaS evaluate identically. Maven's clinical-trust mandate is different from Front's customer-operations mandate. Read the specific posting. Adjust accordingly.
Tier Deviations
Front breaks the healthcare pattern. Its Director posting describes B2B customer operations complexity: queues, routing, rules builders, automation. No clinical or domain-specific consequence. Front's domain consequence is operational, not existential. The evaluation still gates on workflow complexity and expert trust, but the stakes are customer-relationship quality, not patient outcomes. If you engage Front, lead with Allē (dual-surface, operational complexity) over Thermo Fisher (pharma, clinical-adjacent).
Ontra straddles this tier and the AI-native tier. Their VP posting asks for someone who enjoys designing systems "where intelligent agents do the work and humans guide, supervise, and trust the outcomes." That's agentic-system language inside a vertical-SaaS context. The posting confirms the role reports to SVP Product. Lead with your AI trust framework alongside domain-consequence evidence. Ontra is one of the few companies on your list where the essay and the consequence-literacy cases carry equal weight.
Maven has a current Head of Design posting that is one of the strongest matches on your target list. It reports to the CPO, leads ~15 designers, and explicitly mandates building an AI-native design practice across research, prototyping, production, and quality review. The posting asks for AI-native product experiences including conversational interfaces, agent-driven workflows, and adaptive onboarding. This is an act-now target. Lead with Thermo Fisher for domain consequence, deploy the essay for Gate 3, and name the Allē dual-surface pattern for the consumer+clinical split.
Aurora Solar has a current Head of Design posting for a five-person team. The role asks for 12+ years in product design and 5+ years building design teams that ship technical SaaS products. The posting emphasizes AI-capability evolution, design operations, and KPIs tied to company OKRs. This is a function-building mandate in vertical SaaS. Lead with your 0→1 build evidence and name the Thermo Fisher domain-learning narrative.
Amae Health does not currently have a confirmed active design leadership posting, though one appeared briefly on their Greenhouse board. The absence is itself a signal: the company may be pre-search, using referral networks, or reconsidering scope. Monitor but don't outreach until a posting confirms. Moderate confidence on this read. The posting may have been pulled for revision, which would tell us something about what the first search taught them.
Recognition cue: If the first conversation focuses on "how do you learn a new domain" rather than "what's your AI thesis," you're in a domain-consequence evaluation. If it focuses on agent-human interaction patterns, the company evaluates closer to AI-native regardless of its vertical.
Two-Directional Red Flags
Ontra's VP of UX reports to the SVP, Product. Maven's Head of Design reports to the CPO. Ask: "Does the design leader participate in product strategy decisions, or execute on strategy set by product leadership?" The answer determines whether this is a leadership role or a service function.
If the company has no design leader on the executive team and no design representation in product-strategy forums, the Head of Design role is a service function with a leadership title. Amae's clinical and engineering postings reference partnering with the CPO, but the design-leadership posting's status is uncertain. That ambiguity is worth resolving before investing time.
Healthcare companies where the interview loop includes clinicians are testing for domain humility. This is a positive signal. It means clinical expertise shapes product decisions. Companies where the loop is entirely product and engineering may be building healthcare products without clinical governance, which creates design-mandate problems downstream.
Player-Coach Diagnostic — Healthcare / Vertical SaaS
"Hands-on" at this tier usually means the team is small and the leader designs. Front's posting says "senior hands-on design leader" who will improve throughput and raise quality across complex product areas. Ontra's posting describes building a UX organization, which means the organization doesn't fully exist yet.
This is where the player-coach diagnostic matters most, because the gap between "hands-on leader building a function" and "solo designer with a VP title" is widest in this tier.
Diagnostic question: "What's the current design headcount, and what's the budget authority for this role to hire?" If the answer is zero designers and no confirmed headcount, you're being hired as an IC. That might be fine if you want to build from scratch. Your Red Cross and Thermo Fisher 0→1 experience becomes the lead story in that scenario. But know what you're walking into. If the answer is a small team with approved growth, that's genuine function-building authority, and it's the scenario where your full range of experience, from hands-on craft to org design, is most valuable.
What the Side-by-Side View Reveals
Four patterns emerge when you lay the tiers next to each other.
| AI-Native | Regulated Growth | Enterprise | Healthcare / Vertical SaaS | |
|---|---|---|---|---|
| Gate fires at | Profile scan | First conversation | Portfolio review | First conversation |
| Lead pillar | Essay + Agentic Labs | Thermo Fisher + Red Cross | Alibaba | Thermo Fisher + Red Cross |
| Alibaba position | Third move | Subordinate | First move | Supporting |
| Essay position | Gate 1 (lead) | Supporting signal | Gate 3 (supporting) | Gate 3 (closer) |
| "Player-coach" means | Prototype the paradigm | Own the hardest edge cases | Set the bar via creative direction | Design the product (small team) |
Where the gate fires determines how you prepare. AI-native gates at profile scan. The essay needs to be in your outreach before they ever open your portfolio. Enterprise gates at portfolio review. Alibaba needs to be the first case they see. Regulated growth and healthcare/vertical SaaS gate in the first conversation. You have a few minutes of talking to establish consequence literacy or domain understanding. The preparation investment should match: for AI-native, spend your time on the outreach message and the essay link. For enterprise, spend it on portfolio sequencing. For the other two, spend it on your opening narrative.
Leadership credibility is weighted last at AI-native, first at enterprise. This is the most important structural difference across tiers. At AI-native, Alibaba is your third move. At enterprise, it's your first. The same evidence, positioned differently, serves different gates. If you're preparing for multiple tiers in the same week, consciously reset which version of yourself you're sending into each room.
The player-coach question means four different things. Same phrase in four postings, four completely different expectations. The diagnostic question for each tier is the fastest way to determine which version you're facing.
Your AI trust framework is universally relevant but tier-sequenced. It leads at AI-native (Gate 1). It supports at enterprise (Gate 3). It closes at healthcare/vertical SaaS (Gate 3). It's secondary at regulated growth (supporting signal only). The framework is the same. Where it falls in the conversation changes how it lands.
Quick-Take Cards
AI-Native (OpenAI, Anthropic, Suno)
Hiring posture: Buying someone who already has a thesis about how humans interact with AI systems. Developing one on the job is too slow for their timeline. Lead pillar: "Trust Is the New Interface" essay, then Agentic Labs live systems. Top landmine: Leading with team scale or org-building narrative. These teams are small by design. Opening question: "When a new interaction pattern needs to be explored, does the design leader prototype it directly or brief someone?" Pattern-break cue: If the conversation centers on how the product feels rather than how the system works, you're in a consumer-creative evaluation (Suno), not an infrastructure one.
Regulated Growth (Ramp, Gusto, Headway)
Hiring posture: Buying consequence literacy packaged inside speed. When automation fails, a human owns the outcome, and the design leader needs to have felt that weight before. Lead pillar: Thermo Fisher mySupply and American Red Cross. Consequence literacy, constraint dissolution, durability. Top landmine: Describing your design process in phases. These companies ship continuously. Opening question: "How many designers are on this team today, and what's the hiring plan for the next two quarters?" Pattern-break cue: If the posting mentions "agentic" or "AI-native" alongside financial/compliance language (Ramp), the company straddles two tiers. Lead consequence literacy, surface AI thesis early.
Enterprise Platform (Salesforce, Atlassian, Adobe)
Hiring posture: Buying coherence across surfaces, systems, and org boundaries, plus AI modernization that doesn't break what works. Lead pillar: Alibaba. Head of Design and Research, $50B+ GMV, cross-surface coherence, business outcomes. Top landmine: Leading with 0→1 build stories. Enterprise evaluators hear "startup person who can't navigate org complexity." Opening question: "What does a typical design review look like, and who gives the final creative call?" Pattern-break cue: If the posting names specific emerging interaction patterns (chat + canvas, agentic workflows) rather than platform coherence, the role evaluates like AI-native regardless of company size (Adobe GenStudio).
Healthcare / Vertical SaaS (Maven, Aurora Solar, Amae Health, Ontra, Front)
Hiring posture: Buying someone who respects that the user is a domain expert and designs to make them faster, without overriding their judgment. Lead pillar: Thermo Fisher and Red Cross for domain consequence. Allē for multi-role workflow simplification. Essay for Gate 3 AI-native practice building. Top landmine: Presenting yourself as AI-first. This tier wants AI-capable, not AI-led. Opening question: "What's the current design headcount, and what's the budget authority for this role to hire?" Pattern-break cue: If the first conversation focuses on agent-human interaction patterns rather than "how do you learn a new domain," the company evaluates closer to AI-native (Ontra). Maven's Head of Design posting is an act-now target with explicit AI-native practice-building mandate.
-
Gusto's service design language: Their Senior Staff Service Designer posting says "customers do not experience product, CX, and AI as separate things" and names poorly designed handoffs as the cost customers absorb when escalation ownership or AI failure paths are unclear, which is unusually strong external validation for your consequence-literacy framing.
-
Anthropic's design-adjacent roles: Anthropic's current openings distribute design-relevant work across web infrastructure and design systems and policy design for age-appropriate design and child safety, which means design leadership signals at Anthropic may surface outside conventional product-design postings.
-
NIST's Human-AI Configuration risks: NIST AI 600-1 defines specific risks including automation bias, overreliance, and algorithmic aversion in human-AI configurations, giving you sharper regulatory vocabulary than generic "AI trust" when speaking to regulated-growth or healthcare evaluators.
-
Intuit's VP-level AI+HI framing: Intuit's VP, Design - Virtual Expert Platform posting frames the mandate as designing for the intersection of AI capabilities and human expertise across a services ecosystem, with success measured through platform adoption, expert productivity, and experience quality, making it one of the clearest enterprise examples of Gate 3 AI modernization language.

