Nine companies cleared the rubric this cycle. One — Headway — hits Act on both dimensions (company 10/12, role 12/15), but the anchoring signal is aging and the consolidated posting's remote eligibility needs confirmation before you treat it as live. Two others scored Watch with the strongest role-level numbers on this geographic list — Gray Swan AI (13/15) and Replicant (14/15) — and both warrant outreach this week. The remaining six are Watch with blocking questions you should resolve before spending positioning time.
The non-Bay Area signal set at Director+ and Staff IC level is thin. The prior geographic scan documented this, and the pattern holds. What changed: the two highest role-level scores on this list surfaced outside the Bay Area — one remote-US out of Pittsburgh, one fully remote.
Gray Swan AI (role 13/15) and Replicant (role 14/15) carry the highest role scores on this geographic list. Both are actively recruiting. Apply through Ashby before the first slate closes.
Strong Watch — move this week
Both entries carry company scores of 9/12, one point below Act. In both cases the gap is stage risk — early capital structure at Gray Swan, organizational risk at Replicant. The role scores are the highest on this list. If you have outreach capacity this week, these two come first.
1. Gray Swan AI. AI safety evaluation platform — builds tools that test whether AI models behave safely before and after deployment. Ashby posting.
- Stage / Size / Capital: Series A, $40M raised (May 2026, led by Wing and Madrona). Design team described as "small but mighty." Headcount and ARR not publicly disclosed.
- Offices: Remote — US. Head of UX based in Pittsburgh.
- Value chain: Upstream infrastructure between model providers and enterprise deployers. Gray Swan provides the evaluation layer that determines whether a model is safe to ship. Design scope covers the full evaluation workflow: how safety researchers configure tests, interpret results, and communicate risk to enterprise buyers who are not ML experts.
- Best at: Making AI safety evaluation usable by non-researchers — translating probabilistic safety assessments into decisions enterprise customers can act on.
- Signals: Wellfound flags "Actively Hiring" and "Recruiter recently active." Interview process includes technical screen and live coding exercise. Meredith McDermott (Head of UX) and product designer Gaurav Nemade are publicly identifiable — not a founding-design-leader hire, but the team is very small. Fresh signal — actively sourced.
- Tier: AI-native. Tier Playbook: Read AI-native playbook before outreach.
- Rubric scores: Company 9/12 — AI centrality 3, Stage/equity 1 (Series A execution risk), Design influence 2, Trajectory 3. Role 13/15 — Total comp 2 ($205–255K + bonus and equity), Scope 2, Craft depth 3, AI exposure 3, Portfolio value 3.
- Portfolio match: Gray Swan's hard problem: an enterprise buyer who isn't an ML expert needs to understand why a model failed a safety test and decide what to do about it. A pass/fail score is useless to this buyer. They need to trace the failure back through model behavior to the triggering input. Carrier IQ's attribution model — tracing an insurance quote back through agent reasoning to source data — is the same traceability problem in a different domain. Lead with Carrier IQ provenance, then bridge to the Trust essay's Watch-Verify-Delegate ladder as the framework for how enterprise customers graduate from manual safety review to automated monitoring. You shipped audit trails in production AI systems at TinyFish. Candidates who can articulate a trust framework rarely have production experience building the infrastructure underneath it.
- Verdict: Watch (strong). Company score (9) falls one point short of Act on stage risk — Series A with a small team, and the equity window depends on whether they reach their next milestone. Role score (13/15) is among the highest this cycle. Apply through Ashby this week. Ask in the first conversation whether the Staff role reports to Meredith McDermott or operates alongside her — that answer tells you whether the scope ceiling is real or open. Confirm whether the team operates on Pittsburgh-aligned hours. Active recruiting means the first slate may close soon; in a team this small, early candidates set the evaluation bar.
2. Replicant. AI voice agent platform for contact centers — automates customer service calls with AI agents that handle full conversations. Ashby posting.
- Stage / Size / Capital: Growth-stage. Funding details, headcount, and ARR not publicly disclosed. Compensation not disclosed.
- Offices: Remote — US and Canadian time zones.
- Value chain: Midstream platform between enterprise contact centers and their customers. Design scope covers deployment configuration through the AI agent conversation to operator analytics.
- Best at: End-to-end AI voice conversations that resolve issues without human handoff — and the escalation design when they can't.
- Signals: Posted 28 August — 1 day old. Reports to VP of Product. Planned 3-person practice: Director + Lead/Principal (to hire) + Conversation Design Lead. 80% hands-on, 20% practice leadership. The posting says "product design has been attempted twice and under-resourced both times." Read that as both opportunity and risk — a greenfield mandate with real organizational commitment, but you need to find out what killed the previous attempts.
- Tier: AI-native. Tier Playbook: Read AI-native playbook before outreach.
- Rubric scores: Company 9/12 — AI centrality 3, Stage/equity 2, Design influence 2, Trajectory 2. Role 14/15 — Total comp 2 (undisclosed), Scope 3, Craft depth 3, AI exposure 3, Portfolio value 3.
- Portfolio match: The role has two distinct problems. The craft problem: when does the AI voice agent escalate to a human operator, and what context does the human need to take over a conversation mid-stream without the customer repeating themselves? That's a delegation boundary with a real-time constraint. The organizational problem: building a design function that has failed twice. Lead with American Red Cross ($847K disbursed, 6 systems consolidated to 1, national deployment, Product Design Director at BCG DV) — proof you can unify fragmented systems under organizational resistance. Bridge to TinyFish's 0-to-1 in 3 months as proof you ship AI products without waiting for organizational permission. In the first conversation, establish what killed the previous design efforts. If the answer is resourcing, that's solvable. If the answer is organizational will, be cautious.
- Verdict: Watch (strong). Role score (14) is the highest on this list. Company score (9) keeps the overall verdict at Watch. Apply this week through Ashby. The company's history of failed design hires means they'll move fast on candidates who demonstrate both craft and organizational resilience.
Watch — verify before investing outreach effort
Five companies with live signals, each carrying a specific blocking question. Resolve the question before you spend time on positioning or outreach.
3. Headway. Therapist-matching and insurance-navigation platform for mental healthcare. Prior coverage.
- Stage / Size / Capital: Series C, valued above $3B. Strong equity window.
- Offices: New York City. Confirm remote eligibility for the current Provider posting.
- Value chain: Downstream service connecting patients, providers, and insurance. Provider vertical covers scheduling, billing, patient communication, and AI-assisted workflows.
- Best at: Removing administrative burden from therapists — billing and credentialing infrastructure that lets providers focus on therapy.
- Signals: The earlier three-role provider cluster (Aug 3–4) has consolidated into a single Provider Staff posting, and a newer Group Practices role has appeared. Consolidation suggests the committee refined what it needs. No hiring-manager promotion of the current posting in the last 14 days. Signal age: original cluster is 26 days old; consolidation is more recent but exact date unconfirmed.
- Tier: Healthcare/Regulated. Tier Playbook: Read Healthcare/Regulated playbook before outreach.
- Rubric scores: Company 10/12 — AI centrality 2, Stage/equity 3, Design influence 2, Trajectory 3. Role 12/15 — Total comp 2 ($212–265K + equity), Scope 2, Craft depth 3, AI exposure 2, Portfolio value 3.
- Portfolio match: Headway's design problem is three-party delegation: provider, AI agent, and patient. A therapist can't review every AI-assisted scheduling decision or patient communication individually — per-action review is impossible at the scale Headway is targeting. The Trust essay's three-party framework addresses this directly. The specific production match is Brand Pulse's visibility model: what does the human see about what the AI did, and when? That maps to the provider dashboard — how does a therapist know what the system communicated to their patients on their behalf? You have the three-party framework published. No competing candidate has articulated this problem in public writing.
- Verdict: Watch. The scores clear Act on paper (company 10+, role 12+), but the anchoring signal is aging into the 22–42 day zone and remote eligibility is unconfirmed. Verify the consolidated posting is live and remote-eligible before treating this as Act. The Group Practices role may share the same delegation design problems — ask about both.
4. Hightouch. Customer data activation platform — syncs data from warehouses to marketing tools, now expanding into AI-powered creative generation. Greenhouse posting.
- Stage / Size / Capital: Series C, $2.4B valuation. Strong equity window.
- Offices: San Francisco. Geographic eligibility for this non-Bay Area list is unconfirmed — the role's remote/hybrid status was not verified in this research pass. If the role requires SF presence, it belongs on the Bay Area list.
- Value chain: Midstream platform between enterprise data warehouses and downstream marketing/sales tools. The AI Creative Tools role extends this into content generation.
- Best at: Reverse ETL — moving customer data from warehouses to operational tools without engineering. The AI creative expansion is new territory.
- Signals: 200+ applicants on LinkedIn. Design lead Elliot Dahl publicly promoted the role as a "new and unique AI design role" within the last week. Title is Senior Product Designer, but Hightouch has no formal levels — "Senior" is a placeholder, scoped to the candidate. Compensation not disclosed.
- Tier: Growth-stage platform. Tier Playbook: Read Growth-stage platform playbook before outreach.
- Rubric scores: Company 10/12 — AI centrality 2, Stage/equity 3, Design influence 2, Trajectory 3. Role 11/15 (incomplete) — Total comp ? (undisclosed), Scope 2, Craft depth 3, AI exposure 3, Portfolio value 3. If comp clears $200K: role 13/15.
- Portfolio match: The user of Hightouch's AI creative tools is thinking "I need a campaign for this segment," not "I need to engineer a prompt." Allē's dual-surface redesign (30M members, 3.2× redemption, $42 CAC, Product Design Director at BCG DV) is the match — you designed a system where the same underlying capability serves two different user mental models, and the $42 CAC proves the design worked commercially. That acquisition cost metric speaks directly to what Hightouch's AI creative tools need to deliver for their marketing buyers. You've designed for marketers at scale and built AI tools in production at TinyFish. Most AI designers have never optimized for acquisition cost; most marketing designers have never shipped AI.
- Verdict: Watch — two blocking unknowns. Comp is undisclosed. Geographic eligibility is unconfirmed. Resolve both before investing outreach effort: confirm the role is remote-eligible or hybrid-flexible, and confirm comp clears $200K through the interview process. If both clear, this upgrades toward Act. Apply through Greenhouse and reference Elliot Dahl's public promotion. 200+ applicants means the queue is already long.
5. Microsoft. AI products division — Vibe Kit and Vibe Studio design systems for Microsoft AI's product suite. Careers posting.
- Stage / Size / Capital: Public. ~220,000 employees.
- Offices: Redmond (HQ), Mountain View. 4 days/week onsite within 50 miles of office. SF-operable through Mountain View.
- Value chain: Upstream infrastructure — design systems powering interaction patterns across Microsoft's AI product suite.
- Best at: Scale. Microsoft's AI design systems reach more users than almost any other AI product surface.
- Signals: Posted 18 August — 11 days old. $188K–$304K for SF/Bay Area. Principal IC level.
- Tier: Enterprise platform. Tier Playbook: Read Enterprise platform playbook before outreach.
- Rubric scores: Company 9/12 — AI centrality 3, Stage/equity 1, Design influence 2, Trajectory 3. Role 12/15 — Total comp 2, Scope 2, Craft depth 3, AI exposure 3, Portfolio value 2.
- Portfolio match: The design systems problem here is codifying interaction patterns for AI products that don't yet have stable conventions — loading states for inference, confidence displays, correction flows, delegation boundaries. The Trust essay's five handoffs framework maps as a reusable component pattern: Watch-Verify-Delegate encoded as system-level primitives, not one-off flows. You've articulated the framework and shipped the production patterns at TinyFish. Design systems candidates who can build component libraries rarely have production experience defining AI-specific interaction patterns from scratch.
- Verdict: Watch. AI exposure is high. Public company equity and 4-day onsite are real constraints. Apply through Microsoft careers if the design-systems-for-AI mandate is compelling enough to offset the equity limitation. Posting is 11 days old; the first review cycle may be approaching.
6. Outreach. Sales engagement platform with conversational intelligence — analyzes sales calls using AI. Seattle HQ.
- Stage / Size / Capital: Late-stage private. ~1,000 employees.
- Offices: Seattle (HQ). US role location not independently verified — confirm whether Seattle-based or remote-US.
- Value chain: Midstream platform between CRM systems and sales execution. Conversational Intelligence adds AI-powered call analysis.
- Best at: Sales workflow automation at enterprise scale — embedded in the daily workflow of enterprise sales teams.
- Signals: Senior Director of Product Design Jan Srutek publicly promoted a Staff Product Designer opening within the last week, but the promoted role is Prague. Active design recruiting confirmed; recruiting intensity for the US role specifically is not.
- Tier: Growth-stage platform. Tier Playbook: Read Growth-stage platform playbook before outreach.
- Rubric scores: Company 8/12 — AI centrality 2, Stage/equity 2, Design influence 2, Trajectory 2. Role 11/15 — Total comp 2, Scope 2, Craft depth 3, AI exposure 2, Portfolio value 2.
- Portfolio match: The design problem: making AI-generated conversation insights actionable for salespeople who are focused on closing, not analyzing transcripts. Brand Pulse (Agentic Labs) is the closest analog — you designed how AI-generated intelligence surfaces in a workflow where the user's primary task is something else entirely. Brand Pulse's visibility model (what changed, why it matters, what to do) maps to how a salesperson should consume conversation intelligence. You've designed intelligence surfaces at this level and have enterprise platform experience at Alibaba.com's scale — a combination few candidates bring together.
- Verdict: Watch. Confirm US role location before investing outreach effort. If remote-US, apply and reference Jan Srutek's public hiring activity.
7. Babylist. Registry and commerce platform for new parents — product discovery, registry management, and AI-powered recommendations. Greenhouse posting.
- Stage / Size / Capital: Growth-stage. Headquartered in Emeryville, CA — Bay Area by any geographic definition. Babylist appears on this list because the company is remote-first across US and Canada, the role is US-wide, and the team meets in person only twice annually. Included under the remote-first branch.
- Offices: Emeryville (HQ), remote-first US/Canada.
- Value chain: Downstream consumer application. The AI Builder Director role suggests the company is adding AI-powered recommendation and personalization capabilities.
- Best at: Universal registry — parents can add items from any retailer, creating a data advantage over single-retailer registries.
- Signals: Greenhouse posting is live. Posted 3 August — 26 days old, in the Watch decay window. The "AI Builder" title signals the company wants someone who can prototype and ship AI features, not manage designers. ~$245–306K.
- Tier: Growth-stage platform. Tier Playbook: Read Growth-stage platform playbook before outreach.
- Rubric scores: Company 9/12 — AI centrality 2, Stage/equity 2, Design influence 3, Trajectory 2. Role 11/15 — Total comp 2 (~$245–306K), Scope 3, Craft depth 2, AI exposure 2, Portfolio value 2.
- Portfolio match: A bad product recommendation here erodes parental confidence in the platform — the cost is trust, not dollars. Equinox+ (0 to MVP in 3 months, 5 brands, 4.8 stars, Product Design Director at BCG DV) is the match: you built a consumer product from zero under time pressure where quality perception was the primary success metric. The "AI Builder" title maps to your TinyFish trajectory — you moved to product specifically to build AI-natively from zero and shipped the agentic platform in 3 months.
- Verdict: Watch. Signal is 26 days old and aging. AI centrality (2) and trajectory (2) keep this at Watch regardless. Ask early: does the Director own the AI product roadmap or execute against someone else's? That answer determines whether the scope is real.
Watch — marginal
Both entries sit at or near the rubric threshold floor. Neither has a time-sensitive action attached. They stay on the list because the role scores qualify, but they should receive outreach effort only after the entries above are in motion.
8. Coinbase. Cryptocurrency exchange and financial services platform. Financial Services Lead posting.
- Stage / Size / Capital: Public (COIN). ~3,500 employees.
- Offices: Remote-first — USA. Periodic in-person surges.
- Value chain: Vertically integrated exchange, custody, and financial services. The Financial Services Lead sits at the downstream application layer where AI-assisted products meet retail users.
- Best at: Regulatory navigation in crypto — institutional credibility that competitors lack.
- Signals: Role is live on Coinbase careers, Remote — USA. Posted 8 August — 21 days old, crossing from hot to Watch on August 30. A second role (Experience and Engagement) is also live but scores below the role threshold at 9/15.
- Tier: Enterprise platform. Tier Playbook: Read Enterprise platform playbook before outreach.
- Rubric scores: Company 7/12 — AI centrality 2, Stage/equity 1, Design influence 2, Trajectory 2. Role 10/15 — Total comp 2, Scope 2, Craft depth 2, AI exposure 2, Portfolio value 2.
- Portfolio match: AI-assisted financial decision-making for retail users who need to trust the system's recommendations with their money. Alibaba.com's B2B platform redesign (Head of Design and Research, North America; $50B+ GMV; +20% daily transactions; +2.2pt NPS) is the match — you designed trust in a transactional context where the stakes are financial and the user needs confidence before committing. Few candidates combine AI production depth with financial trust design at that transaction volume.
- Verdict: Watch — marginal. Company score (7) is below the 8+ threshold; the role score carries this entry. Public company equity limits upside. Posting crosses into Watch decay tomorrow. Pursue only if the financial services AI mandate extends into product strategy.
9. Redesign Health. Health tech venture studio — builds, launches, and scales healthcare companies from within a studio model. Ashby posting.
- Stage / Size / Capital: Venture studio (not a traditional startup stage). NYC headquarters.
- Offices: New York City (HQ), Los Angeles. Role is hybrid from NYC or SF Bay Area.
- Value chain: Upstream — creates healthcare companies, sitting above the individual ventures it launches. Design scope spans multiple portfolio companies at different stages and regulatory contexts.
- Best at: Systematic healthcare company creation — applying learnings across portfolio companies in a way individual startups can't.
- Signals: Ashby posting is live. Hybrid NYC/SF. Posted 13 August — 16 days old, still in the hot window.
- Tier: Healthcare/Regulated. Tier Playbook: Read Healthcare/Regulated playbook before outreach.
- Rubric scores: Company 8/12 — AI centrality 2, Stage/equity 2, Design influence 2, Trajectory 2. Role 10/15 — Total comp 2 (undisclosed), Scope 2, Craft depth 2, AI exposure 2, Portfolio value 2.
- Portfolio match: The design problem is designing across multiple healthcare ventures simultaneously, each with different regulatory constraints, user populations, and maturity levels. Your BCG DV track record is the direct analog — you designed across multiple ventures (Thermo Fisher, Red Cross, Equinox+, Allē) at different stages and in different regulated domains. Thermo Fisher mySupply (pharma, 6 partners, 100% adoption, 12 months, Product Design Director at BCG DV) demonstrates you can ship in a regulated healthcare-adjacent domain with multiple stakeholders. Single-venture candidates can show depth in one domain; your BCG DV portfolio proves you can operate across the pattern itself — multiple regulated domains, multiple stages, simultaneously.
- Verdict: Watch — marginal. Company and role scores sit at the threshold floor. The posting is 16 days old and still hot, so timing isn't the constraint — the question is whether the scope justifies the investment. The venture studio model is unusual. Scope could be very broad or very narrow depending on whether you own design across the portfolio or within one venture. Ask that question before investing further.
Window status changes
| Company | Status | Key detail |
|---|---|---|
| Gray Swan AI | New → Watch (strong) | Highest AI exposure on this list. Series A stage risk keeps company at 9/12. |
| Replicant | New → Watch (strong) | Highest role score (14/15) on this list. Posted yesterday. |
| Headway | Remains Watch | Scores clear Act on paper, but signal aging (26 days) and remote eligibility unconfirmed. |
| Hightouch | Remains Watch | Comp unknown; geographic eligibility unconfirmed. 200+ applicants. |
| Microsoft | New → Watch | Posted 11 days ago. Principal IC, AI design systems. |
| Outreach | Remains Watch | US role recruiting intensity unconfirmed. |
| Babylist | New → Watch | Posted 26 days ago, in Watch decay window. |
| Coinbase | New → Watch (marginal) | Posted 21 days ago — crosses to Watch decay August 30. |
| Redesign Health | New → Watch (marginal) | Posted 16 days ago, still hot. |
| Brex | Watch → Ignore-until-trigger | 51 days old. Posting remains live but past Watch threshold; Capital One acquisition eliminates pre-IPO equity premise. See Capital One battlecard. Trigger: new Brex design posting or repost with fresh language. |
Ignore-until-trigger
- Affirm (Staff Product Designer, Agentic Experiences): 25 days old, remote-US, relevant — but past the 21-day hot window. Trigger: repost or fresh recruiting promotion.
- Vetcove (Staff Product Designer): Fully remote, live posting, but AI centrality is 1 and comp range ($120–210K) may not clear the floor. Trigger: AI-specific product launch or confirmed comp above $200K.
- Boulevard (Staff Product Designer): Remote-US, but disclosed comp ($140–180K) is below the $200K floor. Trigger: leveling change or comp adjustment.
- Givebutter (Director, Product Design): Remote, $200–225K, but AI centrality is 1 and the role is management-heavy with a VP above. Trigger: AI-specific product mandate.
- Coinbase Experience & Engagement: Remote-US, role score (9/15) falls below Watch threshold. AI exposure is peripheral. Trigger: role language update adding AI-specific mandate.
- Replicant's two failed attempts: The posting's admission that "product design has been attempted twice and under-resourced both times" is unusually candid — the Ashby listing may tell you more about what killed the function than the hiring manager will in a first call.
- Brex's orphaned recruiting route: Capital One completed the Brex acquisition in April, but Brex's Staff AI posting remains live through its own careers page with no visible connection to Capital One's design org — watch for a repost under Capital One branding or a fresh Brex job ID as the integration matures.
- Gusto's Watch window expires September 3: The Unified Service Platform Head of Design role crosses from Watch to Ignore-until-trigger in five days, and its mandate — uncertainty, graceful failure, human escalation — is among the closest matches to your Trust framework on any active list.
- Gray Swan's coding screen: The live coding exercise is a material candidacy gate worth preparing for before the first conversation, since the team is small enough that early candidates define the evaluation bar.

