SOURCES · AI-NATIVE DESIGN

Read the field.
In the order I would.

Canonical readings for AI-native design, sequenced as a path. Every note is mine, written after reading the piece. Suggest what is missing: ayodhyarammohanthy@gmail.com.

OPEN AI-NATIVE DESIGN SYSTEM02 / Learn
01

Foundations

What AI-native design is, and why it is not a skin on old software.

1.1

Eight Dimensions of AI-Native Design

Adam Kinney · 2026

The cleanest break with the old frame: components become conversations, states become spectrums. My demos stream and blend instead of flipping for exactly this reason.

1.2

People + AI Guidebook

Google PAIR · 2023

The textbook. User needs, mental models, trust - less fashionable than the agent discourse and more useful when you actually ship.

02

Interfaces

How the surface changes when software acts on your behalf.

2.1

The Agent is the Interface

Aaron Sagray · 2026

Users evaluate an agent the way they evaluate a colleague - reliable, restrained, honest about what it did. Restraint as a design material. That one idea rewires a roadmap.

2.3

AI UX Patterns for AI-Native Products

Lazarev.agency · 2026

A pattern playbook with the right spine: show the reasoning, admit the uncertainty, hand control back. Closest to how I grade my own demos.

03

Agents

Your next user may not be human. Design for both readers.

3.2

Govern the fleet as a product surface

Karo Zieminski & ToxSec · 2026

Once a team has several agents, the design object is the fleet. Give every agent an identity, owner, purpose, access boundary, review cadence and off switch. Make delegation preserve the initiating identity and pass only task-scoped authority with its own expiry. The control plane should answer what exists, who owns it, what it can reach, who it can call, what it did and how to stop it. Governance works when those answers are part of the build, not a document beside it.

3.3

Bind approval to the task revision

Shengcheng Yu, Chunrong Fang & Zhenyu Chen · 2026

An approval is not durable when the task keeps moving. Bind every effect to the exact task revision, referenced objects, role authority, operational controller and outcome evidence. When a person edits the shared application, a policy changes or a provider response arrives late, invalidate only the claims whose premises changed and ask again where needed. A correct component call can still compose into the wrong effect; continuous assurance belongs at the intent-to-effect boundary. This is a strong specification with worked examples, not yet evidence of production benefit.

3.5

Size governance to model capacity

Michael Ray Johnson & Linda Naimi · 2026

Governance consumes reasoning capacity. A full operating procedure can help a frontier model and stall a weaker one; a single verification rule may outperform the larger scaffold. The stronger move is an answer-blind definition of done derived from the task's policy and published standards, not one generic process for every workflow. Choose the model and the governance burden together, then measure each workflow separately. The results are exploratory - partial implementation, small cells and single trials - but the boundary condition is the useful decision.

04

Case studies

How shipped AI products actually got designed - told by the teams that built them.

4.1

Design the interface for both readers

Jin Gao / Affora · 2026

Affora reverses the usual question. Instead of asking how an agent can cope with an interface built for people, it asks which interaction meanings the interface must preserve for both readers. The operating rule is substrate invariant, skin variable: keep controls, names, state, choices and outcomes explicit from component to site, while visual identity remains free to change. The checks are a conformance floor, not proof of task success.

4.2

Generate against approved objectives

Gal Elidan & Yael Haramaty / Google Research · 2026

The strongest generative UI pipeline starts before generation and ends after it. Teachers approve the objectives; levels, hints and feedback derive from those objectives; agentic checks test pedagogy, mechanics, solvability and visual noise; and a teacher decides whether the result can enter the library. Dynamic UI becomes credible when authority, evaluation and publication gates are designed as one system.

4.3

Make long-running work legible and correctable

Xuan Zhao, Jiwoong Sohn, Qinyue Zheng & Michael Moor · 2026

Supervision should shorten the path from 'what happened?' to a safe correction. Separate activity, elapsed time, tool traffic, workspace artifacts and subagent traces; then let people interrupt, revise the task or hand the same run to a stronger model without rebuilding context. AgentGUI's small study suggests that structured trajectories improve both lookup speed and accuracy, while its automated audit shows the largest completion gains on workers that can do most, but not all, of the job. The evidence is promising, not broad: eight participants and one synthetic steering task.

4.4

Separate exploration from implementation

Jian Zhao et al. · 2026

Propose several structured design directions first; choose one explicitly; then hold downstream code generation fixed. This architecture makes exploration a controllable product decision instead of adding randomness to both aesthetics and syntax. The evaluation lesson matters just as much: measure coverage, rendered difference, adherence, execution, accessibility and human preference separately, because broader variation did not consistently mean better judged quality.

4.5

One design truth, three agent interfaces

Sudharsanam Narasimhan, Kylor Hall & Michael Abrahamian / Atlassian · 2026

The important move is not adding a CLI. It is keeping one typed design truth behind skill, MCP and CLI, then judging the whole task instead of the payload. Atlassian held agent, model and tasks constant, read transcripts as well as metrics, and let observed behavior change routing, batching and follow-up guidance. Distribution stays flexible because governance does not fork.

4.6

Designing Grok Bot for a world of persistent agents

SpaceXAI · 2026

I keep returning to the subtraction test: does this help someone delegate, or give them another surface to manage? Grok Bot uses that question to govern the whole system - persistent Bots instead of disposable chats, presence that shows state, graduated status/preview/takeover supervision, role-bound context, structured responses and routines that keep work moving. The useful lesson is coherence: primitives, coordination and autonomy all answer to the same product principle.

4.7

How Sierra’s design team keeps up with an engineering org 20x its size

Designer Fund / Sierra · 2026

Sierra’s leverage comes from refusing to become the 51st engineer. A tiny design team supports an engineering org nearly 20x its size by moving from individual polish to the systems, patterns and components that raise everyone’s floor, while keeping the judgment-building parts of craft deliberately human. Prototype against real data early; automate the infrastructure, not the taste.

4.8

Figma to Code at Scale: Building with Strands Agents

Tim Moreton / Strands Agents · 2026

Amazon Ads turns Figma into production React across 20+ marketplaces with a governed agent team: specialists extract the design system, generate code, test overlap and overflow, research local constraints and measure quality. The leadership pattern is allocation by consequence - shared rules for every agent, the strongest model on code generation, deterministic checks on the rendered result, and measurement as a first-class role.

4.9

Inside the AI stack of Together AI’s product team

Aakash Gupta / Together AI · 2026

The leadership asset is shared context, not a clever individual setup. Together AI gives every product area the same current context, encodes repeatable research and PRD work as team skills, and closes the loop with agent journeys that find product gaps, leave transcripts and re-verify the fix. AI-native product operations become collective memory plus continuous evidence.

4.10

Teaching agents product design at Vercel

Vercel · 2026

Design governance becomes code: a repo skill routes judgment, linters enforce deterministic rules, evals test the guidance, and a weekly collector/judge/review-packet loop keeps human approval in charge. The operating metric matters: agents failed to invoke an available skill in 56% of Next.js evals, so trigger reliability must be tested separately from rule-following.

4.11

How Stripe built Kai: reusable skills and enterprise governance

Sharadh Krishnamurthy / How I AI · 2026

The leadership pattern is appropriate friction at scale: a named owner sets the models, skills, backends and approval rules for each project, while telemetry and deprecation keep a 2,000-skill ecosystem usable. Governance is product infrastructure, not a warning added at the end.

4.12

How we made our design system bilingual

David Cai / BILL · 2026

One design truth, two delivery shapes: Storybook for people and structured CLI/MCP plus workflow skills for agents. The leadership move is the hard stop - if neither source can verify the system, the agent does not improvise frontend code.

4.13

Building products for humans and agents

Tamar Yehoshua / Atlassian · 2026

Treating agents as direct users makes design an operating-system problem: give them organizational context, keep a named human accountable across agent chains, and make inference spend a product decision. That is the leadership layer above any single AI interface.

4.14

When agents own the middle, design leads at the edges

John Moriarty / Mind the Product · 2026

The useful leadership shift is not "designers code more." Design moves upstream into roadmap judgment and downstream into the shipped interface, while agents compress the execution middle. That is the operating shape I want this portfolio to prove.

4.15

How we built Linear Agent

Matthijs Wolting · Linear · 2026

Linear designed boundaries instead of paths: system prompt, tool design, run scope and the harness underneath are where the behaviour is shaped. Script it too tightly and you dilute the flexibility that makes an agent useful - the agent-design trade in one line.

4.16

Fin over email: How we built a multichannel AI agent

Intercom Product & Design · 2026

Intercom rebuilt Fin for email from first principles instead of porting the chat surface: an Intermission problem doc, an alpha with no bells, then iteration. The channel changes the contract - a case study in not copying your own pattern blindly.

4.17

Building a Robust Harness for an Agent in Production

Kamie Shami-Schnitzer · monday.com · 2026

The bowling-bumper model of agent design: walk the flow, ask what can go wrong at each phase, build the guard there. The alias trick is the gem - swap raw IDs for aliases and every hallucinated citation becomes detectable instead of silent.

05

Craft

Pattern-level playbooks for the daily work.

5.1

Nad Chishtie: Lovable’s Design System for Agents

Tommy Geoco / Lovable · 2026

Lovable treats AI-native design as an org and governance problem, not a tool rollout. End-to-end ownership expands when everyone can build, while half the design system is written for agents. The useful failure is the leadership lesson: a two-week background-agent push did not hold, so the team kept the deterministic linters that did. Automate where the control is testable; kill the fashionable mechanism when it is not.

5.2

Crafting the last mile of delight

Joel Lewenstein / IDEO · 2026

When build time collapses, design leadership moves to the two ends: deciding what deserves to exist and protecting the last mile of quality. Anthropic’s 30-person design, research and content team makes that operating shift concrete - code-first prototypes in the middle, stronger judgment before and after, and agents enforcing routine content standards in production.

5.3

Org Design for an AI-Native Design Org

Aaron Sagray · 2026

Three accountable functions turn AI-native design into an operating model: skill librarian, eval owner and AgentOps. The draft RACI, fractional starting capacity, conversion thresholds, artifacts and cadences make the decision rights concrete enough to staff before they become full-time roles.

5.4

How to lead design teams through the AI era

Jen Dunnam / Figma · 2026

The leadership job is not to add panic to a team already moving fast. Steady the work, sharpen the principles, hire for dissent, and name which kind of speed matters before asking for more of it.

Systems I return to

8 bookmarked systems, shared with the ayodhya. design system.

S.1

Backpack

Skyscanner · Multi-platform coherence

For keeping one product language coherent across platforms without forcing sameness.

S.2

Wanda

Wonderflow · Open, inspectable craft

For pairing a precise token system with component guidance you can inspect, test, and use.

S.7

Polaris

Shopify · Product guidance

For product guidance that gives a component judgment, not only anatomy.

AGENT DESIGN SKILLS · 11 STUDIED

Agent design skills.

The skills, guardrails, and workspaces I study to keep agent-made interfaces specific, usable, and true to their product.

Featured review stack

7

UI prompting skillactive

Design-first UI prompting

Meng To · MIT
  • type direction
  • image and crop
  • rhythm
  • variation
Why it is here

For specifying type, image, crop, rhythm, and variation before the model reaches for a template.

Prompt like a system, not a moodboard.
Open source ↗

Interface audit skillactive

Hallmark

Nutlope · MIT
  • audit
  • redesign
  • review
  • improve
Why it is here

For giving an agent clear verbs to inspect and improve a rendered interface.

Critique must lead to a change
Open source ↗

Agent skill and commandsactive

Impeccable

Paul Bakaus · Apache-2.0
  • 1 skill
  • 23 commands
  • 61 detector rules
Why it is here

For turning aesthetic critique into explicit commands and detector rules.

Taste needs an operational vocabulary
Open source ↗

Interaction skill collectionactive

Skills for Designers and Engineers

Emil Kowalski · MIT
  • emil-design-eng
  • animate
  • review animations
  • improve animations
  • animation vocabulary
Why it is here

For translating interaction craft into a shared vocabulary designers, engineers, and agents can act on.

Polish needs shared language
Open source ↗Followed creator: Emil Kowalski ↓
1 relatedSource repository ↗

Aesthetic guardrailv2 experimental

Taste Skill

Leonxlnx · MIT
  • taste constraints
  • anti-slop critique
Why it is here

For making aesthetic constraints visible before generic interface habits take over.

Guardrails before decoration
Open source ↗
1 relatedSource repository ↗

UI quality skill setactive

UI Skills

Ibelick · MIT
  • baseline-ui
  • improve-ui
  • fixing-accessibility
Why it is here

For turning spacing, hierarchy, interaction, and accessibility into a repeatable finishing pass.

Taste needs a quality floor.
Open source ↗
1 relatedSource repository ↗

Browser quality skill setactive

Web Quality Skills

Addy Osmani · MIT
  • performance
  • accessibility
  • browser evidence
Why it is here

For proving the finished interface works in the browser, not only in the design review.

A beautiful screen still has to survive evidence.
Open source ↗

Foundations & workspaces

2

Agent skill

Frontend Design

Anthropic
  • frontend design
  • visual judgment
Why it is here

The baseline that made visual judgment portable inside a coding-agent workflow.

Portable visual judgment
Open source ↗

Local-first design workspaceactive

OpenDesign

nexu-io · Apache-2.0
  • design context
  • agent collaboration
Why it is here

For keeping product design context local, legible, and close to the work.

Context belongs beside the product
Open source ↗
1 relatedSource repository ↗

Reference library

2

Design reasoning skillv2.0

UI UX Pro Max

Next Level Builder · MIT
  • 192 reasoning rules
  • 119 UX guidelines
  • 74 font pairings
Why it is here

For exposing a wide design reasoning surface an agent can consult before it renders.

Breadth still needs product judgment
Open source ↗

PEOPLE IN DESIGN · 1 FOLLOWED

People I follow in design.

Designers and design engineers whose work sharpens how I think about systems, interaction, and craft.

Emil Kowalski

Design Engineer · Linear
  • motion
  • interaction
  • design engineering

For the care he brings to how interfaces move, feel, and behave - and for making that judgment legible to designers and engineers.

WHAT TO STUDY

Purposeful motion, timing and easing, interaction detail, and the bridge between design judgment and production code.

Then touch it

Reading holds when it meets running software. The patterns, the map and the starter kit are one click away.