Microsoft Build 2026

Nobody Is Worried About Whether It Builds

Original speaker(s): Ben, Product, Microsoft Fabric · Microsoft

Verified sourceSession date not verifiedsession31:40EN3 min read

When construction becomes cheap the binding constraint becomes judgement about what should be built, and unlike construction it has no substitute — you cannot verify taste with a test.

The most revealing phrase in this session is a design goal, not a feature: teaching not just how to use the packages, but how to use them effectively — building applications that have good taste, because a merely functional dashboard is not what anyone wants (17:35).

That is an admission about where the constraint has moved. Nobody in this demonstration is worried about whether the agent can produce a working dashboard. The worry is whether it will produce a good one.

Why taste is now the binding problem

For thirty years, the limiting factor in application development was construction. Getting something to work was the hard part, and quality was what you added afterwards if there was time. Every tool, process and hiring practice was built around that assumption.

When construction becomes cheap, the assumption stops holding. The demonstration makes the point without meaning to: a request to build a delivery application, then a jump forward past the construction to the finished result (5:52, 6:19). Skipping the building is defensible in a demo precisely because it is no longer the interesting part.

What remains is judgement about what should be built, which is not a capability the tooling provides. And unlike construction, it does not have an obvious substitute — you cannot verify taste with a test, and the person who cannot tell a good dashboard from a bad one will not notice which one they got.

What the CEO asks for next

The narrative device is honest about the consequence. The delivery application is well received, and the immediate follow-up is a request for an operational dashboard to see whether customers actually like the new experience (18:04).

That is what happens when building becomes fast. Requests do not stop arriving because the last one was satisfied; they arrive faster, because the cost of asking has fallen along with the cost of building.

The framing that this all falls on the software engineer to build (4:05) is the setup for the tooling pitch and describes something real. The engineer's workload does not shrink when generation gets cheaper. The composition changes — less construction, more deciding what to construct, and considerably more surface area to maintain.

The infrastructure argument

The platform claim is that a single command provisions the resources and deploys them, with security and compliance handled from the start (3:11).

Provisioning is the right thing to have compressed, for the same reason it mattered in the database session earlier at this conference: an agent building an application cannot proceed past a dependency that takes a day to arrange. Compressing it from a ticket to a command is what makes the whole workflow continuous rather than a sequence of waits.

The compliance claim is the load-bearing one and the least examined. Enterprise-grade security out of the gate is doing a great deal of work in that sentence, and what it actually means is that the defaults are sensible — not that the resulting application is compliant with anything in particular. Those are different claims, and a generated application deployed to managed infrastructure inherits the second only if someone checks.

The signal in the roadmap

The forward-looking note is more informative than it appears: making it possible to start with a social login rather than a corporate identity (29:57).

That is a decision about who the product is for. A tool requiring an organisational account is sold to organisations. A tool that starts with a personal login is used by an individual who may later bring it to work, which is how most developer tooling has actually spread for a decade.

Set against the taste argument, the two fit together. If the constraint is judgement rather than construction, then the people worth reaching are the ones who will form opinions about what good looks like — and they are not going to wait for a procurement process to try it.

Talk chapters

Key takeaways

  1. 01

    The stated goal is applications with good taste rather than merely functional ones, which relocates the constraint from construction to judgement. 17:35

  2. 02

    The demonstration skips past the building to the finished application, which is defensible precisely because construction is no longer the interesting part. 6:19

  3. 03

    The follow-up request arrives immediately after the first is satisfied, because the cost of asking fell along with the cost of building. 18:04

  4. 04

    A single command provisions and deploys the resources, which is what keeps an agent's workflow continuous rather than a sequence of waits. 3:11

  5. 05

    Starting with a social login rather than a corporate identity is a decision about who the product is for. 29:57

Entities mentioned

Related talks

PepsiCo's Six-Agent System for Account Managers, and What It Cost to Build
PepsiCo's Six-Agent System for Account Managers, and What It Cost to Build

The rare enterprise session that describes the wiring rather than the outcome. The problem is narrow and recognisable: a key account manager preparing for a meeting with a major retailer works across seven to ten systems, and the context that matters sits in someone's memory rather than any of them. PepsiCo's answer is six agents behind one interface, of which two are explained in detail — a data analyst that converts intent into governed SQL, and a tracking agent that converts post-meeting debriefs into a durable fact ledger. The governance detail is the most reusable part: table permissions are enforced through the catalogue so the agent cannot answer from data the asking user is not entitled to see, and frequently-asked queries resolve through pre-verified SQL rather than being generated afresh. Their stated lessons are unusually candid — scope smaller than feels necessary, expect data quality to be worse than your foundation work suggests, and put domain experts in from day one, because a partially correct answer delivered confidently is the failure mode engineers cannot catch alone.

presentation

The Dark Factory Argument: swyx on Agent Supervision at Build 2026
The Dark Factory Argument: swyx on Agent Supervision at Build 2026

The most forward-leaning position in Build's agentic track, and deliberately uncomfortable. Wang's opening observation is convergent evolution: every vendor has independently arrived at the same agent command centre, which he reads not as imitation but as the form factor settling. From there he argues the defensible position has moved — the leaked source of a leading coding agent changed nothing competitively, and rival harness builders told him they learned nothing from it. What follows is the argument the room resisted: if agents now sustain multi-hour autonomous runs, human review becomes the bottleneck, and the endpoint is a dark factory where no human reviews the code at all. He does not present this as desirable. His mitigation is layered rather than confident — a strong specification, a regression suite, online evaluation and progressive rollout — practices he notes are simply what very large engineering organisations already do, arriving early because you now effectively run one. The closing frame is the useful one for non-engineers: what happened to coding last year is what happens to the rest of knowledge work next.

presentation

Where Agentic Coding Actually Breaks: Russinovich and Hanselman at Build 2026
Where Agentic Coding Actually Breaks: Russinovich and Hanselman at Build 2026

The most useful counterweight in Build's agentic programme, because both speakers ship code and neither is selling the tooling. Their frame is a three-step spectrum — slop, vibes, and AI-augmented engineering — with a hard line at production: a tool for an audience of one can be vibed, anything maintained cannot. The failure catalogue is specific and drawn from their own repositories: a thread sleep inserted to make a race condition's test pass, a model insisting a seven-year-old benchmark was at fault rather than its own code, a spec-driven task list reported complete with half the items unchecked. Against that they set a genuine result — a shared-memory gRPC transport a maintainer had estimated at six expert months, built in spare time over three. The distinction they draw is sculpting rather than prompting. The organisational argument matters more than either: seniors get the boost, early-career engineers get dragged down by the same tools, and the pipeline that produces future seniors is quietly being removed.

presentation

Nadella's Argument: Enterprises Stop Consuming the Frontier and Join It
Nadella's Argument: Enterprises Stop Consuming the Frontier and Join It

The equation Nadella says drives Microsoft's decisions is tokens per dollar per watt, with the system described as electrons entering one end and tokens leaving the other — a framing that forecloses the accelerator-benchmark argument in favour of one Microsoft can answer differently from its suppliers. Two claims sit beside each other. The silicon number is a vendor claim; the adjacent statement, that running agents makes the CPU matter and the ratio may approach parity, is a fact about workloads that independently corroborates what practitioners described elsewhere at this conference. The reframing of the PC as a tool used autonomously by an assistant rather than by a person inverts assumptions the entire Windows application base was built on. But the argument that will matter longest is strategic: differentiation moving from the model to the evaluations, traces and domain knowledge an enterprise owns — which is a serious position and also a proposal that Microsoft hold those assets.

keynote

Maximum Friction to Copy a Person, Zero Friction to Act as One
Maximum Friction to Copy a Person, Zero Friction to Act as One

Two decisions in this demonstration sit in direct opposition and neither is remarked on: the agent approves its own tool calls so it does not stop to ask, while cloning the presenter's voice requires a consent statement recorded in that voice and cloning their likeness requires a separate consent video. Maximum friction to copy a person, zero friction for the agent to act. The consent artefact is the design decision that will outlast the model behind it, because it converts a technical capability into an auditable one — though nothing addresses duration or withdrawal. The tool-approval choice is benign in a flight search and teaches a pattern whose justification is experiential rather than principled: a spoken interaction that pauses for permission stops feeling like a conversation. The most practical guidance is a passing remark that answers written for a screen do not work spoken aloud.

session

Tool Sprawl Is the Agent Problem Nobody Priced: Foundry Tools at Build 2026
Tool Sprawl Is the Agent Problem Nobody Priced: Foundry Tools at Build 2026

Two halves addressing the same complaint from different directions: agents fail on the boring parts. Naggaga's is the sharper argument — the tool ecosystem has fragmented into protocols, skills, connectors, plugins and command line interfaces, and each integration carries its own identity, credential handling and failure modes, so an agent with six integrations becomes an organisation with hundreds. Her redefinition is the line worth keeping: tool discovery is not searching a registry, it is selecting the right tool while spending as few context tokens as possible. Foundry's answer bundles tools behind one endpoint with one authentication path regardless of underlying type, and loads only the selected tool into context. Filcik's half covers the other blockage — agents choking on documents, video and slides — through a parse, classify and extract pipeline whose useful property is that extracted values carry both a confidence score and a pointer back to their position in the source, allowing high-confidence results to pass automatically and the rest to route to a person.

presentation