AI Product Builder — turning frontier models into products people trust. Zero to one, shipped to hundreds of millions.
Varshine builds products at platform shifts. Founding PM of Microsoft’s first frontier reasoning agent, Researcher. Now redefining how products get built in the AI era — across Enterprise and Consumer tech.
▸query · “who is varshine sridharan?”▸plan · scan web + enterprise context, then synthesize✓found · founding PM of Microsoft’s Researcher agent — 0→1 to ~50M users in 11 months✓found · built Model Council — frontier models deliberate, her judge prompt decides✓found · led the first joint Microsoft × OpenAI deployment safety review◆synthesis · Founding PM for Researcher agent in Microsoft Copilot, Microsoft’s first frontier reasoning agent. An AI builder who loves to operate at startup velocity, evaluate models, prototype and pitch new ideas, write prompts to production, build evals and tune models, and build side projects during weekends.→full report below
The Work
Enterprise scale to the frontier.
01
Products Used by Hundreds of Millions
Researcher Agent. Microsoft 365 Copilot. Dynamics 365 Customer Service. Twilio.
From enterprise data quality to a frontier reasoning agent to everyday work surfaces like Copilot — shipped at Microsoft scale.
02
Zero to One AI Agents in Production
Microsoft Copilot — Researcher Agent: 0→1 MVP to production in 5 months. Model Council: hack week to production.
New products stood up from scratch — Microsoft’s first agentic AI system, grown to millions of users. Built the playbook — behavior specs, quality bars, launch rigor — where none existed.
03
AI-Native Product Craft
Researcher: Microsoft’s first frontier reasoning agent. Model Council: multi-model deliberation, in production.
System prompts as product surface. Evals as the steering wheel. I write the production prompts, build the LLM-as-judge frameworks that score them, and make ship calls on evidence — not vibes.
Work
Frontier agents, shipped.
Nine years building products; six building 0-to-1 AI, agentic, ML, and data products. Hands-on builder at the intersection of product, UX, research, engineering, safety, and GTM. Founding PM on Microsoft’s first agentic AI system — now one of the most-used deep research products in the world. Before that: AI for customer service, platform products, and enterprise solutions.
Researcher · Founding PM
Microsoft’s First Frontier Reasoning Agent
Founded and launched Researcher 0→1 in Microsoft Copilot — a long-horizon agent for multi-step deep research across web and enterprise data, built with OpenAI Research. Own product strategy, UX, core quality, and AI safety, including the first joint Microsoft × OpenAI Deployment Safety Board review. Now one of the most-used enterprise deep research agents in the world.
Researcher in Microsoft Copilot · model picker featuring Model Council
Multi-Model Deliberation — Model Council, Critique & Claude in Copilot
Lead Researcher’s multi-model strategy: launched model choice by bringing Claude into Researcher, shipped the Critique capability, and envisioned and built Model Council — multiple frontier models reason over the same task; a judge model analyzes their work and synthesizes the final answer. I prototyped it as a hack project, then shipped it to production — authoring the judge’s system prompt myself and tuning it against deterministic and LLM-as-judge metrics that define what a high-quality response tastes like. Pioneered PM-led prompt tuning at scale: ideation → prototype → prompt iteration → automated LLM-as-judge + human scoring → production readiness.
Built and operationalized the agent-quality flywheel — eval sets, rubric graders, golden outputs, and calibrated LLM-as-judge scoring — and scaled the expert human-evaluation program into a post-training data pipeline. Shipped Microsoft’s first in-house post-trained deep research model, deployed across public and Gov clouds.
Beyond the product
01
Forward-Deployed
Drove the Gates Foundation’s adoption of Researcher for vaccine research, operating hands-on as a forward-deployed engineer.
02
Claude in Copilot
Launched the Claude model family in Microsoft Copilot as part of leading Microsoft’s multi-model strategy.
03
Now Building
Model orchestration for intelligent auto-routing across model families by user task, while optimizing compute.
Customer Service Copilot · Support AI/ML
AI for Customer Service, on Early GPT Models
Led productization and growth of Dynamics 365 Customer Service Copilot — conversational AI for enterprise support built on early GPT models. Ideated and launched AI ticket summarization, case timeline highlights, and inline email assist within case management. Built ML-based automated ticket creation across ~50M annual support cases.
Built integrations with Google, Salesforce, Meta, and Adobe, and developed the Integrations Contributors program opening Segment’s platform to partner teams. Launched Multi-Instance Destinations — a cross-team effort spanning 30+ services and one of the platform’s largest launches.
Windows Developer Technologies
Where Product Sense Started
Advised and architected solutions for 25 large enterprises across Asia — 400+ hours of technical consultation. Translated recurring user needs into product capabilities and presented brownbag sessions at 30+ customer sites on web development, virtualization, and Azure.
Education
01
University of Illinois Urbana-Champaign
Master of Science, Technology Management · 2019–2020
02
SASTRA University
Bachelor of Technology, Computer Science · 2012–2016
Method — how I ship agents
A dot grid. One unbroken line.
A kolam is drawn fresh each morning: dots first, then a single continuous line that loops through every one without lifting. Good agent development is the same shape — the loop only works if it never breaks.
01
Prototype
Build the smallest real thing. Working demos beat decks — hack week is a legitimate ship vehicle.
02
Prompt
The system prompt is product surface. Behavior specs live in the prompt, and I author the ones we ship.
03
Evaluate
LLM-as-judge plus calibrated human panels. Taste becomes rubrics; rubrics become graders.
Plus a running stack of weekend hacks — local model fine-tuning, hardware and robotics with GPT models, and more.
A note on depth
Much of this work ships behind the enterprise wall.
Frontier agent quality, model orchestration, and safety work is confidential by nature. The real depth is in the conversations — happy to go there.
Currently
Let’s build agents people trust.
Leading product for frontier AI agents at Microsoft AI — Researcher, multi-model strategy, and agent quality in Microsoft Copilot. Writing the prompts, building the evals, making the ship calls.
Open to conversations about frontier AI product leadership, speaking, and collaborations in AI × product.