I work on how AI capabilities become useful products, and what happens when people use them.

At Microsoft, I help turn AI research into products people can use and models others can build on. My work spans consumer AI, foundation-model releases, deployment decisions and the internal workflows that support them, including building a new function as a division took shape.

As a Principal Technical Program Manager, I work with researchers and engineers to define requirements, interpret evaluation results and resolve gaps before release. I connect model behavior to the experience a team is trying to deliver, then help translate those findings into product changes and release decisions.

Hands-on data science and political economy inform how I approach that work. I use statistical analysis to understand capabilities and uncertainty, and examine the incentives, institutions and dependencies shaping adoption. A model's performance matters alongside who can adapt it, what deployment requires and how an organization can change course.

My current work focuses on open-weight models and their ecosystems. I bring together use cases, evaluation findings and remediation options, analyze the evidence, and advise on deployment choices. That includes the commercial and geopolitical implications of building on a particular ecosystem, as well as what an organization can actually inspect or adapt.

I'm also exploring world models and their potential to support creation, learning and planning. I'm interested in how simulated experience could help people anticipate consequences, and in what it would take to test a model's predictions before relying on them in a real task.

Writing

Openness Is a Safety Property

Everyone argues about who can download a model. The question that decides outcomes is who can fix one.

World Models and the Limits of Simulated Experience

World models could help AI systems anticipate consequences before acting. What would it take for the people relying on those predictions to test their assumptions, correct their limitations, and understand what remains uncertain?

On the Subtle Differences within Model Censorship

Three very different problems produce exactly the same silence. Telling them apart changes everything about what you can do next.

Owning the Infrastructure Doesn't Necessarily Mitigate Model Risks

Hosting a model on your own soil settles one real question and leaves the harder one completely untouched.

Eligibility, Not Trust

We never ask whether steel is safe. We ask what load it's rated to carry. Models deserve the same question.

Selected work

Research-to-product delivery

Microsoft

I've led model-behavior requirements and release-readiness work across Microsoft's foundation-model releases and consumer AI integrations, from Copilot Voice and Vision to OneDrive photo search. With research and engineering teams, I interpret evaluation findings and translate them into feedback and product changes for the intended experience.

Model and ecosystem strategy

Microsoft

I lead work to consolidate use cases, evaluation findings and remediation options for open-weight models. Interpreting results across teams, I advise Microsoft executives on deployment and ecosystem choices. That advice connects model limitations and practical use cases to commercial fit, sovereignty concerns and the geopolitical dependencies an organization takes on.

Building new functions and workflows

Microsoft

When Microsoft AI formed in 2024, I built its Responsible AI program from zero. I also worked directly with developers and engineering users to define requirements and test the migration of an internal review workflow, resolve edge cases and retire legacy dependencies without disrupting active cases.

Applied data science and visualization

World Economic Forum · Carnegie Mellon University CREATE Lab

Earlier, I led data science at the World Economic Forum, building the analytical layer behind Transformation Maps, available through its Strategic Intelligence platform. In parallel, I collaborated with Carnegie Mellon University's CREATE Lab on geospatial and climate-projection stories for EarthTime.

Contact

If you're working on AI products, research or deployment, I'd welcome a conversation. andrew@andrewberkley.com