The best AI doesn't just scale, it adapts to the languages, cultures, people and communities it serves. Explore three playbooks created to help teams design, deploy, and evaluate AI systems with inclusivity in mind.
We advance science and technology to benefit humanity.
Joined February 2009
- Can AI learn pathology through clinical dialogue? Introducing PRISM2, a multimodal foundation model trained on pathology images and language from real pathology reports. Using simple question-answering, PRISM2 matches specialized cancer-detection systems across several benchmark
- Small language models learn to negotiate with SocialRL, PazaBench V2 expands speech AI evaluation across African languages, and EvoLib helps agents turn experience into knowledge. Plus: new methods for more reliable A/B testing and advances in AI-driven precision oncology.
- Orchard is an open-source framework for the research community to train and evaluate AI agents across task types. It reduces complexity while supporting strong performance from smaller models by enabling researchers to reuse the same infrastructure. msft.it/6019a8fqP
- Computer-use AI agents struggle with multi-step workflows like email and customer support. Echoverse trains agents in realistic environments rather than simply providing more training tasks, helping them improve as the tasks, tests, and environments evolve.



