Make personal computing accessible through the ways people naturally communicate. You will study speech, vision and other inputs under real device constraints.
Not open yet: Research program
This work starts when that stage arrives, so there is no application to submit today and we will not pretend otherwise. What is written below is what the role is for and what would make somebody right for it, published early on purpose so you can decide whether it is worth watching.
The work
Develop or adapt models for speech, documents, images and grounded interaction. Evaluate diverse accents, environments and accessibility needs with appropriate participant consent. Measure local performance and make capture, retention and activation behavior understandable.
The milestone
In your first 90 days, deliver a bounded multimodal capability with evaluation across realistic conditions and a clear record of its limitations.
Evidence
Bring research or advanced engineering in speech, vision or multimodal learning. Show how you distinguish apparent fluency from correct understanding.
Evidence, not credentials. We are describing work you can point at, in whatever form it exists.
The exercise
Design an evaluation for a voice-controlled task in a shared room where only one person has authorized the action.