๐Ÿ“‹Stalecollected in 27m

Microsoft Tests Voice Avatars Sage & Pax

Microsoft Tests Voice Avatars Sage & Pax
PostLinkedIn
๐Ÿ“‹Read original on TestingCatalog

๐Ÿ’กCopilot's new voice avatars Sage/Pax + OneDrive sync testing โ€“ multimodal upgrade for devs.

โšก 30-Second TL;DR

What Changed

Testing Portraits in Voice Mode for Copilot

Why It Matters

Enhances Copilot's multimodal voice capabilities, making interactions more personalized. OneDrive integration improves file access during voice sessions. Could boost adoption among productivity users.

What To Do Next

Join the Copilot insider program to preview Sage and Pax voice avatars.

Who should care:Developers & AI Engineers

Key Points

  • โ€ขTesting Portraits in Voice Mode for Copilot
  • โ€ขNew voice avatars Sage and Pax introduced
  • โ€ขOneDrive sync option added
  • โ€ขPublic rollout planned for later this year

๐Ÿง  Deep Insight

Background and context from public sources โ€” not the original article. 7 sources cited.

๐Ÿ”‘ Enhanced Key Takeaways

  • โ€ขMicrosoft's VASA-1 research model from 2024 generates lifelike talking faces from a single image and audio clip using a diffusion-based architecture for realistic lip sync, expressions, and head movements[1][2][5].
  • โ€ขAzure AI Speech service offers text-to-speech avatars with standard and custom options, supporting real-time synthesis, batch processing, and photorealistic videos in MP4 or VP9 formats[4].
  • โ€ขVASA-1 was not released publicly due to deepfake misuse concerns, with Microsoft withholding demos, APIs, or products until responsible safeguards are ensured[1].

๐Ÿ”ฎ Future ImplicationsAI analysis grounded in cited sources

Copilot voice avatars will leverage VASA-1-like diffusion models for enhanced realism
VASA-1's face latent space and diffusion techniques for expressions and motions align directly with the tested Portraits in Voice Mode features[1][2][5].
OneDrive sync will enable persistent avatar personalization across devices
Integration with OneDrive suggests cloud-based storage for user-specific voice and portrait data to maintain continuity in Copilot interactions[original article].

โณ Timeline

2024-03
Microsoft Research releases VASA-1 paper and demos for audio-driven talking faces[1][5]
2024-04
VASA-1 coverage highlights diffusion model for hyper-realistic avatars[2]
2026-03
Microsoft begins testing Copilot Portraits in Voice Mode with Sage and Pax avatars[original article]
๐Ÿ“ฐ

Weekly AI Recap

Read this week's curated digest of top AI events โ†’

๐Ÿ‘‰Related Updates

AI-curated news aggregator. All content rights belong to original publishers.
Original source: TestingCatalog โ†—

This is a summary, not the original. Read the source, or get the weekly briefing.

Weekly AI briefing

One email a week. Unsubscribe anytime.