SpeechMike Ambient Wearable AI Assistant

PSM5000 Series

High-performance beamforming microphones
ensure accurate speech recognition

Specific audio modes
for optimal performance across environments and use cases

Patented ambient sound intelligence
ensures natural, intelligible recordings

Secure, encrypted communication
with IT-compliant design for safe clinical use

Turn your speech into results

Transforming professional workflows with ambient AI and speech recognition

Professionals today face increasing administrative workloads, documentation demands, and time pressure. The SpeechMike Ambient Wearable AI Assistant builds on the proven Philips SpeechMike family, now expanded for today’s challenges: a next-generation, professional-grade dictation microphone that supports both ambient AI use cases and traditional speech recognition workflows. From meeting transcription, legal documentation, and report creation to AI-powered note generation, multilingual communication, and virtual assistant functions, the device is designed to optimize daily work. By improving documentation, communication, and workflow efficiency, the SpeechMike Ambient helps professionals focus on high-value tasks, reduce administrative burden, and improve overall productivity.

Features

High-performance beamforming microphones ensure accurate speech recognition

The SpeechMike Ambient features four high-performance beamforming microphones for ambient recording. Its wearable design allows full mobility while keeping the microphones close to the speaker at all times, ensuring greater accuracy in dynamic, noisy clinical environments—far outperforming smartphones or generic microphones.

Specific audio modes for optimal performance across environments and use cases

Different workflows call for different audio setups, and the SpeechMike Ambient is designed to meet those needs. Ambient Mode suits both conversation and single-speaker scenarios, preserving natural sound, while Dictation Mode focuses on the primary speaker and suppresses background voices. Two additional modes are available to integrators for speaker separation in conversation settings, ensuring maximum flexibility.

Patented ambient sound intelligence ensures natural, intelligible recordings

The built-in patented ambient sound intelligence automatically detects individual speakers and creates two distinct audio streams. This results in natural, easy-to-follow playback, making the device ideal for automated documentation, conversation transcription, and protocol generation.

Purpose-built design for superior performance of AI-powered assistants and tools

The SpeechMike Ambient is engineered specifically to capture high-fidelity, single and multi-speaker audio in real-world environments—perfect for AI transcription, conversational AI and ambient scribe scenarios in clinical, and virtual assistant functions. Its superior input quality significantly enhances AI model accuracy, making it ideal for generating clinical notes, transcriptions, and real-time documentation.

Secure, encrypted communication with IT-compliant design for safe clinical use

The SpeechMike Ambient ensures secure data transmission with encrypted Bluetooth LE technology, and a secure “Passkey” pairing method. It avoids over-the-air pairing and allows only one active connection at a time to prevent unauthorized access, interference, and man-in-the-middle attacks. Tested for reliable coexistence with Bluetooth, Wi-Fi, and other 2.4 GHz devices, the system meets global standards like CE, FCC, and RCM.

Developer SDK for fast, native integration into custom software ecosystems

The SpeechMike Ambient is fully backwards compatible with existing SpeechMikes, ensuring seamless integration into current workflows while unlocking new opportunities. With it’s comprehensive software development kit (SDK) and API access, third-party software providers can natively integrate the device into desktop and mobile applications. With full control over button mapping, device settings, and audio input, the SDK enables rapid onboarding for speech recognition platforms, AI tools, and enterprise software ecosystems.

Compact, wearable design for seamless stationary, mobile, and hands-free use

Designed for seamless transitions between mobile and stationary use, the device can be worn with a magnetic clip or neck strap for hands-free operation. It stores up to 10 wireless profiles, automatically connecting to the nearest workstation when moving between patient rooms or work areas, eliminating the need for manual pairing. The docking station enables charging and desktop recording, while the ergonomic, lightweight build with tactile feedback ensures comfort and durability for long shifts.

All-day power with simple, flexible charging

The energy-efficient design delivers up to 10 hours of continuous recording time, ensuring full-shift battery life in a lightweight body. Users can recharge quickly via the docking station, a standard USB-C cable, or a PC USB port. This ensures reliable uptime and supports the mobile demands of clinical environments.

Hygienic, low-maintenance design for safer handling and reduced operating costs

Built with infection control in mind, the device uses hygienic, medical-grade materials and avoids the contamination risks common with handheld devices. Its smooth, polished surface resists germs and fingerprints while minimizing handling noise during operation for clearer audio capture. Compared to smartphones, it also lowers operational costs with fewer updates, simpler maintenance, and no need for mobile device management.

Wireless connectivity

Wireless technology: 2.4 GHz Bluetooth Low Energy
Maximum power: ≤ 10mW
Maximum range: up to 25 m / 82 ft (in clear view)

Audio recording

Microphone type: MEMS 4-way microphone array
Characteristic: omni-directional and beam forming
Frequency response: 200 – 8000 Hz

Sound

Speaker type: built-in rectangular, dynamic speaker
Acoustic frequency response: 300 – 8000 Hz
Speaker output power: > 200 mW

Power

Battery type: Li-polymer
Rechargeable: via docking station or USB-C power supply
Battery lifetime: up to 10 hours continuous talk time
Charging time: 3 hours

Product dimensions

Product dimensions (W × D × H): 32 × 104 × 15 mm / 1.3 × 4.1 × 0.6 in
Weight: 42 g / 1.5 oz

Wireless Adapter

Product dimensions (W × D × H): 14 × 7 × 42.5 mm / 0.6 × 0.3 × 1.7 in
Weight: 4 g / 0.1 oz

Docking Station

Product dimensions (W × D × H): 85 × 85 × 32 mm / 3.4 × 3.4 × 1.3 in
Weight: 140 g / 4.9 oz
USB-C: for charging and data connection
USB-C: for Wireless Adapter
Kensington lock

System requirements Philips SpeechControl Device and Application Control Software

Processor: Intel dual core or equivalent AMD processor, 1 GHz or faster processor
RAM: 2 GB (32 bit)/4 GB (64 bit)
Hard-disk space: 30 MB for SpeechControl software, 4.5 GB for Microsoft .NET Framework
Operating system: Windows 11, Windows 10 (64 bit)
Graphics: DirectX-compliant graphics card with hardware acceleration recommended
Sound: Windows-compatible sound device
Free USB-C port

Supported speech recognition software

Philips SpeechLive
Microsoft Dragon Copilot (Flex)
Microsoft Dragon Medical One
Microsoft Dragon Legal 16
Microsoft Dragon Professional 16
Solventum Fluency Direct
Solventum Fluency for Imaging
Dolbey Fusion Narrate powered by nVoq
Tandem Health

Green specifications

Compliant to 2011/65/EU (RoHS)
Lead-free soldered product

Operation conditions

Temperature: 5° – 45° C / 41° – 113° F
Humidity: 10 % – 90 %

Design and finishing

Material: high-class polymers
Color: dark grey pearl metallic / black