Apple has acquired Israeli artificial intelligence startup Q.ai, bringing technology that can interpret whispered and silent speech by analysing subtle facial micromovements.
The deal, valued at roughly $1.6 to $2 billion represents Apple’s largest acquisition since Beats in 2014 and one of the clearest signals yet that the company is betting on new ways for users to interact with AI beyond traditional voice and touch.
Around 100 Q.ai employees, including Chief Executive Aviad Maizels and co-founders Yonatan Wexler and Avi Barliya, will join Apple’s hardware technologies group.
Apple said the startup has been working on new applications of machine learning for understanding whispered speech and enhancing audio in challenging environments, though it did not disclose detailed product plans.
The acquisition comes as Apple faces intensifying competition from rivals including Google, Meta and OpenAI, all of which are racing to embed conversational AI into devices and emerging form factors such as smart glasses and dedicated AI hardware.
For Apple, which has faced criticism for lagging in conversational AI, the deal points to a strategy focused on owning the interface layer as much as the AI models themselves.
From Voice to Facial Interfaces
At the core of Q.ai’s technology is the ability to detect facial skin micromovements associated with speech.
Even when a person produces no audible sound, the muscles used to form words still move in consistent patterns.
By combining imaging, audio processing and machine learning, the system aims to map those subtle movements to words and intent.
This approach goes beyond traditional lip-reading, which relies primarily on visible mouth shapes. Q.ai’s systems are designed to capture subtler cues across the face that may not be visible to the human eye, allowing devices to infer commands even when speech is whispered or silent.
For users, this could make interaction with digital assistants more discreet and socially acceptable, particularly in meetings, open-plan offices, healthcare environments and noisy workplaces where speaking commands out loud is impractical or disruptive.
A Foundation for Wearables and Spatial Computing?
The implications for wearables are particularly significant.
Apple has positioned Vision Pro as a major step into spatial computing and is widely expected to pursue lighter, more everyday smart glasses over time.
In those form factors, relying solely on voice control presents both technical and social limitations.
Silent speech and facial intent detection could become a key control layer for head-worn devices, enabling users to interact with digital overlays, assistants and collaboration tools without speaking out loud.
For enterprise users, this could support hands-free access to information, task management and real-time guidance in environments where noise, privacy or safety make voice interaction difficult.
In UC scenarios, silent controls could also allow participants to trigger actions, retrieve information or manage meetings without interrupting discussions, potentially reshaping how AI is embedded into everyday workplace workflows.
Emotional and Biometric Signals Raise Privacy Stakes
Q.ai’s patents also point to capabilities that extend beyond speech.




