Anthropic opens up to voice collection to train Claude
The company introduces an optional feature to share audio and spoken conversations, separate from the consent already required for text
Anthropic has introduced a new option for Claude users concerning voice. Those who use the assistant in voice mode can now choose to share their recordings and audio conversations so that the company can use them to improve its models. The feature is not active by default: explicit opt-in is required from the person using the service.
The most significant choice, from the standpoint of those managing their own data, concerns the separation introduced by Anthropic between the different types of content generated while using Claude. Authorization to share voice recordings is distinct from that concerning text conversations and from Claude Code sessions, the company’s tool dedicated to code writing. A user could therefore, in theory, authorize the use of their spoken input without this automatically implying consent for the texts exchanged with the assistant, or vice versa.
The move fits into a broader industry trend: companies developing large language models are constantly seeking new data sources to train subsequent versions of their systems, and audio represents a category that has so far been less exploited compared to written text, which is available in enormous quantities thanks to decades of digital content. Voice recordings make it possible to train models on nuances that text does not capture: intonation, pauses, ways of expressing oneself that change depending on the spoken context.
It remains to be seen how many users will be willing to voluntarily activate this sharing, given that the feature requires an additional step beyond simply using the service. The news currently comes from a single source (Anthropic press release reported by Quotidiano Nazionale); no independent confirmation is currently available on details such as the number of users involved or the timeline for implementing the feature.
The topic nonetheless fits into a broader debate on the relationship between conversational assistants and personal data: the distinction between separate consents for voice, text and code is one of the few concrete measures adopted so far by a major company in the sector to fragment, rather than aggregate, the authorizations required of users.
← Archive · Front page · Past editorials · Report an error · Original article (in Italian)