HeAR
A model for health sounds: coughing, breathing, throat clearing. It provides features on top of which a researcher builds their own model and draws no conclusions about illness itself. It does not replace a doctor; decisions are made by a specialist. Voice recordings are personal data, so check the procedure with a lawyer.
- Developer
- Google, USA
- First release
- Dec 2024
- Latest release
- Apr 2025
- Sizes
- a ViT-Large class model, trained on more than 300 million two-second clips
- License
- Commercial use with conditionsHealth AI Developer Foundations Terms of Use - open weights with restrictions, access on request
- Running
- On your own serverAlso runs without a GPU
- Industries
- Healthcare, Science and research
What it does
- Features of cough and breathing sounds for your own model
- Research projects in health acoustics
- Selecting cough and breathing fragments from a recording
- Pilots with a small amount of labelled data
Where it is used
Hardware requirements
Versions
- Обновление карточки и артефактов
- HeAR, открытые веса на Hugging Face
How to run it
I can set this up end to end: pick the model size, deploy it on your server and connect it to your systems.
Frequently asked questions
Can HeAR be used in a commercial project?
With conditions. License: Health AI Developer Foundations Terms of Use - open weights with restrictions, access on request. Restrictions vary — region, company revenue, attribution requirements. Have a lawyer check the terms before a commercial launch.
What hardware does HeAR need?
At minimum: Laptop or regular PC, up to 8 GB of VRAM — smaller versions. Some versions also run on an ordinary CPU, without a GPU. You can calculate the exact VRAM for your model size and context in the hardware calculator.
Does HeAR support Russian?
Language does not matter for this model: it does not work with text.
Where can I download HeAR and what does it cost?
The HeAR weights are open and free to download. You only pay for the hardware it runs on and for the setup. Source links are at the bottom of this page.
How I deploy it for clients
- SelectionI pick the model size for your task and hardware and test it on your examples.
- DeploymentI deploy it on your server or in a closed network and provide an API.
- Fine-tuningI fine-tune it on your data (LoRA) or connect a knowledge base — whichever is cheaper for the task.
- IntegrationI connect it to your CRM, ERP, bot, website or team chat and set up monitoring.
Similar models
Google's medical version of Gemma: reads medical texts and images (X-ray, dermatology, histology). A tool for doctors and developers; does not replace a doctor, decisions are made by a specialist.
DetailsAcoustic monitoringPANNsUniversity of Surrey · UKCommercial use allowedThe classic set of convolutional networks that label sound across the 527 AudioSet categories, from machinery noise and alarms to breaking glass and screams. The system does not make a diagnosis, it gives you a reason to check the unit before it fails. A microphone can also capture people voices, which is personal data, so check the procedure with a lawyer.
DetailsAcoustic monitoringCEDXiaomi · ChinaCommercial use allowedCompact sound-labelling models from 5.5M to 86M, with ONNX and INT8 builds that run on an ordinary CPU and on a board next to the equipment. The system does not make a diagnosis, it gives you a reason to check the unit before it fails.
DetailsSource: huggingface.co/google/hear. Data checked against the model card on 22 Sep 2026. Have a lawyer review the license before commercial launch.


