whisper_ros

This repository provides a set of ROS 2 packages to integrate whisper.cpp into ROS 2 using audio_common. Besides, silero-vad is used to perform VAD (Voice Activity Detection).

Related Projects

chatbot_ros → This chatbot, integrated into ROS 2, uses whisper_ros, to listen to people speech; and llama_ros, to generate responses. The chatbot is controlled by a state machine created with YASMIN.

Installation

To run llama_ros with CUDA, first, you must install the CUDA Toolkit.

$ cd ~/ros2_ws/src
$ git clone https://github.com/mgonzs13/audio_common.git
$ git clone https://github.com/mgonzs13/whisper_ros.git
$ sudo apt install portaudio19-dev
$ pip3 install -r audio_common/requirements.txt
$ pip3 install -r whisper_ros/requirements.txt
$ cd ~/ros2_ws
$ colcon build --cmake-args -DGGML_CUDA=ON # add this for CUDA

Usage

Run Silero for VAD and Whisper for STT:

$ ros2 launch whisper_bringup whisper.launch.py

Demos

Send a goal action to listen:

$ ros2 action send_goal /whisper/listen whisper_msgs/action/STT "{}"

Or try the example of a whisper client:

$ ros2 run whisper_demos whisper_demo_node

Name		Name	Last commit message	Last commit date
Latest commit History 145 Commits
whisper_bringup		whisper_bringup
whisper_cpp_vendor		whisper_cpp_vendor
whisper_demos		whisper_demos
whisper_msgs		whisper_msgs
whisper_ros		whisper_ros
.gitignore		.gitignore
CITATION.cff		CITATION.cff
LICENSE		LICENSE
README.md		README.md
requirements.txt		requirements.txt

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

Repository files navigation

whisper_ros

Table of Contents

Related Projects

Installation

Usage

Demos

About

Releases 30

Packages

Contributors 3

Languages

License

mgonzs13/whisper_ros

Folders and files

Latest commit

History

Repository files navigation

whisper_ros

Table of Contents

Related Projects

Installation

Usage

Demos

About

Topics

Resources

License

Stars

Watchers

Forks

Releases 30

Packages 0

Contributors 3

Languages

Packages