Evasion test time attacks on Mozilla DeepSpeech (STT).
10K+
Docker image to run test time optimisation attacks against Mozilla DeepSpeech.
These images are built from this repo on each commit, with latest referring to the master branch of the code.
These images contain everything you need to run some attacks. The DeepSpeech checkpoints & scorer are included. Different data sets are also included with the mozilla common voice V1 and V7 valid test data sets living in the samples directory.
Please note that these images require a CUDA compatible NVIDIA GPU, a suitable nvidia driver version and the nvidia-container-runtime.
You can start testing the code out by using
docker run -it --rm \
--name cleverspeech \
--gpus all \
dijksterhuis/cleverspeech:latest
The cleverspeech/scripts directory then contains a bunch of help scripts to try things out for attacks with different loss functions and graphs etc.
If you want to bind mount a local directory to generate some results then you'll need to provide your user and group ID as environment variables. Please note that using docker run --user=${USER} will break things as docker will not chown any files. Instead you'll need to run this command
docker run -it --rm \
--name cleverspeech \
--gpus all \
-v $(pwd)/adv/:home/cleverspeech/cleverSpeech/adv:rw \
-e LOCAL_UID=$(id -u ${USER}) \
-e LOCAL_GID=$(id -g ${USER}) \
dijksterhuis/cleverspeech:latest
Note that the chown and changing of user file permissions can take up to 15 minutes depending on where your docker install stores image data (nvme SSDs take less than 5 minutes).
Content type
Image
Digest
Size
3.6 GB
Last updated
over 4 years ago
docker pull dijksterhuis/cleverspeech