Skip to content
View will-rice's full-sized avatar

Highlights

  • Pro

Organizations

@Hugging-Face-Supporter

Block or report will-rice

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
will-rice/README.md

Will Rice

ML engineer, nine years in speech and audio.

I work on real-time lip sync — audio-driven talking heads, and the inference work that gets them running in real time: 2-step distillation, distilled VAE, FP8 quantization, TensorRT. Before that, production TTS and STT serving on multi-GPU infrastructure, streaming pipelines with endpointing, voice activity detection and CTC decoding, and on-device ASR and NLU under embedded constraints.

Selected work

TFHubert / TFWav2Vec2 TensorFlow implementations of Meta's self-supervised speech models, in huggingface/transformers
vadkit Multi-provider voice activity detection — FireRedVAD, Silero, FSMN-VAD, WebRTC — behind one streaming API for endpointing
pe-av-syncnet SyncNet on Meta's Perception Encoder audio-visual representations, for lip-sync evaluation
whisper-rl Reinforcement learning on Whisper, past what supervised fine-tuning reaches
denoisers PyTorch waveform denoisers for speech enhancement
spokestack-python Python library for embedded speech: on-device wake word, ASR and natural language understanding
ai-agent-security-2026 Kaggle gold — red-teaming LLM agents for multi-step tool-misuse attacks

Consulting

Available for consulting on speech, ASR/TTS and inference optimization.

wrice20@gmail.com · LinkedIn · Kaggle

Pinned Loading

  1. spokestack/spokestack-python spokestack/spokestack-python Public archive

    Spokestack is a library that allows a user to easily incorporate a voice interface into any Python application with a focus on embedded systems.

    Python 141 15

  2. denoisers denoisers Public

    Simple PyTorch Denoisers for Waveform Audio

    Python 43 2

  3. ai-agent-security-2026 ai-agent-security-2026 Public

    AI Agent Security - Multi-Step Tool Attacks

    Python 1 2

  4. pe-av-syncnet pe-av-syncnet Public

    SyncNet based on Meta's Perception Encoder Audio-Visual (PE-AV)

    Python 1 1

  5. vadkit vadkit Public

    Multi-provider voice activity detection for the browser — FireRedVAD, Silero, and WebRTC VAD behind one TypeScript API

    TypeScript 1

  6. whisper-rl whisper-rl Public

    Example of Using Reinforcement Learning to Improve Whisper

    Python 1