Skip to content
View brijmohan's full-sized avatar
:octocat:
:octocat:

Organizations

@fi-io

Block or report brijmohan

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
brijmohan/README.md

Brij Mohan Lal Srivastava

Voice AI in production. On-premise, low-latency, and measured.

PhD in privacy-preserving speech (Inria / University of Lille). Co-creator of the VoicePrivacy Challenge, the community benchmark for voice anonymization. Co-founder of Nijta, where we deploy speech systems inside regulated environments, mostly public safety, transport police and healthcare.

Most of my production work lives behind customer NDAs. This account is where I rebuild it in the open, on public models and public data.

Currently building

A four-part series on what it actually takes to run a voice agent in production:

  • Latency lab. A cascaded voice agent (LiveKit Agents, Deepgram, Cartesia) with barge-in, semantic turn detection and SIP telephony, instrumented end to end. Where do the 800ms go?
  • On Kubernetes. The same agent deployed to GKE with Terraform, plus the on-premise variant with zero egress. Real-time audio breaks most of the assumptions a normal web service makes.
  • Evaluation harness. Simulated callers, scored on task completion, tool correctness, groundedness, turn-taking failures, latency and cost. Run in CI.
  • Agents and compliance. Multi-agent with tool calls and warm human handoff, built in both LangGraph and Google ADK, with real-time PII redaction in the pipeline.

Background

  • 20+ peer-reviewed publications, 1,500+ citations, h-index 15 (Jul 2026) · Google Scholar
  • Reviewer for NeurIPS, ICLR and Interspeech
  • Organizing Committee, ISCA SPSC Symposium
  • Previously: Microsoft Research (code-switched Hindi-English ASR), Inria

Elsewhere

Website · LinkedIn

Popular repositories Loading

  1. pocketsphinx.js pocketsphinx.js Public

    Forked from syl22-00/pocketsphinx.js

    Speech recognition in JavaScript

    JavaScript 18 6

  2. iremedy iremedy Public

    HTML 16 7

  3. proneval-service proneval-service Public

    This is a web service to accept features extracted by pocketsphinx.js and extract pronunciation evaluation score

    Python 10 4

  4. voice-anonymization-legal-eval voice-anonymization-legal-eval Public

    Python 4 1

  5. lid-convex-comb lid-convex-comb Public

    Convex combination of phonotactics for large-scale spoken language identification

    Python 2 2

  6. kaldi kaldi Public

    Forked from kaldi-asr/kaldi

    This is the official location of the Kaldi project.

    Shell 2