K.Boopathi
BlogNotesSecond BrainProjectsHire MeYour reading file☕0 pts
BlogNotesSecond BrainProjectsHire MeYour reading file☕0 pts
Home

❯

notes

❯

AI ML

❯

Generative AI

❯

Voice Multimodal

15 items under this folder.

  • Mar 01, 2026

    Audio Processing - How ML Models Understand Sound

  • Mar 01, 2026

    Inference Optimization - Speed & Cost Reduction

  • Aug 17, 2026

    Simulation

  • Mar 01, 2026

    Speech Language Models - Voice LLM Architecture

  • Mar 01, 2026

    Speech-to-Speech Models - Real-Time Voice Interaction

  • Mar 01, 2026

    Text-to-Speech (TTS) Models - Architecture & Implementation

  • Nov 10, 2025

    voice agent deployment

  • Mar 01, 2026

    Voice Agent Evaluation & Testing

  • Mar 01, 2026

    Voice Agents Deployment - Infrastructure & Cost

  • Mar 01, 2026

    Voice Eval

  • May 04, 2025

    Voice LLM

  • Mar 01, 2026

    Voice LLM Fundamentals - Audio & Core Concepts

  • Aug 27, 2025

    Voice LLM processing

  • Mar 01, 2026

    Voice LLM Resources - Tools, Libraries & References

  • Mar 09, 2026

    Voice Roadmap


    Kept in this browser. Nothing is sent anywhere.
    How the ledger works
      ☕
      A small way to keep this goingFound something useful?

      Fuel the next write-up
      One coffee, no subscription—and no interruption if you pass.

      Graph View

      Build with ♥ K.Boopathi © 2026

      • GitHub
      • Linkedin
      • Twitter