Skip to content
View rtj1's full-sized avatar

Block or report rtj1

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
rtj1/README.md

Tharun Jagarlamudi

AI Engineer · LLM Agents · Voice AI · Testing & Evaluation

I build LLM agents and voice AI products, and I measure them before I trust them. Right now I'm building Katha, an iOS reader that gives every character in a book its own AI voice. Before that I launched indes.ai, a multi-LLM platform with agentic routing.


Current work

Katha: Multi-Voice AI Reader for iOS · in development

  • Agentic voice casting: Claude profiles each character from the book and writes a voice prompt; Qwen3-TTS renders candidates until one passes a pitch check, then it becomes that character's voice
  • Speaker attribution (NER, alias resolution, Claude as LLM fallback), measured on a 380-quote annotated benchmark: explicit-quote accuracy 0.68 → 0.82 at 95% precision
  • 4 TTS engines routed by language: Kokoro on-device (MLX), self-hosted Qwen3-TTS and Indic Parler-TTS, Apple as fallback, with consent-gated voice cloning
  • 1,500+ Swift tests and 150+ Pytest tests in CI · Swift, Python, Claude API, MLX

indes.ai: Multi-LLM Platform with Agentic Routing

  • Streams GPT, Claude, Gemini and DeepSeek side by side and merges them into one answer (web, iOS/Android)
  • Agentic routing: 11-type query classifier with LLM fallback, plus a 7-tool orchestrator that answers simple queries without an LLM call
  • Fine-tuned a T5 synthesis model on Claude outputs, evaluated with BLEU/ROUGE
  • 650+ Jest/Supertest tests gating every PR, Playwright E2E across Chromium, Firefox and WebKit · Node.js, React Native, Python

Agents

aria: LLM red-teaming agent

  • 10 attack strategies, 77 variants
  • Reflexion-based learning from failed attempts

reflexive-llm-agent: self-reflective tool-using agent

  • ReAct + Reflexion loop with long-term ChromaDB memory
  • FastAPI backend and Streamlit UI

Open-source contributions


ML systems


Skills

  • Agents & LLMs: Claude, GPT, Gemini APIs · tool calling · agentic routing · RAG · ReAct · LangChain · MCP
  • Voice AI: Qwen3-TTS · Indic Parler-TTS · Kokoro · Chatterbox · MLX
  • Testing & Evaluation: Playwright · Jest · Pytest · Swift Testing · load testing · eval sets · BLEU/ROUGE
  • Languages: Python · Swift · JavaScript/Node.js · SQL · C++ · Rust
  • Infrastructure: Google Cloud · AWS · Docker · GitHub Actions · Redis · PostgreSQL · Kafka

Pinned Loading

  1. vllm-project/llm-compressor vllm-project/llm-compressor Public

    State-of-the-art LLM compression, built for production inference with vLLM

    Python 3.8k 673

  2. instructlab/training instructlab/training Public

    InstructLab Training Library - Efficient Fine-Tuning with Message-Format Data

    Python 57 84

  3. aria aria Public

    ARIA: Automated Red-teaming & Iterative Attack Agent - Systematic LLM adversarial robustness testing

    Python

  4. triton-kernels triton-kernels Public

    GPU kernel implementations in Triton - Vector Add, MatMul, Softmax, LayerNorm, FlashAttention

    Python 1