AI Engineer · RAG, LLM products & voice

Sepehr Radmard

I ship the product around the model.

RAG / retrieval · workspace · meetings · OCR · voice / telephony

Shiraz, Iran GitHub LinkedIn

public projects on GitHub
21
products shipped to live users
7
years shipping for clients
10

Selected work

6 projects I would show first

  1. polymind

    Self-hosted multi-model AI workspace with side-by-side arenas, judged debates, a workflow canvas, sandboxed code execution and a DLP gate.

    FastAPI, React, OpenRouter Code for polymind on GitHub

  2. voice-chess-coach

    Talk to a chess coach and play by voice in Persian or English: LiveKit voice agent + tool-calling model driving a live board, python-chess owns the rules.

    LiveKit, React, python-chess Code for voice-chess-coach on GitHub

  3. true-anomaly

    Fly a 6DOF probe through a real-scale solar system in the browser, every planet placed from NASA/JPL Horizons data.

    Three.js, WebGPU, Vite Code for true-anomaly on GitHub

  4. phone-agent

    Persian AI phone receptionist: answers a real call over Asterisk AudioSocket, STT → LLM → TTS in natural Persian, with an Android companion app.

    Asterisk, STT/TTS, Android Code for phone-agent on GitHub

  5. pr-agent

    Seven AI reviewers on every GitLab merge request, one human who decides what gets posted. Multi-agent code review with a CTO dashboard.

    n8n, GitLab, Next.js Code for pr-agent on GitHub

  6. rag-service

    Hybrid retrieval API for LLM agents: multi-query fusion, dense + sparse search over LlamaCloud vector indexes, Cohere rerank.

    FastAPI, LlamaIndex, Cohere Code for rag-service on GitHub

All GitHub projects

21 public repos, filter by topic or stack

About

Sepehr Radmard, photograph

I ship live AI products for Iranian companies: workspace, RAG, meetings, helpdesk, OCR, and realtime voice / telephony agents. Full path (API, UI, auth, Docker/CI), then vibe-code the rest with Claude Code, Cursor, Hermes, and MCP.

How I ship

  1. 01

    Design the loop

    Who types, what the model sees, when a human steps in.

  2. 02

    Build the path

    API + UI + auth + DB. FastAPI, Next.js, Keycloak, Postgres. RTL when needed.

  3. 03

    Ship live

    Docker / GitLab CI / Nginx → live URL. Then iterate with agent passes.

Skills

AI / LLM
Python, FastAPI, Next.js, OpenRouter, Ollama, agents, DLP. Gemini 3.6 / Claude Opus 5 / GPT-5.6, pick by quality, latency, cost.
RAG / vectors
Hybrid dense+sparse, multi-query fusion, RRF, Cohere rerank, LlamaCloud, Pinecone, pgvector, embeddings
Voice / telephony
LiveKit Agents, Vapi, ElevenLabs, OpenAI Realtime, in-call tool calling, barge-in latency, conversation QA
Eval / obs
Langfuse, DeepEval, Sentry, Datadog: traces, agent evals, production errors
DevOps / MLOps
Docker, GitLab CI/CD, Nginx, GPU serving, Unsloth LoRA / QLoRA, 4-bit quant, distillation, Ollama / GGUF
AI coding
Vibe coding: Claude Code, Cursor, Grok, Hermes Agent, OpenClaw, MCP, skills / subagents

Have a model that needs a product?

Write to me on LinkedIn

Sepehr Radmard, Shiraz, Iran