AI Engineer · RAG, LLM products & voice
Sepehr Radmard
I ship the product around the model.
- public projects on GitHub
- 21
- products shipped to live users
- 7
- years shipping for clients
- 10
Selected work
6 projects I would show first
All GitHub projects
21 public repos, filter by topic or stack
Showing 21 of 21 projects
-
polymind
PythonSelf-hosted multi-model AI workspace with side-by-side arenas, judged debates, a workflow canvas, sandboxed code execution and a DLP gate.
- FastAPI
- React
- OpenRouter
-
pm-assistant
PythonLocal-first AI copilot for project managers: MCP tools across Jira, GitHub, Slack and more, approval-gated writes, natural-language rules engine.
- MCP
- FastAPI
- React
-
voice-chess-coach
TypeScriptTalk to a chess coach and play by voice in Persian or English: LiveKit voice agent + tool-calling model driving a live board, python-chess owns the rules.
- LiveKit
- React
- python-chess
-
voice-agent-testgen
PythonPaste a voice agent's prompt, get a judged test suite back: LLM-generated scenarios run against a LiveKit agent, scored turn by turn by an LLM judge.
- LiveKit
- LLM-as-judge
-
agent-studio
TypeScriptLocal, Persian-first AI workbench where agents call native tools and MCP servers, and every step shows up in the chat.
- Next.js
- AI SDK
- MCP
-
true-anomaly
TypeScriptFly a 6DOF probe through a real-scale solar system in the browser, every planet placed from NASA/JPL Horizons data.
- Three.js
- WebGPU
- Vite
-
steel-market-analyst
PythonSteel price report PDFs in; charts, spreads and a grounded AI market read out, with every number traced to a deterministic parser.
- FastAPI
- React
- Multi-agent
-
sepicode
TypeScriptFull-screen terminal coding agent on Bun + OpenTUI: reads, edits, searches and runs code with hard step and dollar caps per turn.
- Bun
- OpenTUI
- OpenRouter
-
rag-service
PythonHybrid retrieval API for LLM agents: multi-query fusion, dense + sparse search over LlamaCloud vector indexes, Cohere rerank.
- FastAPI
- LlamaIndex
- Cohere
-
rag-evaluator
PythonFind where your RAG chatbot is wrong: LLM-as-judge on four criteria plus side-by-side human expert scoring.
- Flask
- LLM-as-judge
-
product-assistant
TypeScriptProduct Talk: scan a supermarket product and ask it out loud about protein, sugar or allergens. Realtime voice agent grounded in pack facts.
- LiveKit
- OpenAI Realtime
- Next.js
-
pr-agent
TypeScriptSeven AI reviewers on every GitLab merge request, one human who decides what gets posted. Multi-agent code review with a CTO dashboard.
- n8n
- GitLab
- Next.js
-
phone-agent
PythonPersian AI phone receptionist: answers a real call over Asterisk AudioSocket, STT → LLM → TTS in natural Persian, with an Android companion app.
- Asterisk
- STT/TTS
- Android
-
meeting-assistant
TypeScriptPersian meeting recorder: chunked speaker diarization with speaker stitching, schema-validated LLM summaries, searchable minutes.
- FastAPI
- Next.js
- Diarization
-
livekit-fa-agent
PythonPersian speech-to-speech voice agent on LiveKit that answers only from retrieved passages (RAG), built for filtered networks.
- LiveKit
- RAG
- Realtime
-
invoice-extractor
PythonExtract invoice fields with OCR + an LLM, then check every value against the exact box it came from. Human-in-the-loop review UI, English and Persian.
- PaddleOCR
- FastAPI
- Structured output
-
feedo
TypeScriptPlain-Persian bug reports in, structured AI drafts out: vision-LLM drafting, duplicate detection and truncated-JSON recovery.
- Next.js
- FastAPI
- Vision LLM
-
erp-sql-agent
PythonAsk your ledger questions in Persian, get audited numbers back: read-only accounting dashboard plus an LLM tool-calling SQL agent over an ERP warehouse.
- Text-to-SQL
- Hasura
- FastAPI
-
dongeto
TypeScriptPersian bill splitter: type the expense in Farsi, an LLM drafts it via tool calls, code does the math (min-transfer settlement, tested).
- Next.js
- Drizzle
- Vitest
-
ai-explain
TypeScriptAsk any question, get an interactive HTML/SVG explainer page planned and designed by a multi-stage LLM pipeline and streamed into a sandboxed iframe.
- Next.js
- AI SDK
- Generative UI
-
live-media-tavus
JavaScriptRealtime voice assistant with a live Tavus virtual avatar in the room: OpenAI Realtime, MediaPipe face detection triggers, LiveKit tokens.
- LiveKit
- OpenAI Realtime
- Tavus
No project matches that search and topic.
About
I ship live AI products for Iranian companies: workspace, RAG, meetings, helpdesk, OCR, and realtime voice / telephony agents. Full path (API, UI, auth, Docker/CI), then vibe-code the rest with Claude Code, Cursor, Hermes, and MCP.
How I ship
- 01
Design the loop
Who types, what the model sees, when a human steps in.
- 02
Build the path
API + UI + auth + DB. FastAPI, Next.js, Keycloak, Postgres. RTL when needed.
- 03
Ship live
Docker / GitLab CI / Nginx → live URL. Then iterate with agent passes.
Skills
- AI / LLM
- Python, FastAPI, Next.js, OpenRouter, Ollama, agents, DLP. Gemini 3.6 / Claude Opus 5 / GPT-5.6, pick by quality, latency, cost.
- RAG / vectors
- Hybrid dense+sparse, multi-query fusion, RRF, Cohere rerank, LlamaCloud, Pinecone, pgvector, embeddings
- Voice / telephony
- LiveKit Agents, Vapi, ElevenLabs, OpenAI Realtime, in-call tool calling, barge-in latency, conversation QA
- Eval / obs
- Langfuse, DeepEval, Sentry, Datadog: traces, agent evals, production errors
- DevOps / MLOps
- Docker, GitLab CI/CD, Nginx, GPU serving, Unsloth LoRA / QLoRA, 4-bit quant, distillation, Ollama / GGUF
- AI coding
- Vibe coding: Claude Code, Cursor, Grok, Hermes Agent, OpenClaw, MCP, skills / subagents
Have a model that needs a product?
Write to me on LinkedInDownload CV (PDF) GitHub LinkedIn
Sepehr Radmard, Shiraz, Iran