CodePilot
An autonomous AI engineering system that turns natural-language specs into full applications: repository-level understanding, architectural planning, RAG-based context retrieval, and self-correcting test/lint loops.
View on GitHub →From a 100K+ devotional content creator to a Generative AI Engineer at Zencia AI — the story of how Saksham Pathak, known online as Parthmax, rebuilt himself around voice agents, agentic AI, and production LLM systems.
Creator → AI Engineer
Every chapter looks unrelated to the last — devotional editing, pure math, freelance AI evaluation, shipped open-source tools, production voice AI — until you see the thread running through all of it: understanding what makes something actually work for people.
While studying B.Sc. Mathematics at Dr. Ram Manohar Lohia Avadh University, Saksham built a 100K+ following on Instagram and YouTube editing devotional content — an early lesson in what makes content resonate at scale.
Began an M.Sc. in Artificial Intelligence & Machine Learning at IIIT Lucknow, trading the editing timeline for transformer architectures and systems thinking.
Freelance AI developer work in LLM evaluation, prompt engineering, and response-quality assessment — judging how well models actually reason, from the inside.
FALCON, DocuMind AI, and parth-dl ship under the Parthmax name — public on GitHub and Hugging Face, not just claimed but checkable.
Generative AI Engineer, building real-time conversational voice agents and agentic AI systems in production — M.Sc. completed Jun 2026 alongside the role.
Not a buzzword list — this is the actual stack behind FALCON, DocuMind AI, CodePilot, and the voice agents running in production at Zencia AI.
Since August 2025, Saksham has been building production voice AI and agentic systems at Zencia AI — here's the actual stack behind that work.
Speech-to-Text (Whisper, Deepgram), LLM reasoning, Text-to-Speech, streamed over WebRTC and Twilio — with barge-in support and Voice Activity Detection for natural, low-latency conversation.
Adapting transformer architectures for speech generation by replacing text token embeddings with audio token representations.
LangGraph-based systems with tool calling and structured outputs, integrated into Google Calendar, CRM systems, and databases.
Async FastAPI services, WebSockets, background jobs, and logging — deployed via Docker to GCP Cloud Run and Vercel.
Public, checkable projects under the Parthmax name — not just claimed on a page, but live on GitHub, Hugging Face, and PyPI.
An autonomous AI engineering system that turns natural-language specs into full applications: repository-level understanding, architectural planning, RAG-based context retrieval, and self-correcting test/lint loops.
View on GitHub →A multi-stage fact-checking pipeline: extract the claim, cross-reference it against Wikipedia and live search, verify with an LLM, and generate an explainable confidence score.
Try on Hugging Face →Document intelligence over large PDF collections — BGE embeddings, FAISS and ChromaDB retrieval, BM25 and semantic reranking, orchestrated across Gemini, Claude, GPT-4.1, and Llama-family models.
View on GitHub →A lightweight, zero-dependency Instagram media downloader CLI — no login, no API key. Built and maintained in pure Python.
View on PyPI →A Python CLI to search, download, and stream movies and TV series — cyberpunk interactive shell, async API calls, subtitle support, and MPV/VLC streaming.
View on GitHub →A curated collection of 100+ high-impact open-source LLM, RAG, and AI agent projects — actively maintained as the ecosystem moves.
View on GitHub →Formal credentials, algorithmic rigor, and hands-on model-evaluation experience behind the applied AI engineering work.
350+ Data Structures & Algorithms problems solved, strengthening the algorithmic problem-solving behind production system design.
View Profile →M.Sc. Artificial Intelligence & Machine Learning, Aug 2024 – Jun 2026 (completed). B.Sc. Mathematics, Dr. Ram Manohar Lohia Avadh University, 2021 – 2024.