Open to remote hire

AI Full-Stack Engineer
for Hire

Generative AI & LLMs  ·  RAG Pipelines  ·  Agentic Workflows  ·  FastAPI

I'm Ramesh Kumar Das. Over 6.5+ years I've built AI features that sit inside real products — backends, web apps, and the glue between them. My focus is Generative AI, LLMs, RAG pipelines, and Agentic AI workflows: the parts that have to work under load, not just in a notebook. At Kliky AI (Kuwait) I lead AI on WinstaAI, our live agentic SaaS with multi-model routing, text-to-image, video AI, and autonomous jobs. If you need someone who can go from prompt design to a deployed API, let's talk.

AI Full-Stack Engineer Generative AI Developer LLM Engineer RAG Pipelines Agentic AI Nepal  ·  Remote 6.5+ Years
Portrait of Ramesh Kumar Das, AI full-stack engineer

6.5+
Years Experience
10+
AI Systems Built
5+
LLM Integrations
100%
Remote Ready

AI & Full-Stack Tech Stack

Tools I use from model wiring through deploy

Agentic AI & Generative AI
LangChainLlamaIndexOpenAI GPT APIsDeepSeekHugging FaceOllamaRAG PipelinesVector EmbeddingsSemantic SearchAI AgentsAgentic WorkflowsTool CallingPrompt EngineeringMulti-Model Systems
AI Media & Processing
Text-to-ImageComfyUIReplicateFAL.aiRunPod (GPU)Speech-to-Text / TTSOCRFace SwapVideo AI
Backend & APIs
FastAPIDjangoNode.jsPython3REST APIsWebSocketsCeleryAsync ProgrammingMicroservices
Frontend
React.jsNext.jsTypeScriptTailwind CSS
Databases & Vector Stores
PostgreSQLMongoDBRedisPineconeWeaviateSupabase
Cloud & DevOps
DockerAWS EC2/S3GPU ServersNginxGitHub ActionsCI/CD

AI Engineering Services

Concrete work I take on when a product needs AI built in

🤖 Generative AI Platform Development

Full AI SaaS builds: LLM hooks, credits and billing, admin panels, usage stats, and routing across several models. FastAPI services on the back, Next.js on the front.

🔍 RAG Pipeline Architecture

Retrieval stacks with LangChain or LlamaIndex, vector stores like Pinecone or Weaviate, and semantic search so the model answers from your own docs — not generic training data.

🧠 Agentic AI Workflows

Agents that call tools, hit APIs, and run multi-step jobs with light human oversight. LangChain agents, tool calling, and custom orchestration where frameworks fall short.

🎨 AI Media Generation Systems

Text-to-image (ComfyUI, Replicate, FAL.ai), face swap, video and audio processing, OCR, and speech-to-text / TTS. GPU runs on RunPod and similar hosts.

🌐 AI-Powered SaaS Backends

APIs built for AI workloads: Celery queues, WebSockets for streaming output, usage meters, rate limits, and sensible GPU budgeting.

🔗 AI API Integration

Wire OpenAI, DeepSeek, Anthropic, Hugging Face, and others into apps you already have. Prompts, response checks, cost control, and fallbacks when a provider flakes.


Work Experience

Where I've shipped AI into live systems over 6.5+ years

Lead AI Full-Stack Developer
Kliky AI
2024 – Present

WinstaAI — Agentic AI SaaS Platform (Kuwait)

I own the AI platform build: generative features, agent workflows, and routing across multiple LLMs in a product customers use every day.

  • Shipped agentic AI pipelines for tool calling, branching decisions, and multi-step jobs with LangChain plus custom orchestration
  • Connected text-to-image (ComfyUI, Replicate), face swap, OCR, video/audio AI, and model routing across OpenAI, DeepSeek, and Ollama
  • Built the admin dashboard for service toggles, credits, usage stats, and billing
  • Hooked up GPU inference on RunPod and FAL.ai with dynamic allocation so spend stays under control
Full-Stack Developer
AJATH INFO TECH PVT. LTD.
2021 – 2023

Dibbly — Multilingual AI Video Translation  |  Paperport — AI Content Generator

  • Built multilingual video translation with speech-to-text, TTS, DeepL, and live transcription
  • Added LLM content generation that turns rough text into presentations and roadmaps
  • Hardened API traffic with sessions, CSRF protection, and JWT auth around the AI endpoints
Full-Stack Developer
SARAJ SYSTEM PVT. LTD.
2019 – 2021

Khantailor & PathHub

  • Shipped full-stack platforms with React.js, Node.js, and Django
  • Added real-time sync with WebSockets and Redis (about 30% better synchronization in practice)
  • Cut API latency roughly 20–30% with backend tuning and caching

Frequently Asked Questions

Questions I get from hiring managers and founders

What's the difference between Generative AI and Agentic AI work?
Generative AI is about plugging in models like GPT or Stable Diffusion to create text, images, or audio. Agentic AI goes a step further: systems that decide next steps, call APIs, and finish multi-step jobs with little babysitting. I build both.
Can you add AI to software we already have?
Yes. Chatbots, document Q&A with RAG, auto-generated content, recommendation flows — I'll look at your current stack and wire AI in without tearing the product apart.
Which LLMs have you used in production?
OpenAI GPT-4 / GPT-4o, DeepSeek, Anthropic Claude, Hugging Face models, Ollama for local runs, and specialist models via Replicate. I often route tasks across models so you get the right quality for the price.
How do you keep AI costs and rate limits under control?
Token budgets, response caching, fallback model chains, Celery queues, and per-user or per-org usage tracking. That combo usually drops API spend while keeping the product responsive.
Do you offer consulting or architecture reviews?
Yes — model and tool selection, RAG design, agent workflows, and whether a feature even needs an LLM. WhatsApp +977-9700864900 is the fastest way to start.

Get in Touch

AI engineering roles, product contracts, and consulting — remote worldwide

💼 Also explore:  Hire me as Full-Stack Developer →  |  Hire me as Backend Engineer →  |  View full portfolio →
50% off — direct hire

Solid AI engineering. Direct rates.

LLM, RAG, and agentic systems without agency markups. Current rates already include 50% off for direct clients.

Hourly Consulting
Market Standard: $60 - $80
$29-50% Off
Pay-as-you-go · Billed weekly

  • LLM wiring and RAG pipeline setup
  • Architecture and code review
  • Prompt work and cost tuning
  • Minimum 4 hours per booking
  • Async chat or video calls
Book hourly work
Fixed-Price MVP
Market Standard: $3,000+
From $1,499Fixed
Milestone-based · Clear deliverables

  • AI product MVP in about 4–6 weeks
  • Custom RAG chatbot ready to run
  • Secure LLM API integrations
  • Scope locked before work starts
  • Invoices at finished milestones
Request a quote

💡 Direct clients only: These prices are for hiring me straight — no agency cut. They already include my current 50% discount.

Global Tech Hiring Hubs

Hire Developers USA Hire Developers Canada Hire Developers UK AI Engineer San Francisco FastAPI Developer NYC Hire Developers Middle East Hire Developers Saudi Arabia AI Engineer Dubai FastAPI Developer Riyadh Hire Developers Singapore Hire Developers Australia Hire Developers Germany Hire Developers Japan Backend Architect Berlin

Hire AI Engineer · Related Expertise

Hire Agentic AI Engineer Hire LLM Engineer Hire Generative AI Developer Hire AI Integration Expert LLM Orchestration Expert RAG Pipeline Developer LangChain Developer LangGraph Engineer CrewAI Specialist AutoGen Developer LlamaIndex Expert Multi-Agent Systems Engineer Prompt Engineering Expert ReAct Pattern Developer OpenAI API Integration Anthropic Claude API Expert AI Chatbot Developer Voice AI Developer AI Product Engineer LLMOps Engineer MLOps Engineer vLLM Deployment Expert LoRA · PEFT Fine-Tuning Mistral · Llama Fine-Tuner LLM Inference Optimization Model Evaluation Engineer Vector Database Engineer Pinecone · Weaviate · Qdrant ChromaDB Developer Semantic Search Expert Embedding Pipeline Expert Hybrid Search Developer RAG Evaluation Expert Full-Stack AI Engineer Agentic Workflow Automation FastAPI AI Backend AI Business Automation AI for Fintech Healthcare AI Automation AI Startup Developer Freelance AI Developer Remote AI Engineer · Nepal On-Demand AI Engineer Hire · Contact Ramesh Das
WhatsAppHire & projects