Available for remote AI engineering work — New York, EST, PST, and US-friendly hours

Agentic AI Engineer
for New York Teams

LLM Applications  ·  RAG Systems  ·  Autonomous AI Workflows  ·  FastAPI

I am Ramesh Kumar Das, a dedicated AI Full-Stack Engineer bringing over 6.5 years of experience in architecting and deploying resilient, production-ready web platforms and AI-driven SaaS applications. I partner with New York-based startups and established product teams to transform ambitious AI concepts into reliable software. My technical expertise spans advanced LLM integrations, sophisticated RAG pipelines, autonomous agentic workflows, FastAPI microservices, and highly responsive Next.js interfaces. Currently, my primary focus at Kliky AI is the development of WinstaAI—a comprehensive agentic SaaS platform that features intelligent multi-model routing, advanced media generation pipelines, workflow automation, and robust operational tooling.

The value I bring is practical engineering, not AI demos that fall apart in production. I focus on clean architecture, controlled API costs, secure data handling, reliable background jobs, useful admin controls, and deployment paths that make sense for real users.

Agentic AI Engineer Generative AI Developer LLM Engineer RAG Pipeline Developer AI Product Engineer New York / US Remote 6.5+ Years
Ramesh Kumar Das, agentic AI engineer available for New York and US remote teams

6.5+
Years Building Software
10+
AI Features and Systems Delivered
5+
LLM and Model Provider Integrations
100%
Remote-Ready for US Teams

AI and Full-Stack Engineering Stack

Technologies I use to plan, build, deploy, and maintain production AI applications.

Agentic AI and LLM Systems
LangChainLlamaIndexOpenAI GPT APIsDeepSeekHugging FaceOllamaRAG PipelinesVector EmbeddingsSemantic SearchAI AgentsAgentic WorkflowsTool CallingPrompt EngineeringMulti-Model Systems
AI Media and Document Processing
Text-to-ImageComfyUIReplicateFAL.aiRunPod GPUSpeech-to-Text / TTSOCRFace SwapVideo AI
Backend and APIs
FastAPIDjangoNode.jsPython3REST APIsWebSocketsCeleryAsync ProgrammingMicroservices
Frontend
React.jsNext.jsTypeScriptTailwind CSS
Databases and Vector Search
PostgreSQLMongoDBRedisPineconeWeaviateSupabase
Cloud and Deployment
DockerAWS EC2/S3GPU ServersNginxGitHub ActionsCI/CD

AI Engineering Services for New York Companies

Hands-on AI development for SaaS teams, founders, agencies, and businesses adding automation to existing products.

🤖 Generative AI Product Development

I build AI SaaS products with LLM APIs, account systems, credits or billing logic, admin panels, usage analytics, and model orchestration. A typical build uses FastAPI or Django on the backend with a React or Next.js frontend for customers and internal teams.

🔍 RAG Pipeline Architecture

I design retrieval-augmented generation systems that help users query private documents, internal knowledge bases, product data, and support content. This includes document cleanup, chunking, embeddings, vector search, prompt design, access control, and answer quality checks.

🧠 Agentic AI Workflows

I create AI agents that can plan steps, call tools, use APIs, process files, and complete structured tasks. These workflows can support research assistants, internal operations, content pipelines, customer support processes, and multi-step business automation.

🎨 AI Media Generation Systems

I work with text-to-image, face swap, OCR, video AI, speech-to-text, and text-to-speech workflows using tools such as ComfyUI, Replicate, FAL.ai, and RunPod. I also help structure GPU usage so media generation features stay practical to operate.

🌐 AI SaaS Backend Development

I build backend systems for AI products, including queues, WebSocket streaming, rate limits, user permissions, usage metering, payment-aware credits, retries, logs, and background workers for long-running model tasks.

🔗 AI API Integration

I connect OpenAI, DeepSeek, Anthropic, Hugging Face, Replicate, and other AI providers to new or existing applications. My focus is clean integration, safe output handling, fallback logic, cost visibility, and prompts that can be tested instead of guessed.


Relevant Work Experience

Production experience across AI platforms, SaaS products, backend systems, and full-stack applications.

Lead AI Full-Stack Developer
Kliky AI
2024 – Present

WinstaAI — Agentic AI SaaS Platform

I lead development across the AI product stack, including LLM workflows, media generation pipelines, backend services, dashboards, and admin tooling.

  • Built agentic AI workflows for tool use, autonomous task execution, branching decisions, and multi-step orchestration with LangChain and custom backend logic.
  • Integrated text-to-image, face swap, OCR, video AI, audio AI, and LLM routing across OpenAI, DeepSeek, Ollama, ComfyUI, and Replicate.
  • Developed an AI operations dashboard for service control, user credits, billing visibility, usage monitoring, and admin management.
  • Connected GPU inference services such as RunPod and FAL.ai with workflows designed for reliability and controlled operating cost.
Full-Stack Developer
AJATH INFO TECH PVT. LTD.
2021 – 2023

Dibbly and Paperport — AI Video Translation and Content Generation

  • Built multilingual AI video translation workflows using speech-to-text, text-to-speech, DeepL, and transcription processing.
  • Implemented LLM-assisted content generation for converting unstructured inputs into presentations, summaries, and structured roadmaps.
  • Worked on secure AI API communication, authentication, session handling, CSRF protection, and JWT-based access control.
Full-Stack Developer
SARAJ SYSTEM PVT. LTD.
2019 – 2021

Khantailor and PathHub

  • Built and maintained full-stack platforms using React.js, Node.js, and Django.
  • Added real-time features with WebSockets and Redis to support live updates across users.
  • Improved backend performance through API optimization, caching, and database-aware engineering.

Frequently Asked Questions

Answers for New York founders, hiring managers, and product teams evaluating an AI engineer.

Can you work with New York teams remotely?
Yes. I work remotely with US-friendly availability, including overlap with New York business hours. Communication can happen through WhatsApp, email, Slack, GitHub, task boards, or scheduled calls depending on the team’s workflow.
How is agentic AI different from regular generative AI?
Generative AI creates content such as text, images, code, audio, or summaries. Agentic AI goes a step further by planning actions, choosing tools, calling APIs, retrieving data, processing files, and completing multi-step tasks with less manual input.
Can you add AI features to an existing product?
Yes. I can add AI chat, document Q&A, RAG search, content generation, internal assistants, recommendation features, workflow automation, and backend model integrations to existing applications. I usually begin with a technical review so the integration path is safe and realistic.
Which LLMs and AI providers have you worked with?
I have worked with OpenAI GPT models, DeepSeek, Anthropic Claude, Hugging Face models, Ollama for local inference, and specialized AI services through Replicate and similar platforms. I can also design routing logic based on task type, response speed, quality, and cost.
How do you manage AI API cost and rate limits?
I use token budgeting, caching, fallback models, queues, usage tracking, rate limits, retries, batching, and model selection rules where appropriate. These controls matter because an AI feature can work well in testing but become expensive or unstable under real production traffic.
Do you offer AI consulting before development begins?
Yes. I can review your product idea, existing architecture, data sources, model choices, RAG design, agent workflow, deployment plan, and cost risks before a full build begins. For quick discussions, WhatsApp is the fastest way to reach me at +977-9700864900.

Need an Agentic AI Engineer for a New York Product?

I’m available for remote AI engineering roles, product builds, architecture consulting, and fixed-scope MVP work.

💼 Related hiring pages:  Full-Stack Developer  |  Backend Engineer  |  Full Portfolio
Direct-Hire AI Engineering

AI Development Support Without Agency Overhead

Flexible options for LLM integrations, RAG systems, agentic workflows, AI SaaS builds, and architecture review for New York and US teams.

Hourly Consulting
New York Agency: $60 - $80
$29-50% Off
Pay as you go · Billed weekly

  • LLM integration and RAG setup guidance
  • Architecture review and code review
  • Prompt design, testing, and refinement
  • Minimum 4 hours per session
  • Async support or video call support
🚀 Book an Hourly Session
Fixed-Price MVP
New York Agency: $3,000+
From $1,499Fixed
Milestone based · Scope agreed upfront

  • AI product MVP build in 4–6 weeks
  • Custom RAG chatbot or AI assistant
  • Secure LLM API integrations
  • Clear deliverables before development starts
  • Invoices tied to completed milestones
🚀 Request a Project Quote

💡 Direct-hire note: Work with me directly for AI engineering, product development, and consulting without adding an agency layer. Final pricing depends on scope, timeline, integrations, and technical complexity.

More New York hiring profiles

United States hiring hub Full-Stack Developer in New York AI Full-Stack Engineer in New York Backend Engineer in New York FastAPI Developer in New York LLM Engineer in New York RAG Developer in New York Backend Architect in New York

Global tech hiring hubs

Hire Developers in the USA Hire Developers in Canada Hire Developers in the UK AI Engineer in San Francisco FastAPI Developer in New York Hire Developers in the Middle East Hire Developers in Saudi Arabia AI Engineer in Dubai FastAPI Developer in Riyadh Hire Developers in Singapore Hire Developers in Australia Hire Developers in Germany Hire Developers in Japan Backend Architect in Berlin

Common AI and Software Hiring Needs in New York

New York companies often need engineers who can connect AI strategy with reliable product development, secure APIs, and scalable SaaS architecture.

🤖 AI and Agentic Systems

  • • Agentic AI engineering
  • • LLM application development
  • • RAG pipeline implementation
  • • Generative AI product features
  • • Multi-agent workflow design

⚙️ Backend and Architecture

  • • FastAPI backend development
  • • Backend architecture review
  • • Python microservices
  • • Secure API engineering
  • • Node.js backend support

💻 Full-Stack and Frontend

  • • React and Next.js development
  • • Full-stack SaaS engineering
  • • Remote software development
  • • Web app development
  • • Tailwind CSS implementation

📈 SaaS and Consulting

  • • AI automation consulting
  • • SaaS MVP development
  • • Enterprise software support
  • • Technical product partner work
  • • Direct-hire alternative to agency teams

Related AI engineering expertise

Agentic AI Engineer LLM Engineer Generative AI Developer AI Integration Specialist LLM Orchestration RAG Pipeline Development LangChain Development LangGraph Workflows CrewAI Projects AutoGen Development LlamaIndex Systems Multi-Agent Systems Prompt Engineering ReAct Pattern Workflows OpenAI API Integration Anthropic Claude API AI Chatbot Development Voice AI Development AI Product Engineering LLMOps Support MLOps Support vLLM Deployment LoRA and PEFT Fine-Tuning Mistral and Llama Fine-Tuning LLM Inference Optimization Model Evaluation Vector Database Engineering Pinecone · Weaviate · Qdrant ChromaDB Development Semantic Search Embedding Pipelines Hybrid Search RAG Evaluation Full-Stack AI Engineer Agentic Workflow Automation FastAPI AI Backend AI Business Automation AI for Fintech Healthcare AI Automation AI Startup Development Freelance AI Developer Remote AI Engineer · Nepal On-Demand AI Engineer Contact Ramesh Das
WhatsAppHiring and projects