I will engineer a rag system that does not hallucinate

H
huzaifa_flutter
H
huzaifa_flutter
Huzaifa Furqan

About this gig

Need a chatbot that actually works, not one that confidently makes stuff up the moment someone asks something unexpected?


That's what I build. RAG and LLM systems engineered to hold up in production, not just look good in a demo.


What I can do for you:


-Build a full RAG chatbot for your website, WhatsApp, or Telegram

-Fast retrieval with semantic caching, repeat questions return in under 50ms

-Failover between multiple LLMs so your bot keeps working if one provider goes down

-Full instrumentation so you can see cost and failures instead of flying blind

-Vector database setup (ChromaDB, Pinecone) and backend

-Clean, documented code


Where this comes from:

I'm currently lead engineer on a production RAG platform a company has adopted internally. I built the query routing, caching, failover, and monitoring myself, fixing problems most chatbot gigs ignore until a client complains.


If you've been burned by a chatbot that hallucinated or broke under real traffic, message me before ordering so we can scope it properly.

Get to know Huzaifa Furqan

Huzaifa Furqan

Production RAG Engineer for Chatbots That Do Not Hallucinate

  • FromPakistan
  • Member sinceJan 2026
  • Languages

    Urdu
✅ RAG & LLM Systems Engineer ✅ Production Chatbot Development ✅ Multi-LLM Failover: Gemini, Llama, Mixtral ✅ Vector Databases: ChromaDB, Pinecone ✅ LLM Observability & Monitoring ✅ Multi-Channel Bots: Web, WhatsApp, Telegram ✅ Semantic Caching & Query Routing 🚀 Recent: Multi-channel RAG chatbot, 4-route query router, sub-50ms cached responses, 4-model failover chain 💼 You Get: Production-grade code, on-time delivery, clear communication Let's build something that actually works!