I will deploy, finetune, and optimize local llms and vlms

A
akashbhansali
A
akashbhansali
Akash Bhansali

About this gig

High-Performance AI, Optimized for Your Infrastructure


Looking to run powerful open-source models without breaking the bank on cloud costs? I specialize in making Large Language Models (LLMs) and Vision-Language Models (VLMs) faster, lighter, and smarter.


️ What I Do:

  • Custom Fine-Tuning: Tailoring models (Llama, Mistral, Qwen, Phi) to your specific dataset using advanced techniques like LoRA/QLoRA (PEFT).
  • Quantization: Reducing model size (GGUF, AWQ, EXL2) to run efficiently on consumer hardware or smaller cloud instances without sacrificing accuracy.
  • Inference Optimization: Setting up ultra-fast inference frameworks (vLLM, TensorRT-LLM, Ollama) for maximum throughput.

Why Choose Me?

  • End-to-End Expertise: From raw dataset preparation to production-ready deployment.
  • Cost-Conscious: Focused on reducing your production VRAM footprint and compute costs.
  • Clean, Documented Code: Full delivery with deployment scripts.


Get to know Akash Bhansali

Akash Bhansali

Whatever it Takes

5.0(8)
  • FromIndia
  • Member sinceJan 2021
  • Avg. response time5 hours
  • Last delivery2 years
  • Languages

    English, Hindi, German
Hi, I'm Akash, an AI & Robotics Engineer specializing in bridging the gap between simulation and real-world deployment. With an M.Sc. in AI & Robotics, I build high-performance pipelines. What I do: 🔹 Physical AI: Sim2Real & Real2Sim Robotics, MuJoCo, Isaac Sim/Lab, ROS, VLA and control workflows. 🔹 Applied ML & Vision: Real-time tracking (SAM2), detection, tracking, ViTs, etc. 🔹 LLM/VLM Systems: Fine-tuning (LoRA), optimization, quantization, and intelligent agents. 🔹 Research & Implementation 📩 Message me today to discuss your project requirements before placing an order!