Back to LeaderboardLast snapshot: Sep 9
Musharaf A.

Musharaf A.

Generative AI Engineer | Production RAG, LLM Agents & Vector Search

I build RAG systems and LLM agents that hold up in production. On my last enterprise conversational AI platform, retrieval accuracy went up 37% and response latency dropped by 17 seconds after I rebuilt the chunking and retrieval layer. I'm Lead AI Engineer at a London-based company and take a small number of Upwork contracts alongside it; most recently an 18-month, 716-hour engagement that one client kept renewing. WHAT I BUILD - RAG and conversational AI over private documents: hierarchical, recursive and topic-based chunking, hybrid retrieval (BM25 + dense vectors), re-ranking, sensitive-data guards, and multi-provider LLM failover - LLM agents with real tools: SQL, graph queries, web search, internal APIs using LangGraph, MCP and n8n - Vector database architecture โ€” Milvus, Pinecone, Chroma, MongoDB Atlas, Elasticsearch, index design, quantization, query cost tuning - LLMOps: embedding pipelines, CI/CD for models, MLflow, monitoring, drift detection, inference autoscaling on Azure, GCP and AWS - No-code ML platforms so non-technical teams can upload data, tune models and ship predictions without engineering support RESULTS FROM RECENT WORK Enterprise Conversational AI Platform: Rebuilt chunking and retrieval for a document-grounded assistant, added sensitive-data guards and multi-provider LLM switching. Retrieval accuracy +37%, response latency โˆ’17s. LangChain, LangGraph, MongoDB Atlas, Azure AI Foundry. Crypto Transaction Intelligence: Modeled entities in Memgraph with sanctions screening and IP/VPN detection, and enabled LLM-to-Cypher natural language querying. Cut investigation time from days to hours. Memgraph, LLMs, Binance/OKX/Kraken APIs. Logistics ML: Packing-box selection model that improved packing efficiency 32% and reduced material waste; distributed Random Forest ETA predictor that improved delivery-time accuracy 18%. PySpark, scikit-learn, distributed training. Medical Text Summarization API: Fine-tuned Pegasus for long-form clinical text with multi-GPU distributed training, deployed as a real-time inference API. Pegasus, Hugging Face, multi-GPU. BACKGROUND MS in Computer Science (Information Technology University, Pakistan): Computer Vision, Deep Learning, Big Data. Gold Medalist in BS Information Technology. Fully funded PEEF graduate scholarship. Five years across FinTech and crypto, healthcare, logistics, e-commerce, media analytics and sports. HOW WE START Message me with what you are building and what's blocking you. I'll reply within a few hours with an honest read on whether it's a fit, including when it isn't, and who you should talk to instead.

๐Ÿ‡ต๐Ÿ‡ฐSahiwal, Pakistan$45/hr100% JSS4.09 (3)Since Jun 2016
MRR
$48
๐ŸŒ#140K/363K
๐Ÿ‡ต๐Ÿ‡ฐ#20.8K/47.0K
Recent Earnings
$288
๐ŸŒ#140K/363K
๐Ÿ‡ต๐Ÿ‡ฐ#20.8K/47.0K
Total Earnings
$19,047
๐ŸŒ#135K/363K
๐Ÿ‡ต๐Ÿ‡ฐ#13.3K/47.0K
Avg. Per Project
$2,381
๐ŸŒ#62.4K/363K
๐Ÿ‡ต๐Ÿ‡ฐ#4,518/47.0K
Recent Projects
1
0f1h
๐ŸŒ#203K/363K
๐Ÿ‡ต๐Ÿ‡ฐ#33.1K/47.0K
Total Projects
8
5f9h
๐ŸŒ#203K/363K
๐Ÿ‡ต๐Ÿ‡ฐ#33.1K/47.0K
MRR Performance Over Time
$50k$25k$0
6 mo ago3 mo agoNow
Coming SoonGathering historical data
World Skill Rankings
of 48
#18
Graph DatabaseTop 35%
#28Generative Model
#115ML Automation
#135Neural Network
#164MLOps
#488Vector Database
#530Data Engineering
#1,071LLM Prompt Engineering
#1,328Retrieval Augmented Generation
#1,583LangChain
#1,900Generative AI
#2,662AI Development
#2,935Machine Learning
#3,257n8n
#9,886Python
๐Ÿ‡ต๐Ÿ‡ฐPakistan Skill Rankings
of 10
#5
Graph DatabaseTop 40%
#8Generative Model
#29Neural Network
#48ML Automation
#52MLOps
#110Data Engineering
#185Vector Database
#351LLM Prompt Engineering
#561Retrieval Augmented Generation
#608Generative AI
#662LangChain
#696Machine Learning
#774AI Development
#984n8n
#2,195Python
Musharaf A. โ€” Top 44% in Pakistan | UpworkMRR