Back to LeaderboardLast snapshot: Sep 8
Ross F.
Badge

Ross F.

AI Integration Engineer | MCP Servers - Agents - RAG - Python

Agents don't fail because the model is bad. They fail because the tools they're handed are unnavigable, the retrieval is unmeasured, and nothing catches a regression before the client does. I build the layer underneath: production MCP servers (Model Context Protocol), RAG pipelines with measured accuracy, and the eval harnesses that keep both honest. ๐Ÿ† Top Rated Plus ยท $350K+ earned ยท 8,000+ hours billed โ”โ”โ”โ”โ”โ”โ”โ”โ”โ”โ”โ”โ”โ”โ”โ”โ”โ”โ”โ”โ”โ” RECENT PRODUCTION RESULTS ๐Ÿ”Œ Built and shipped a production MCP server (FastMCP, Python) that gives a client's analysts direct agent access to their own domain data โ€” questions that used to need an engineer now get answered in the chat window. It runs in production behind an agent service on AWS Fargate over stdio, and in Claude Desktop and Claude Code. I designed the tool surface for progressive discovery โ€” broad list, then filter and count, then drill down โ€” so models navigate 15K+ records without blowing their context, and built credential-gated tool registration with a read-only-by-design data layer so pointing an agent at live data is safe. 850+ tests, and an eval harness I run across model versions before shipping changes, so a model upgrade can't silently break agent behaviour. ๐Ÿค Wrote the agent that drives it, too โ€” a project-level Claude Code subagent with a curated tool allowlist, in-prompt gates derived from real production failures, and a cost-tier routing matrix that picks the cheapest transport likely to work. Building the tool surface and the agent that consumes it is a different skill from wiring up one API. ๐Ÿค– Built a RAG extraction service (FastAPI + Celery + Pinecone, two-tier model routing with a per-model cost estimator) turning messy documents into structured data across ~11K projects from 11 registry sources. Measured on a golden dataset, accuracy went from 33% to 91% F1 on one extraction task and 43% to 84% on another โ€” the difference between a pipeline nobody trusted and one the team runs unattended. Every output traces back to its source document. ๐Ÿง  Built a second production RAG system on a medical knowledge graph (FastAPI, Neo4j, MongoDB) with character-level span citations, so a disputed claim takes seconds to check rather than an afternoon. Application-layer tenant isolation with dedicated tests proving no cross-tenant leakage. 2,700+ tests; I wrote roughly two-thirds of the codebase. ๐ŸŒ Built and operate a scraping platform covering 20 sources behind enterprise anti-bot protection โ€” Cloudflare-class WAFs and Incapsula, a rotating datacenter proxy pool plus a residential tier for the hardest targets, and fallback transport chains that step up only when they have to. 225K+ documents collected to date. 50+ scheduled pipelines, around 30 of them daily, with per-source error recovery and alerting. โ”โ”โ”โ”โ”โ”โ”โ”โ”โ”โ”โ”โ”โ”โ”โ”โ”โ”โ”โ”โ”โ” WHAT I DO โœ”๏ธ AI integration & agent infrastructure โ€” MCP (Model Context Protocol) servers with FastMCP, tool surfaces designed for how models actually search, Claude Desktop and Claude Code integrations, custom Claude Code subagents, prompt engineering, eval harnesses โœ”๏ธ LLM & RAG backends โ€” Anthropic/OpenAI/Gemini APIs, Pinecone, vector search with RRF fusion, structured extraction from messy documents, measured accuracy against golden datasets, cost routing that sends the easy 80% to cheap models โœ”๏ธ Web scraping & data extraction โ€” Playwright, Selenium, ZenRows; resilient access to protected sources, proxy management, scheduled fleets via Celery, PDF/Excel/Word extraction, normalization into clean schemas โœ”๏ธ API development โ€” FastAPI, Flask, Django; auth, rate limiting, background jobs, clean documentation โœ”๏ธ Distributed systems โ€” Celery, RabbitMQ, Redis; retries, idempotency, fault tolerance under real load. Redis caching at 85-95% hit rate, typically 10-50x faster responses โœ”๏ธ Production ops โ€” Docker, AWS, PostgreSQL/MongoDB/Neo4j, 110+ zero-downtime migrations on a single project, CI-gated test suites running 2,500-4,900 tests on my largest systems โ”โ”โ”โ”โ”โ”โ”โ”โ”โ”โ”โ”โ”โ”โ”โ”โ”โ”โ”โ”โ”โ” HOW I WORK โœ… I own systems end-to-end: architecture โ†’ implementation โ†’ deployment โ†’ monitoring โ†’ handover docs. โœ… I'll tell you when an LLM is the wrong tool โ€” and what to use instead. Cheaper for both of us than finding out in week three. โœ… Failures surface where you'll see them: Prometheus/AlertManager into Slack with per-alert templates, Grafana and Loki for dashboards and logs, scheduled digests and a daily data-feed tripwire. I get paged, not you. โœ… Most of my $350K+ comes from repeat clients and multi-year engagements. โ”โ”โ”โ”โ”โ”โ”โ”โ”โ”โ”โ”โ”โ”โ”โ”โ”โ”โ”โ”โ”โ” ๐Ÿ“ UK-based (GMT/BST) If you need agent tooling that real models can navigate, an LLM pipeline whose accuracy you can actually check, or scraping infrastructure that survives contact with real anti-bot systems โ€” send me a couple of lines about your project and I'll tell you straight away whether I'm the right fit.

๐Ÿ‡ฌ๐Ÿ‡งLeicester, United Kingdom$60/hr100% JSS4.99 (51)Since Nov 2018
MRR
$14,642
๐ŸŒ#492/363K
๐Ÿ‡ฌ๐Ÿ‡ง#17/9,432
Recent Earnings
$87,854
๐ŸŒ#492/363K
๐Ÿ‡ฌ๐Ÿ‡ง#17/9,432
Total Earnings
$356K
๐ŸŒ#5,190/363K
๐Ÿ‡ฌ๐Ÿ‡ง#112/9,432
Avg. Per Project
$3,745
๐ŸŒ#42.5K/363K
๐Ÿ‡ฌ๐Ÿ‡ง#731/9,432
Recent Projects
9
4f5h
๐ŸŒ#27.9K/363K
๐Ÿ‡ฌ๐Ÿ‡ง#688/9,432
Total Projects
95
59f43h
๐ŸŒ#27.9K/363K
๐Ÿ‡ฌ๐Ÿ‡ง#688/9,432
MRR Performance Over Time
$50k$25k$0
6 mo ago3 mo agoNow
Coming SoonGathering historical data
World Skill Rankings
of 134
#1
Neo4jTop <0.01%
#1Celery
#2Web Scraping
#2Flask
#3Selenium
#5Data Scraping
#5Prompt Engineering
#8AI Model Integration
#12DevOps
#13FastAPI
#16Retrieval Augmented Generation
#17Docker
#18Claude
#26PostgreSQL
#38Automation
#49AI Agent Development
#92Python
๐Ÿ‡ฌ๐Ÿ‡งUnited Kingdom Skill Rankings
of 66
#1
Web ScrapingTop <0.01%
#1Data Scraping
#1Neo4j
#1Celery
#1DevOps
#1AI Model Integration
#1FastAPI
#2Selenium
#2Flask
#2Docker
#2Retrieval Augmented Generation
#2Prompt Engineering
#2PostgreSQL
#3Claude
#4Python
#4Automation
#4AI Agent Development
Ross F. โ€” Top 1% in United Kingdom | UpworkMRR