Dmitri M.
AI Engineer | Flink | Kafka | Python | Java | AWS
Your AI is only as good as the data behind it. I build both. Most AI projects stall because the data isn't ready: it's messy, slow or stuck in silos. I'm an AI Engineer with a strong Data Engineering background. I build LLM apps and AI agents on top of real-time streaming pipelines and cloud data platforms, so your AI features work in production and don't stop at the demo. Check my skills ๐ ๐ AI Engineering: โ LLM application development with OpenAI, Anthropic Claude, and open-source models โ RAG (Retrieval-Augmented Generation) pipelines: chunking, embeddings, hybrid search and re-ranking โ AI agents and tool-calling workflows using LangChain, LlamaIndex and MCP (Model Context Protocol) โ Vector databases: pgvector, Pinecone, Qdrant, OpenSearch and Elasticsearch โ Prompt engineering, LLM evaluation, cost and latency tuning โ Deploying AI services with FastAPI, Docker and AWS Lambda/ECS ๐ Streaming & Real-Time Data: โ Apache Flink (PyFlink and Java) for stateful stream processing at scale โ Apache Kafka for event-driven architectures and real-time ingestion โ Real-time feature and embedding pipelines that feed AI and ML systems ๐ Programming Languages: โ Python, Java, Golang and Node.js โ Complex SQL, query tuning and database performance optimization ๐ Cloud Data Platforms: โ AWS: Lambda, Step Functions, S3, Glue, Athena, EMR, Redshift, EC2, ECS, ECR and CloudFormation โ Lakehouse: Apache Iceberg on S3 โ GCP: BigQuery and Looker Studio ๐ Databases: โ SQL: PostgreSQL, MySQL, MSSQL and SQLite โ NoSQL: MongoDB, Redis, Elasticsearch and DynamoDB ๐ Data Stack & DevOps: โ Spark, dbt and Apache Airflow for batch processing, data modeling and orchestration โ Docker and Kubernetes for containerized deployments โ Git, GitHub and GitLab CI/CD ๐ Data Collection & Automation: โ Web scraping with Scrapy, BeautifulSoup, Selenium and Playwright to collect AI training and RAG data โ Social media data pipelines for analytics and AI enrichment โก What I can build for you: โ AI chatbots and assistants trained on your company's documents and databases โ Real-time streaming pipelines that feed AI models and dashboards โ Moves from legacy ETL to modern cloud data platforms โ End-to-end AI products, from data ingestion to deployed API โก Why clients work with me: โ Production-first: I build systems that are monitored, scalable and cost-aware โ Clear communication, with regular progress updates and documented code โ Flexible hours that overlap with US, EU and APAC time zones ๐ฌ Have an AI idea or a data pipeline problem? Send me a message with a short description of your project. I'll reply with a clear plan, the tech I'd recommend and a realistic timeline, free of charge. Let's turn your data into something your business can use.