Pedro – LLM, RAG, LangChain, experts in Lemon.io

Pedro

From Brazil (UTC-3)flag

AI Engineer|Middle-to-senior

Pedro – LLM, RAG, LangChain

Pedro is a strong Mid-to-Senior AI Engineer with strong expertise in Python, production voice agent engineering, and pragmatic system design. He has hands-on experience with multi-agent orchestration, MLOps, and evaluation methodologies, particularly in voice-driven AI products and enterprise-scale ML services. His strengths include calm communication, adaptability, a product-focused mindset and direct stakeholder engagement.

5 years of commercial experience in
AI
Machine learning
AI software
Chatbots
Main technologies
LLM
4 years
RAG
1 year
LangChain
2 years
Python
8 years
Machine learning
5.5 years
Additional skills
LangGraph
PyTorch
Docker
Kubeflow
Kubernetes
ONNX
FastAPI
GCP
AWS SageMaker
Computer Vision
Data Science
Tensorflow
MLOps
Embedded Systems
SQL
Pandas
Flask
Apache Kafka
NumPy
Matplotlib
Scikit-learn
Keras
GPU
Pipecat
GCP Compute Engine
Hugging Face
Workflow Automation
Fine-tuning
Nvidia GPU
Direct hire
Possible
Ready to get matched with vetted developers fast?
Let’s get started today!

Experience Highlights

AI/ML Engineer
Apr 2024 - Ongoing2 years 3 months
Project Overview

An AI-powered simulation platform that enables users to create and run structured persona-based simulations through self-serve workflows. The product combines agentic LLM capabilities with voice simulation orchestration, integrating speech-to-text, LLM, and text-to-speech services to streamline simulation setup and support scalable customer usage.

Responsibilities:
  • Build a self-serve simulation product with a structured, form-based persona workflow, reducing simulation creation time from 1–2 days to about 1 hour and increasing throughput from 1 to 10 simulations per day.
  • Develop an agentic LLM pipeline that converts free-form user prompts into structured persona configurations, reducing setup time from about 1 hour to 40 seconds.
  • Design and implement voice simulation orchestration across third-party and self-hosted components, integrating Speech-to-Text (STT), LLM, and Text-to-Speech (TTS), contributing to a 5x increase in simulation usage and customer adoption.
Project Tech stack:
LangGraph
LangChain
Python
Docker
FastAPI
Pipecat
Machine Learning Engineer
Oct 2021 - Apr 20242 years 5 months
Project Overview

An enterprise AI and machine learning platform providing services for clustering, classification, vector embeddings, and text and image generation. The platform supports production-grade AI workloads for enterprise customers, with scalable model deployment and optimization capabilities.

Responsibilities:
  • Led development of 20+ ML/AI services across clustering, classification, vector embeddings, and text/image generation, serving 20+ enterprise customers with Python, NumPy, Scikit-Learn, PyTorch, Hugging Face, ChatGPT, and Llama.
  • Built and operationalized MLOps workflows from development to production using Docker, Kubernetes, AWS SageMaker, GCP Cloud Run, FastAPI, ONNX optimization, and model quantization, improving model speed by up to 30%.
  • Delivered scalable AI solutions for enterprise use cases, supporting reliable deployment and production operation across diverse ML workloads.
Project Tech stack:
PyTorch
Python
Hugging Face
Docker
Workflow Automation
GCP
GCP Compute Engine
Machine Learning Engineer
Jun 2021 - Sep 20213 months
Project Overview

A surveillance solution that integrates machine learning applications to support intelligent monitoring and analysis. The platform combines video surveillance capabilities with ML-powered functionality for more efficient and automated monitoring.

Responsibilities:
  • Built NVIDIA GPU-accelerated Kubernetes infrastructure for computer vision workloads, with reusable Kubeflow components for training, inference, and monitoring.
  • Developed reusable Kubeflow components for model training, inference, and monitoring.
  • Deployed a hard-hat detection model on in-house infrastructure with approximately 95% accuracy.
Project Tech stack:
Kubeflow
Kubernetes
Nvidia GPU
Fine-tuning
Computer Vision

Education

2017
Electrical and Electronics Engineering
Bachelor's degree

Languages

English
Advanced

Hire Pedro or someone with similar qualifications in days
All developers are ready for interview and are are just waiting for your requestdream dev illustration
Copyright © 2026 lemon.io. All rights reserved.