Skip to content
View vitorjpc10's full-sized avatar

Block or report vitorjpc10

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please donโ€™t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this userโ€™s behavior. Learn more about reporting abuse.

Report abuse
vitorjpc10/README.md

Hi there! Hi

I'm Vitor Cavalcante, a Senior AI Engineer specializing in LLM systems, Retrieval-Augmented Generation (RAG), and Agentic AI architectures.

I design and deploy production-grade AI systems that combine large language models, vector search, and scalable cloud infrastructure to deliver reliable, explainable, and high-precision AI solutions.


๐Ÿง  About Me

I focus on building intelligent systems that go beyond prompting โ€” architecting full AI pipelines with retrieval, ranking, orchestration, and infrastructure automation.

๐Ÿš€ What I Work On

  • Designing high-precision RAG systems with contextual and hybrid retrieval
  • Building agentic AI workflows with structured reasoning and tool orchestration
  • Optimizing vector search, chunking strategies, and semantic re-ranking
  • Engineering AI-ready data pipelines for LLM consumption
  • Automating AI infrastructure using Terraform and CI/CD

๐Ÿ—๏ธ Core Expertise

AI Engineering

  • LLMs (GPT-style models)
  • Retrieval-Augmented Generation (RAG)
  • Agentic AI & Tool-Using Systems
  • Prompt Engineering
  • Vector Search & Hybrid Retrieval
  • Semantic Re-ranking
  • LangChain & Structured Output Systems

Backend & APIs

  • Python (FastAPI, Pydantic)
  • Java (Spring Boot)
  • Microservices Architecture
  • RESTful APIs
  • Data Modeling & Schema Design

Data & Platforms

  • Databricks (PySpark)
  • SQL
  • Data Quality & Observability
  • Metadata & Lineage Automation

Cloud & Infrastructure

  • Azure (Azure OpenAI, AI Search, AI Foundry)
  • AWS
  • GCP
  • Docker
  • Terraform
  • CI/CD Pipelines

๐ŸŽ“ Education

B.S. in Electrical and Computer Engineering
Minor in Applied Mathematics
GPA 3.9 โ€“ Summa Cum Laude


๐ŸŒŽ Languages

  • English (Native)
  • Portuguese (Native)
  • Spanish (Intermediate)

๐ŸŽง Beyond Tech

When Iโ€™m not building AI systems, Iโ€™m into audio engineering, music production, cooking, and exploring global cuisines.


๐Ÿ”ง Technologies & Tools

Languages

Java Python Scala Go SQL JavaScript HTML5 CSS3

Libraries & Frameworks

SBT Gradle JUnit Spring React

Infrastructure & DevOps

AWS GCP Docker

Airflow Snowflake dbt

Environments & IDEs

IntelliJ IDEA Git Postman VSCode PyCharm


๐Ÿ“Š GitHub Stats



๐Ÿ“ซ How to reach me:

LinkedIn Outlook


Profile views

Pinned Loading

  1. ModularChatBot-Dash ModularChatBot-Dash Public

    A modular chatbot framework with a Plotly Dash interface for building, testing, and visualizing conversational AI agents.

    Python 1

  2. etl-breweries etl-breweries Public

    Brewery Data Pipeline - This project implements a data pipeline to fetch, transform, and persist brewery data from the Open Brewery DB API into a data lake, following the medallion architecture (brโ€ฆ

    Python

  3. milhas_whatsapp_agent milhas_whatsapp_agent Public

    AI Agent scraping promotional websites and sharing thorugh whatsapp

    Python

  4. ETL-Pipeline--dbt--Snowflake--Airflow- ETL-Pipeline--dbt--Snowflake--Airflow- Public

    This project demonstrates how to build an ELT pipeline using dbt, Snowflake, and Airflow. Follow the steps below to set up your environment, configure dbt, create models, macros, tests, and deploy โ€ฆ

    Python 4

  5. Spark-Application-with-Python-Using-MongoDB-and-PySpark Spark-Application-with-Python-Using-MongoDB-and-PySpark Public

    This project is an ETL pipeline that fetches market data from the Albion Online Data API, processes it with PySpark, and stores it in MongoDB. It demonstrates real-time data extraction, transformatโ€ฆ

    Python 4

  6. ETL-GDP-of-South-American-countries-using-the-World-Bank-API ETL-GDP-of-South-American-countries-using-the-World-Bank-API Public

    This project is a comprehensive data pipeline designed to extract Gross Domestic Product (GDP) data for South American countries from the World Bank API, transform it into a structured format, and โ€ฆ

    Python 1