Skip to content
View MarcosMaio's full-sized avatar
😄
Focusing
😄
Focusing

Block or report MarcosMaio

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
MarcosMaio/README.md

Marcos Maio

Typing SVG

LinkedIn Email Location


About

AI Engineer specialized in AI Security & GenAI Governance, with hands-on experience protecting production AI ecosystems in regulated, large-scale environments (financial services, telecom, enterprise ERP). Specialist in LLM red teaming (OWASP Top 10 for LLMs, Microsoft PyRIT, Promptfoo), continuous evaluation of agents in production, and multi-agent architecture — combining senior software engineering (Python, Cloud, CI/CD) with rigorous security controls. Track record of 32 GenAI use cases evaluated and secured before reaching production.

Currently focused on transforming legacy automation into autonomous AI agents, while pursuing a postgraduate degree in Applied AI Engineering (UNIPDS) and an MBA in Machine Learning Engineering (FIAP).


Professional Focus & Impact

Autonomous Agents & Process Automation

  • Led the conversion of legacy RPA flows into autonomous AI agents combining LLMs, RAG, and computer vision, cutting execution time per process by 40% and achieving 95%+ task success rate in production enterprise environments.
  • Implemented security-by-design guardrails and validation from conception, blocking 8 vulnerabilities pre-production through CI/CD pipelines.

GenAI Evaluation & Observability

  • Structured continuous evaluation for 8+ GenAI use cases in production using golden sets, regression testing, and groundedness/hallucination/safety metrics (DeepEval, G-Eval, ROUGE/BLEU/METEOR), reducing quality incidents by 35%.
  • Monitored production agents on task success rate, tool-calling accuracy, latency, and cost per execution, cutting drift/degradation detection time by 50% via thresholds and alerting.
  • Combined automated evaluation (LLM-as-a-judge) with human review in a continuous loop, sustaining 90% adherence to defined quality criteria.

AI Security & Governance

  • Ran red teaming covering OWASP Top 10 for LLMs risks (prompt injection, jailbreak, data leakage, tool abuse) across 12 use cases, blocking 20+ vulnerabilities before production.
  • Automated security testing pipelines with Microsoft PyRIT, Promptfoo, and DeepTeam, cutting evaluation time per use case by 60%, producing evidence, reports, and playbooks aligned with NIST AI RMF.
  • Grew from sub-lead to squad lead within 8 months, coordinating a team of 3 and aligning technical risk with cross-functional stakeholders.

Full-Stack AI Solutions

  • Built a financial document entity-extraction platform from scratch (Python, Django, LangChain), processing 12,000 documents/month at 97% average accuracy, eliminating manual data entry.
  • Developed a multi-agent marketing automation platform (briefing, content, review, validation), reducing campaign creation cycle time by 45%.
  • Delivered full-stack improvements on high-traffic platforms serving 150,000+ users, optimizing SQL queries and reducing operational rework by 30%.

Selected Projects

  • Internal PDF RAG Knowledge Chat — Full-stack RAG pipeline (LangChain, ChromaDB, FastAPI, Docker) with cited answers and confidence scoring in under 120ms per query.
  • Risk Scoring ML Pipeline — Risk-scoring pipeline with feature engineering and supervised models achieving AUC ≈ 0.91 (scikit-learn, PyTorch, structured data validation).
  • Financial Document Entity Extraction — LLM-based structured entity extraction from financial documents (Django, LangChain, REST APIs).

Skills

AI Security & Governance Red Teaming Prompt Injection Jailbreak Guardrails PyRIT Promptfoo DeepTeam OWASP LLM Top 10 NIST AI RMF

LLMs & Generative AI LangChain LangGraph CrewAI RAG Multi--Agent Systems MCP Prompt Engineering

Evaluation & Observability DeepEval RAGAS G--Eval LangSmith OpenTelemetry Prometheus

Software Engineering Python TypeScript FastAPI Django Docker Kubernetes GitHub Actions Terraform

Cloud & Data AWS Azure PostgreSQL Redis Kafka pgvector

Machine Learning scikit-learn PyTorch TensorFlow SHAP


Education

  • Postgraduate, Applied AI Engineering — UNIPDS (May 2026 – May 2027, in progress) LLMs, RAG, MCP, autonomous agents, fine-tuning (LoRA/PEFT), AI security & governance.
  • MBA, Machine Learning Engineering — FIAP (Jan 2026 – Oct 2026, in progress) MLflow, DVC, Docker, Kubernetes, CI/CD, model monitoring, data drift, LLMOps.
  • B.Tech, Systems Analysis and Development — Estácio de Sá (Jul 2022 – Jan 2025, completed)

Activity

GitHub contribution grid snake animation

Contact

📫 Email: marcospaulomaio2607@gmail.com 💼 LinkedIn: linkedin.com/in/marcos-maio

Popular repositories Loading

  1. Beautiful-Form-Register Beautiful-Form-Register Public

    Projeto de site proprio - Finalizado

    TypeScript 4

  2. Animes-Fights Animes-Fights Public

    Projeto de site proprio - ainda não finalizado !

    JavaScript 1

  3. POKEDEX POKEDEX Public

    Projeto de site proprio - ainda não finalizado !

    JavaScript 1

  4. To-Do-List To-Do-List Public

    Own website project - completed - Implementation of a responsive to-do list.

    JavaScript 1

  5. Stone-Challenge Stone-Challenge Public

    Stone Challenge project made by mayself - https://github.com/stone-payments/template-desafio-web

    TypeScript 1

  6. Sunnyside-Agency Sunnyside-Agency Public

    Web Project made to conclude front-end mentor challenge

    CSS 1