Abstract architectural void with light entering from the right

Hello, I'm

Davi Guides

Principal AI Architect

Turning artificial intelligence into tangible, simple, and scalable systems.

My LinkedIn profile My GitHub profile

Get To Know More

About Me

Bridging deep expertise with practical, business-driven solutions.

Abstract architectural void with light entering from the right
Experience icon

Solid Experience

Software Engineer

+22 Years

Domains icon

GenAI Focus

Agents & Assistants

RAG & Orchestration

Domains icon

Broad Domains

Backend & Frontend

Data & Platform

Extensive experience delivering high-performance systems, resilient platforms, creating real business value.


I specialize in GenAI systems that work in production, with reliability, security, and cost awareness.


Arrow icon

Explore My

Stack

GenAI systems specialist with deep ML foundations and over two decades of full-stack engineering.

AI Engineering

GenAI & Agentic Systems

Agno Icon

Agno

+2y

LangChain Icon

LangChain

+4y

CrewAI Icon

CrewAI

+2y

LangGraph Icon

LangGraph

+2y

Claude Agent SDK Icon

Claude Agent SDK

+1y

Claude Code Icon

Claude Code

+1y

Qdrant Icon

Qdrant

+2y

pgvector Icon

pgvector

+2y

Pinecone Icon

Pinecone

+1y

Chroma Icon

Chroma

+1y

Neo4j Icon

Neo4j

+1y

Context Engineering Icon

Context Engineering

+4y

RAG Optimization Icon

RAG Optimization

+4y

Voice Agents Icon

Voice Agents

+3y

LiteLLM Icon

LiteLLM Proxy

+2y

OpenAI SDK Icon

OpenAI SDK

+4y

Anthropic SDK Icon

Anthropic SDK

+2y

Google Gemini SDK Icon

Gemini SDK

+1y

LlamaIndex Icon

LlamaIndex

+2y

AWS Bedrock Icon

AWS Bedrock

+2y

vLLM Icon

vLLM

+1y

Ollama Icon

Ollama

+2y

Phoenix Icon

Arize Phoenix

+1y

LangFuse Icon

LangFuse

+2y

Prompt Engineering Icon

Prompt Engineering

+4y

Multi-Agent Design Icon

Multi-Agent Design

+2y

MCP Servers Icon

MCP Servers

+1y

Evaluation-Driven Development Icon

Evaluation-Driven Development

+1y

Classical ML & NLP

spaCy Icon

spaCy

+6y

Transformers Icon

Transformers

+4y

Sentence Transformers Icon

Sentence Transformers

+4y

BERT Icon

BERT

+5y

Word2Vec Icon

Word2Vec

+7y

BM25 and Lexical Retrieval Icon

BM25 & Lexical Retrieval

+6y

Hybrid Retrieval Icon

Hybrid Retrieval (RRF)

+2y

Text Classification Icon

Text Classification

+7y

Intent Classification Icon

Intent Classification

+6y

Whisper Icon

Whisper

+3y

HuggingFace TTS Icon

Kokoro TTS

+3y

NLTK Icon

NLTK

+9y

scikit-learn Icon

scikit-learn

+9y

PyTorch Icon

PyTorch

+1y

NLI and Cross-Encoders Icon

NLI & Cross-Encoders

+2y

IR Evaluation Icon

IR Evaluation (nDCG, MRR)

+2y

Linear Regression Icon

Linear Regression

+5y

Named Entity Recognition Icon

Named Entity Recognition

+6y

Backend Engineer

Python Icon

Python

+18y

SQL Icon

SQL

+22y

FastAPI Icon

FastAPI

+9y

Flask Icon

Flask

+9y

Pytest Icon

Pytest

+10y

Django Icon

Django

+18y

RabbitMQ Icon

RabbitMQ

+18y

Apache Kafka Icon

Apache Kafka

+9y

GraphQL Icon

GraphQL

+6y

Rust Icon

Rust

+1y

Redis Icon

Redis

+6y

SQLAlchemy Icon

SQLAlchemy

+9y

Docker Icon

Docker

+14y

Amazon EC2 Icon

Amazon EC2

+14y

Amazon RDS Icon

Amazon RDS

+14y

Amazon SQS Icon

Amazon SQS

+6y

Amazon SES Icon

Amazon SES

+4y

Platform Engineer

Linux Icon

Linux

+22y

ShellScript Icon

ShellScript

+22y

AWS Icon

AWS

+14y

GitHub Actions Icon

GitHub Actions

+6y

Terraform Icon

Terraform

+6y

Kubernetes Icon

Kubernetes

+6y

Grafana Icon

Grafana

+7y

Prometheus Icon

Prometheus

+3y

Vault Icon

Vault

+3y

AWS Lambda Icon

AWS Lambda

+6y

Ansible Icon

Ansible

+1y

AWS Route53 Icon

AWS Route53

+3y

Google Cloud Icon

Google Cloud

+4y

CloudWatch Icon

CloudWatch

+14y

Data Engineer

MySQL Icon

MySQL

+22y

PostgreSQL Icon

PostgreSQL

+19y

MariaDB Icon

MariaDB

+8y

MongoDB Icon

MongoDB

+4y

DuckDB Icon

DuckDB

+2y

Elastic Stack Icon

Elastic Stack

+6y

Pandas Icon

Pandas

+4y

AWS Redshift Icon

AWS Redshift

+3y

Polars Icon

Polars

+4y

scikit-learn Icon

scikit-learn

+9y

SciPy Icon

SciPy

+4y

Numpy Icon

Numpy

+9y

SpaCy Icon

SpaCy

+6y

Apache Parquet Icon

Parquet

+3y

Frontend Engineer

JavaScript Icon

JavaScript

+22y

TypeScript Icon

TypeScript

+7y

React Icon

React

+4y

Next.js Icon

Next.js

+1y

Tauri Icon

Tauri

+1y

Vue.js Icon

Vue.js

+4y

HTML5 Icon

HTML

+22y

CSS Icon

CSS

+22y

Tailwind CSS Icon

Tailwind CSS

+2y

Gradio Icon

Gradio

+1y

Markdown Icon

Markdown

+10y

Plotly Icon

Plotly

+3y

Streamlit Icon

Streamlit

+3y

Looker Icon

Looker

+3y

GitHub Pages Icon

GitHub Pages

+3y

Arrow icon

Leverage My

AI Specs

Open specifications and plugins for AI-powered development workflows

Arché

Arché

Eight behavioral principles for Claude Code: research before action, corrections that stick, single source of truth, respect for the user's mode, aiming at the high-value path, done meaning the requirement, autonomous execution, and maximum signal. A dogmatic framework that turns guidelines into gates.

Semantic Docstrings

Semantic Docstrings

Semantic documentation standards for Python projects. Agents and commands for validating and generating meaningful, context-rich docstrings.

YMD Spec

YMD Spec

YMD (YAML Markdown) specification for structured prompt composition. Modular format for building complex, maintainable AI prompts and instructions.

Arrow icon

Read My

Articles

A look into the ideas, experiments, and lessons behind the systems I build.

Above the Model

A minimal Zen-style illustration of a bold line climbing steadily while a ribbon undulates below, running straighter where it passes through an open containment frame.

Capability Compounds, Discipline Doesn't

Why Smarter Models Still Need Guardrails

July 25, 2026

Every release makes models more capable. It does not reliably make them better behaved. Those are different axes, and the industry's roadmap quietly assumes they are one.


A minimal Zen-style illustration of a large enso circle holding a small inner circle, with thin conduits carrying its output to documents outside.

The Model Is Not the System

Build the Structure, Don't Wait for the Model

June 27, 2026

One camp waits for the next model to make agent problems disappear. The other engineers the structure around the model, including the instrument that feeds the system data about itself, so it can evolve on evidence instead of faith.


A minimal Zen-style illustration contrasting a single sketchy checkmark with a ring of many loops, one incomplete.

Loops Need Proof, Not Vibes

Why "It Passed" Isn't Evidence

May 30, 2026

The industry found its word for how agents improve, the loop. What the conversation skips is the step that makes a loop trustworthy when the exit criterion is semantic and the agent underneath is nondeterministic.


A minimal Zen-style illustration of tangled demand passing through three torii gates and emerging as ordered parallel lines feeding clean briefs.

Intent Engineering

Working Above the IDE

April 25, 2026

Code became a byproduct. The scarce input is well-formed intent. Notes on a year of building and daily-driving a post-IDE workstation.


Field Notes

A minimal Zen-style illustration of one operator coordinating thirteen parallel flowing streams, one broken, converging into a single check.

Running Thirteen AI Agents in Parallel

An Operations Post-Mortem

March 28, 2026

One coordinating session drove thirteen parallel agent work units through analysis, implementation, measurement and benchmarking. Here is everything that broke, and the operating rules that came out of it.


Foundations

A minimalist Zen-style illustration representing the intersection of global culture, language, and machine intelligence in the age of LLMs.

When Language Sounds Off

July 09, 2025

Large language models often fail to sound natural in Portuguese and Spanish, not because of lack of fluency, but due to deep cultural, pragmatic, and contextual mismatches. This article explores how context engineering, cultural awareness, and linguistic insight can help bridge that gap and produce outputs that truly resonate with real-world users.


A minimal Zen-style diagram illustrating LLMs translating structured data into natural language.

Why LLMs Prefer Natural Language

May 10, 2025

Why Transformer-based language models respond more fluently and reliably to narrative input than structured formats, exploring the architectural bias toward sequential context, the cost of semantic reconstruction from JSON or YAML, and how to write prompts that align with how LLMs actually reason.


A minimal, horizontal Zen-style diagram illustrating self-attention at the heart of Transformer architecture.

Transformer Architecture for Humans

May 03, 2025

An intuitive, text-first introduction to Transformer models written for technical leaders, engineers, and curious minds. This article explains how Transformers work using systems thinking and practical language, without equations or jargon, making the core mechanics of modern LLMs accessible through context and clarity.


Security & DevSecOps

Zero Trust Local Manifesto visual concept.

Zero Trust Local Env Manifesto

April 25, 2025

A philosophy for CLI development that assumes breach, encrypts everything, and trusts nothing local, ensuring secrets never persist in plaintext on disk.


Diagram showing two-layer token security with 1Password icon, encrypted vault, local key, river lines, arrows, and labels, Unlock, Encrypted Data, Decrypt, Local Key, and the title.

Building a Secure Token CLI in Python

April 24, 2025

Learn how to build a secure and reliable Python CLI for token management by combining encrypted storage and robust local key protection.


Mind map diagram showing SOC 1, SOC 2, and SOC 3 audits, Type 1 and Type 2 reports, and standards SSAE 18 and ISAE 3402.

Understanding SOC Audits in Cybersecurity

April 24, 2025

A practical guide to SOC 1, SOC 2, SOC 3 audits, report types, and standards, designed to enhance security evaluations and vendor trust.


Arrow icon

Browse My

Open Source

See how I craft systems and solutions.

the-saurus - AI literature review pipeline

The Saurus

the-saurus is a literature review pipeline: upload scientific PDFs, AI agents extract themes and claims in parallel, deduplicate semantically across papers, and synthesize a citation-backed review you can then chat with via an embedded conversational assistant. Evaluated with RAGAS and DeepEval; Restate for durable workflow execution; Terraform and Helm for AWS/Kubernetes deployment.

claude-agent-toolkit - Rust port of the Claude Agent SDK

Claude Agent Toolkit

claude-agent-toolkit is an idiomatic Rust port of the Claude Agent SDK: protocol, subprocess transport, MCP server, session management, and client, built under clippy::pedantic with unsafe code forbidden.

KeySentinel - Secure Token Management

KeySentinel Python Library

KeySentinel is a lightweight, secure token encryption library and CLI tool for managing sensitive credentials with strong Zero Trust principles. It features two-layer encryption, memory-only decryption, and enforces safe local environments with automatic cleanup and no plaintext leakage.

Care Gateway – Full-stack healthtech simulation

Care Gateway - All-in-One Healthtech Simulation

Care Gateway is a full-stack backend simulation for healthcare claim processing, combining REST APIs, gRPC, Kafka, and ETL pipelines with PySpark. Designed for clarity and modularity, it showcases real-world skills in Python, data engineering, and cloud-ready system design.

Arrow icon

Get in Touch

Contact Me

I am open to discussing new projects, consulting opportunities, technical leadership roles, or collaborations.


Feel free to connect with me on LinkedIn!