The work behind the papers
Each case study below walks through the problem, the architecture, the evaluation and the results. For the complete list of publications with abstracts and BibTeX, see publications.
Featured projects
-
IJCNN 2026 First author
RG-DermNet
Residual gated attention for multimodal skin lesion classification
A fusion block that lets clinical metadata modulate the visual features without ever cutting off the image pathway — so the model keeps working when the anamnesis is incomplete. AUC 0.95 ± 0.01 on PAD-UFES-20.
Read the case study -
Scientific Reports Co-author
PRISM
Probabilistic reasoning for interpretable diagnosis
A Bayesian framework that lets a clinician watch the diagnosis shift as each piece of evidence arrives, with a stepwise calibration protocol that keeps confidence honest. Peak balanced accuracy 77.2 ± 3.6 %.
Read the case study -
IEEE JBHI Co-author
MetaBlock-SE
Sentence embeddings for missing clinical metadata
Replacing MetaBlock's sparse categorical encoding with a sentence embedding, so the model captures how clinical attributes relate — and survives when some are absent. Better than the original in every scenario tested.
Read the case study -
ICECCME 2025 First author
SRGAN-upscaled YOLO
Super-resolution as a training strategy for aerial detection
Aerial scenes mix very large and very small objects at uneven image quality. Upscaling the weakest images with an SRGAN during training raises detection performance across ten YOLO architectures on DOTA v1.5.
Read the case study
Ongoing work
Research still in progress, with the code public while it runs. The results are not settled yet — treat these as work you can read, not as conclusions.
-
MSc project · PPGI/Ufes
LLM as a NAS controller
Neural architecture search for multimodal skin lesion classification, driven by a local LLM served through Ollama. At each step the model receives the search space and the history of what has already been tried, proposes a new architecture configuration as JSON, and gets the resulting balanced accuracy back as the reward that shapes its next proposal.
llm-as-nas-controller -
Paper in preparation
NAS optimization for a multimodal VLM
The experimental code behind Neural Architecture Search optimization for a multimodal model for skin lesion recognition — a case study: searching over vision-language configurations on PAD-UFES-20 under patient-wise 5-fold cross-validation.
NAS-optimization-for-VLM
Code on GitHub
A selection from my repositories — the systems and experiments I keep public, grouped by what they do.
Search & architecture optimisation
-
NAS-ML-optimization
Architecture search with Microsoft NNI on CIFAR — the groundwork for the LLM-driven controller.
Python -
fine-tune-tokenizer
Adapting a tokenizer to a domain vocabulary before fine-tuning the model on top of it.
Python -
simple-image-classifier-pytorch
A minimal, readable PyTorch training loop — the baseline I start experiments from.
Python
LLM services & document intelligence
-
AUTOMATIC-OCR-SERVICE
A Docker Compose OCR service that watches an input folder and emits the extracted content as JSON.
Python -
RAG-DOCS-with-PGVECTOR
Retrieval-augmented generation over documents, with embeddings stored in Postgres via pgvector.
Python -
RESUME-INTELLIGENTE-ANALYZER
Job recommendation by RAG: sentence-transformer embeddings indexed in FAISS, answers generated with Ollama.
Jupyter -
NER
End-to-end named entity recognition: dataset generation, training and inference scripts.
Python -
agent-ai-with-langchain
A LangChain agent running against a self-hosted Ollama model, configured entirely through env vars.
Python -
ia-transc2form-with-ollama
Turning a transcription into a filled form, with a local Ollama server exposed to the network.
JavaScript
Fine-tuning small models
-
Fine-tunning-Gemma3-1b-unsloth
Fine-tuning Gemma 3 1B on SQuAD with Unsloth, then exporting the result for vLLM inference.
Jupyter -
vlm-fine-tunning
Fine-tuning Qwen2.5-VL — the vision-language side of the same question.
Python -
portuguese-summarization
Summarisation for Brazilian Portuguese text.
Python
Computer vision systems
-
Suspicious-Abandonned-Bag-Detector
Abandoned-baggage detection from a webcam or RTSP stream, built on YOLOv8 with a configurable threshold.
Python -
object-detect-and-transfer-by-gRPC
Detection on one machine, frames streamed to another over gRPC and decoded there — a distributed inference setup.
Python -
telegram-chatbot-and-object-description
A Telegram bot whose
JavaScript/describeImagecommand runs a local vision-language model over the photo you send it. -
clothes-detection
Clothing-type detection trained on a public dataset, wrapped in an inference API.
Python -
Crop-faces-using-MTCNN
Face detection and cropping with MTCNN, as a preprocessing step for downstream models.
Python -
3W-dataset-classifier
Time-series classification on Petrobras' 3W dataset of undesirable events in offshore oil wells.
Jupyter
Other work
Published research without a case study page yet — the abstracts and BibTeX are on the publications page.
-
Verbose metadata descriptions generated by LLMs
LLMs as semantic transducers: structured clinical attributes rewritten as anamnesis-like text, encoded with SBERT and fused through MetaBlock-SE. SBCAS 2026
-
LiwTERM-r
A lightweight transformer for multimodal skin lesion detection, robust to incomplete clinical input and to multiple images of the same lesion. Journal of the Brazilian Computer Society, 2026
-
Virtual reality diagnostic room for lower limb amputees
A multi-camera intelligent space feeding a VR environment where physiotherapists inspect gait parameters, with a neural classifier averaging above 91 % across evaluation metrics. Artificial Intelligence in Medicine, 2023
-
Gait recognition with the StarRGB technique
Condensing the spatial-temporal information of a gait sequence into a single image so a CNN can classify movements in remote physiotherapy. ICECCME 2021
-
Anomaly detection in civil structures
A benchmark of machine learning models for detecting anomalies in structural monitoring data. ICMLA 2023
-
Books on systems analysis, production planning and inventory
Three teaching books published with Atena Editora, used as course support material in technical and undergraduate programmes. Atena Editora, 2023–2024
Every paper, with abstracts and BibTeX
Fifteen entries across multimodal learning, medical AI, computer vision and embedded systems.