Juan Garassino — AI Engineer

Juan Garassino is an AI engineer based in Berlin, Germany, specializing in multi-agent systems, large language model (LLM) applications, and generative AI. He designs and ships production AI systems and teaches machine learning, with a background that bridges architecture (Universidad de Buenos Aires) and applied AI/ML engineering.

What Juan does

Juan builds multi-agent orchestration systems (LangGraph, LlamaIndex, and custom hand-rolled agent loops), retrieval-augmented generation (RAG) pipelines, fine-tuned and quantized LLMs (LoRA/QLoRA, PEFT, RLHF with GRPO/PPO/DPO), and generative models (diffusion, GANs). He works across the full stack — FastAPI backends, Docker/Kubernetes, and GCP, AWS, and Azure — and has a research interest in LLM interpretability and AI red-team security.

Current roles

Core skills

Python, PyTorch, TensorFlow, Transformers, LangGraph, LangChain, LlamaIndex, Hugging Face, FastAPI, multi-agent systems, RAG, LLM fine-tuning, MLOps (MLflow, Weights & Biases), vector databases (ChromaDB, Pinecone, FAISS, pgvector), and cloud ML (Vertex AI, SageMaker).

Work with Juan

Links