- Contributing to client projects that incorporate LLMs and generative AI, taking part in the development and implementation of different solutions.
- Intensive daily use of Claude Code alongside other AI-assisted development environments for analysis, prototyping and implementation.
- Development and integration of software components within client-facing teams, working mainly with Python, APIs and cloud services.
- Fine-tuning and evaluation of BERT models for multilingual classification, with automated model comparisons and experiment tracking in MLflow.
Marcos García Estévez
AI / GenAI Engineer | LLMs, RAG, AI Agents & Applied AI
Professional profile
Applied AI Engineer specializing in Large Language Model (LLM), Retrieval-Augmented Generation (RAG) and agentic systems. Experience on client projects involving LLMs and generative AI at Accenture, alongside self-built products and tools shipped to production and open source. Strong Python, backend and NLP/ML foundation, with regular use of Claude Code and other AI-assisted development environments.
Professional experience
- Systematically evaluated Spanish LLM responses for factual accuracy, relevance, instruction following, reasoning, inconsistencies and hallucinations.
- Reviewed training/evaluation data and model outputs, providing structured feedback to improve model quality and usefulness.
Selected projects
Production-oriented AI guide for Menorca. Work on the conversational layer: RAG over controlled sources, retrieval and reranking quality, evaluation, safety/abuse controls, feedback, latency/cost/quality observability and staged CI/CD releases.
Production deal platform with a public website and a 2.3K+ Telegram audience. End-to-end automation with AI-assisted processing, multi-channel publishing and lifecycle maintenance; Python/PostgreSQL backend, Docker operations, monitoring, backups and staged CI/CD with health checks and rollback.
Automatic permission-review plugin for coding agents: tool-free LLM reviewer, configurable policy, deterministic risk invariants, secret redaction, audit logging, read-only evidence enrichment and fail-safe escalation to human review.
Technical skills
Education & certifications
600-hour dual program · Final grade: 10/10 · Capstone: MIDAS
Languages: Spanish - Native · English - B1 (Cambridge).