JobRaahGet matched free

Jobs

Agentic AI Engineer - RAG Architecture - LLM Systems

Aspenview Technology Partners · Buenos Aires, Autonomous City of Buenos Aires, Argentina, Bogotá, Bogota D.C., Colombia · On-site

Posted Aug 4, 2026

Apply with JobRaah

Sign up free: we match you to jobs like this, tailor your application and fill the form. 2 free applications every day.

Build the Future with AspenView Technology Partners At AspenView, we are passionate about transforming the way organizations approach technology. We specialize in creating high-performing, nearshore IT teams to help North American clients innovate faster and more efficiently. As we continue to grow, we’re looking for exceptional people to join our team and help drive impactful change across industries. Why Join AspenView? At AspenView, we’re more than a nearshore IT partner—we’re a people-first, purpose-driven company that believes great culture drives great outcomes. We’re passionate about connecting talent and technology to deliver measurable value for clients—and meaningful career paths for our people. Here’s what you can expect: Competitive base Comprehensive benefits and wellness support Flexible work model: hybrid, remote, or in-office Real growth opportunities and leadership visibility Inclusive, respectful culture that blends U.S. innovation with Colombian heart A company that listens, invests in you, and celebrates wins together About the Role In this role, you will lead the architecture and implementation of multi-step agentic workflows capable of reasoning, tool execution, and synthesizing complex clinical and operational data. You will engineer sophisticated RAG pipelines that bridge unstructured medical knowledge with enterprise EHR systems, establishing strict safety guardrails, hallucination controls, and HIPAA-compliant AI pipelines. What You Will Do Agentic System Engineering: Architect and deploy autonomous agentic AI workflows and multi-agent systems (using frameworks like LangGraph, LlamaIndex, AutoGen, or CrewAI) capable of multi-step reasoning, dynamic tool usage, and clinical task execution. Advanced RAG Architecture: Design, optimize, and scale production RAG pipelines utilizing hybrid search (dense + sparse retrieval), re-ranking, query transformation, context compression, and semantic routing over multi-modal clinical and research datasets. LLM Fine-Tuning & Evaluation: Evaluate, fine-tune, and benchmark foundational models (e.g., Llama 3, Claude, GPT-4, Med-PaLM) for specialized healthcare tasks, establishing rigorous evaluation frameworks (e.g., Ragas, TruLens) for accuracy, groundness, and latency. Vector Database Infrastructure: Manage and optimize vector stores (Pinecone, Qdrant, Milvus, Weaviate, or pgvector) for high-concurrency, low-latency similarity searches and knowledge retrieval. Guardrails & HIPAA Compliance: Implement enterprise safety controls, input/output sanitization, and hallucination guardrails (e.g., NeMo Guardrails, Llama Guard) to ensure 100% HIPAA compliance and zero unauthorized PHI leakage. Cross-Functional Innovation: Collaborate closely with US-based clinical researchers, cloud architects, and product leads to translate complex healthcare needs into autonomous AI capabilities. What You Bring Experience 4+ years in software production…