Applied AI Engineer
OnebyZero · Bangalore, Karnataka, India · On-site
Posted Sep 28, 2026
Sign up free: we match you to jobs like this, tailor your application and fill the form. 2 free applications every day.
About OneByZero OneByZero is a Frontier Systems Integrator building agentic AI systems for leading banks, telcos, insurers, and retailers across Asia Pacific. We work on mission-critical problems for organizations operating the region's financial and digital infrastructure. Operating exclusively on AWS, we combine deep technical expertise with strong partnership programs that enable innovation, scale, and delivery excellence. We build with people who are technically serious, outcome-focused, and passionate about solving complex enterprise challenges. The Role We are seeking a Applied AI Engineer with 3–4 years of experience to help design and build production-grade GenAI systems. In this role, you will contribute architecture coverage across the team—reviewing system designs, identifying gaps, and guiding technical decisions at the solution level. You will work on end-to-end LLM system design, Retrieval-Augmented Generation (RAG) pipelines, and multi-agent architectures, with a strong focus on production readiness. Strong coding depth is non-negotiable. What You Will Do Design and contribute to end-to-end LLM system architecture for real-world enterprise use cases (from requirements to production). Pre-train, fine-tune LLMs and domain-specific models using techniques such as CPT, SFT, LoRA, and QLoRA for client-specific use cases. Design and run model evaluation pipelines to benchmark performance, accuracy,and cost across different fine-tuning approaches. Optimise models for latency, throughput, token efficiency, and inference cost in production environments. Work alongside agent orchestration and architecture teams to integrate fine-tuned models into multi-agent pipelines. Implement prompt versioning, rollback strategies, and model monitoring to ensure reliability post-deployment. Translate business requirements from client engagements into model adaptation strategies with clear success criteria. Contribute to internal knowledge sharing on fine-tuning best practices, tooling, and emerging techniques. Define enterprise integration patterns for GenAI systems (identity/access controls, auditability, data boundaries, governance, and compliance alignment). Improve production reliability: latency/throughput optimization, token efficiency, cost control, and robust failure handling. Collaborate with cross-functional stakeholders (engineering, data, product, client teams) to deliver high-impact solutions on tight timelines. Contribute hands-on code, perform code reviews, and raise the engineering bar through strong software fundamentals. Requirements What We Are Looking For 3–6 years of experience in ML engineering, LLMs, or model development roles. Hands-on experience with Continual pre-training (CPT), supervised fine-tuning (SFT), LoRA, or QLoRA on LLMs. Strong Python programming skills. Ability to write clean, testable, production-ready code. Experience running model evaluation and benchmarking pipelines in a structured way. Solid understanding of…