Req ID: 392357
We are currently seeking a RAG Arcitect/Engineer to join our team in Dallas, Texas (US-TX), United States (US).
As a RAG Architect / Engineer, you will be responsible for overseeing the design and implementation of enterprise-grade GenAI RAG solutions that solve complex business problems. You will work closely with product managers, software engineers, data scientists, and key stakeholders to conceptualize, prototype, test, and release robust GenAI solutions. In this dual role, you will act as both the technical authority and a core hands-on contributor across all aspects of our GenAI infrastructure, services, and structural vector databases.
Day-to-Day Job Duties:- Design and optimize the foundational structure for vector databases to seamlessly accommodate LLM models across multiple enterprise GenAI tools.
- Architect, develop, and scale the infrastructure for GenAI systems, including data storage, processing, and high-performance retrieval mechanisms.
- Lead the technical decision-making process regarding architectural choices with external GenAI service providers (e.g., Microsoft Azure, AWS) to ensure solutions are scalable, maintainable, and cost-effective.
- Collaborate directly with Business Partners, product owners, and software engineers to translate complex business requirements into enterprise-level GenAI solutions.
- Drive model optimization and fine-tuning strategies for large language models to maximize performance, lower latency, optimize memory utilization, and increase throughput.
- Present and demo architectural designs and analysis results clearly to non-technical audiences, breaking down complex AI/ML concepts into actionable business insights.
- Identify use-case-specific data sources and design robust methods for data collection, integration, cleansing, and transformation in collaboration with data engineers.
- Partner with software engineering teams to integrate GenAI components into existing platforms, ensuring seamless communication and compatibility across software modules.
- Enforce ethical AI practices (fairness, transparency, and accountability) by actively identifying and mitigating biases or risks prior to deployment.
- Conduct proactive research and POC experimentation on the latest advancements in GenAI technologies, frameworks, and algorithms to continuously enhance Enterprise AI capabilities.
- Evaluate and benchmark emerging GenAI models while providing technical guidance, code reviews, and mentorship to development team members.
- Align and support the overall organization's GenAI strategy and technology adoption roadmap.
Basic Qualifications:- Minimum 5+ years of professional, practical experience building and deploying enterprise AI applications. Must have an advanced understanding of LLMs, NLP, and advanced ML algorithms, alongside extensive hands-on experience designing RAG architectures and AI Agents.
- Minimum 3+ years of dedicated experience utilizing pre-trained language models (e.g., GPT, BERT, RoBERTa) via APIs or frameworks like Hugging Face for downstream NLP production tasks.
- Minimum 3+ years of professional experience utilizing core NLP libraries (NLTK, SpaCy, Gensim, BERT, SBERT models) and modern orchestration frameworks (LangChain, LangGraph, or EmbedChain) to build and deploy GenAI applications.
- Minimum 5+ years of professional experience in Python (or similar enterprise programming languages), paired with excellent problem-solving skills and the ability to drive results in a fast-paced, collaborative team environment.
Degree: Bachelor's degree in Computer Science, Data Science, or equivalent practical work experience.
Nice-to-Have Qualifications:- Proficiency in designing and deploying GenAI workloads using Azure or AWS Cloud Platforms.
Whenever possible, we hire locally to NTT DATA offices or client sites. This ensures we can provide timely and effective support tailored to each client's needs. While many positions offer remote or hybrid work options, these arrangements are subject to change based on client requirements. For employees near an NTT DATA office or client site, in-office attendance may be required for meetings or events, depending on business needs. At NTT DATA, we are committed to staying flexible and meeting the evolving needs of both our clients and employees. NTT DATA recruiters will never ask for payment or banking information and will only use @nttdata.com, @nttdatafed.com and @talent.nttdataservices.com email addresses. If you are requested to provide payment or disclose banking information, please submit a contact us form, https://us.nttdata.com/en/contact-us.
NTT DATA endeavors to make https://us.nttdata.com accessible to any and all users. If you would like to contact us regarding the accessibility of our website or need assistance completing the application process, please contact us at https://us.nttdata.com/en/contact-us.