Virtusa Corporation

Architect

Virtusa Corporation$110K — $130K *
Enterprise Technology
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • 4+ years in data engineering, conversation design operations, applied NLP data work, or knowledge-pipeline engineering.
  • Working knowledge of conversational platforms and their data consumption needs.
  • Strong judgment on synthetic-data quality and retrieval safety.
  • Preferred: Experience with LLM-assisted synthetic data generation in production.
  • Familiarity with BigQuery, S3, and document storage for knowledge management.
  • Preferred: Experience in multilingual data generation or evaluation.

Responsibilities

  • Analyze customer domains to identify intents, entities, knowledge topics, and edge cases.
  • Generate synthetic conversation transcripts for voice and chat applications.
  • Create supporting content such as agent profiles, knowledge-base articles, and FAQs.
  • Schedule and document updates to customer knowledge and index refresh processes.
  • Validate the realism and diversity of synthetic data while ensuring privacy compliance.
  • Maintain reusable data generators and ingestion processes tailored for each customer.
  • Collaborate with Conversational Platform Specialists and DevOps for effective data utilization and automation.

Benefits

  • Flexible work arrangements to accommodate personal and professional needs.
  • Opportunity to work on cutting-edge conversational AI technologies.
  • Access to professional development resources and training programs.
  • Collaborative work culture that emphasizes teamwork and innovation.
  • Health and wellness programs to support employees' well-being.
Full Job Description
Role summary
Create high-quality, customer-specific synthetic data and own RAG / knowledge pipelines so each deployment of
CCAI, voice, and chat can be configured, grounded, demonstrated, and validated without using real customer PII.
You design generation and ingestion pipelines and load data into the correct GCP and AWS services.
What success looks like
Each customer engagement has a documented synthetic dataset covering the channels in scope • Each in-scope customer has a working RAG / knowledge pipeline: corpus prepared, indexed, retrievable,
and evaluated.
Data and retrieval quality are good enough for configuration, evaluation, and stakeholder demos, and
safe enough for isolation and compliance expectations.
Generation and indexing are parameterized and repeatable, not a one-off manual copy-paste per
customer.
Key responsibilities
Analyze each customer's domain: intents, entities, knowledge topics, document types, languages, tone,
and edge cases.
Generate synthetic conversation transcripts for voice and chat, plus CCAI training/evaluation dialogues.
Generate supporting content: customer/agent profiles, knowledge-base articles, FAQs, and structured
entity values.
Schedule and document index refresh processes when customer knowledge changes.
Use appropriate techniques while
documenting parameters and limitations.
Validate realism, coverage, diversity, and absence of residual real-world PII in synthetic data and source
corpora.
Maintain reusable generators, ingestion jobs, and quality checklists that can be parameterized per
customer.
Partner with the Conversational Platform Specialist so loaded data and indexes actually drive the
deployed experience.
Partner with DevOps so pipeline jobs, stores, and secrets are automated and isolated per customer.
Required qualifications
4+ years in data engineering, conversation design operations, applied NLP data work, or knowledge-
pipeline engineering.
Working knowledge of how conversational platforms consume training, FAQ, transcript, and retrieval-
grounded knowledge data.
Strong judgment on synthetic-data quality, retrieval quality, and privacy safety.
Preferred qualifications
LLM-assisted synthetic data generation in a production or implementation setting.
Familiarity with BigQuery, S3, and document stores used as knowledge sources.
Multilingual data generation or evaluation experience.

About Virtusa Corporation

Virtusa Corporation is a global provider of digital business strategy, digital engineering, and information technology (IT) services and solutions. The company helps its clients transform their businesses through innovative technology solutions that enable them to improve customer engagement, increase operational efficiency, and drive revenue growth. Virtusa Corporation was founded in 1996 and is headquartered in Southborough, Massachusetts. The company has over 20,000 employees and operates in over 30 countries around the world. Virtusa Corporation is committed to sustainability and has received numerous awards for its environmental and social responsibility initiatives.
Learn more about Virtusa Corporation
Size
22,883 employees
Market Cap
$1.5 billion
Industry
Net Income
$44.6 million
Founded
1996
5 Year Trend
+22.3%
Revenue
$1.2 billion
NASDAQ

Similar Jobs

More Jobs at Virtusa Corporation

  • Virtusa Corporation
    Systems Architect
    $120K — $145K *
    Las Vegas, NV 89110 (Clark County)
    Information Technology
    In-Person
  • Virtusa Corporation
    Lead Software Engineer
    $100K — $120K *
    Toronto, ON M3C 0E3
    Information Technology
    In-Person
  • Virtusa Corporation
    Senior Software Engineer
    $100K — $120K *
    Toronto, ON M3C 0E3
    Finance & Insurance
    In-Person
  • Virtusa Corporation
    Lead Software Engineer
    $130K — $155K *
    New York, NY 10025 (New York County)
    Information Technology
    In-Person
  • Virtusa Corporation
    Architect-Java
    $110K — $130K *
    Tampa, FL 33647 (Hillsborough County)
    Information Technology
    In-Person

More Enterprise Technology Jobs

Find similar Architect jobs: