Scientific Data Architect - Tarrytown, NYRequirementsWhat You Have DoneYou deeply understand the life science R&D data ecosystem - you've felt the pain of brittle, bespoke workflows with fractured data, and you've actively worked to solve this. In our experience, the candidates with this experience bring the following background:
- PhD with +4 years, Masters with +6 years, or Bachelors with +8 years of industry experience in life sciences with extensive domain knowledge in drug discovery (target ID through lead optimization), preclinical development, CMC (all drug modalities), product quality testing, or pharma manufacturing.
- Proven track record of defining, designing, prototyping, and implementing productized AI/ML-driven use cases in cloud environments
- Designed scalable and reusable data architecture
- Collaborated with cross-functional teams, including product managers, software engineers, and scientific stakeholders.
- Performed extensive exploratory data analysis and workflow optimization to enable scientific outcomes not previously possible.
- Engage diverse audiences, from scientists to executive stakeholders using your excellent communication and storytelling abilities.
- Advised scientists in a consulting capacity to further research, development, and quality testing outcomes.
- Augmented your technical, business, and communication work through agentic development and knowledge work
- Nice to have: Hybrid dry lab / wet lab experience
What You Will Do- You will be a critical team member in a unique partnership to industrialize Scientific AI. As such, you will engage directly with customers onsite a couple of days per week in the assigned geographic region, building strong relationships, deeply understanding their scientific data challenges and requirements, and accelerating solutions.
- Design and implement extensible, reusable data models that efficiently capture and organize scientific data for scientific use cases, ensuring scalability and future adaptability.
- Translate scientific data workflows into robust solutions leveraging the Tetra Data Platform.
- Own, scope, prototype, and implement solutions including:
- Data model design
- Python-based pipeline development.
- Lab software (e.g., ELN/LIMS) integration via APIs.
- Data visualization and app development
- Scientific agents
- Leverage agentic development tools like Claude Code and Codex to contribute to discovery, prototyping, and development
- Collaborate with Scientific Business Analysts (SBAs), customer scientists and applied AI engineers to develop and deploy models (ML, AI, mechanistic, statistical, hybrid) and agents
- Interface directly with scientific end users and technical stakeholders to rapidly drive solution development and adoption through regular demos and meetings
- Proactively communicate implementation progress and deliver demos to customer stakeholders.
- Collaborate with the product team to build and prioritize our roadmap by understanding customers' pain points within and outside Tetra Data Platform.
- Rapidly learn new technologies to develop and troubleshoot use cases
- Must be able to travel to client sites in local geographic areas
Benefits- 100% employer-paid benefits for all eligible employees and immediate family members
- Unlimited paid time off (PTO)
- 401K
- Company paid Life Insurance, LTD/STD
- A culture of continuous improvement where you can grow your career and get coaching
We are not currently providing visa sponsorship for this position.
The salary range for this position is $140,000 - $240,000. The salary range posted reflects our target baseline for this role. Final compensation is determined by a thorough evaluation of factors including the candidate's specific experience, localized market data, and internal team equity.