3-5 years of hands-on software engineering experience with end-to-end project ownership
Proficiency in at least one backend programming language (Python, Go, Java, C++, etc.)
Strong experience in creating production systems focusing on reliability and scalability
Interest in AI applications, specifically LLMs and workflow automation
End-to-end ownership mindset with effective collaboration skills across teams
Responsibilities
Engineer and deliver vertical Agent workflows for specific use cases
Design and implement production-ready service architecture
Contribute to testing and quality benchmarking of Agent systems
Collaborate with cross-functional teams to meet delivery milestones
Adopt AI tools to enhance engineering efficiency
Benefits
Opportunity to join a founding team with significant impact
Top-tier healthcare for employees and dependents
401(k) plan with company matching
Unlimited paid time off and 13 paid holidays
12 weeks of paid new parent leave
Hybrid work model with a minimum of 3 in-office days per week
Top-of-the-line laptops and workstation setup
Full Job Description
What You Will Do
Design, build, and ship the batch transcription and streaming ASR API services on Plaud's developer platform - owning the full stack from inference API design and model service integration through authentication, rate limiting, versioning, and developer documentation.
Deploy and operate the STT services across all 4 global regions; establish on-call rotation, monitoring, alerting, and SLO governance from launch; lead incident response and post-mortems.
Architect for high-concurrency, low-latency MaaS-style workloads - streaming/real-time protocols (WebSocket/gRPC), async job pipelines for batch transcription, autoscaling, and cost/latency optimization - benchmarking against SOTA inference API providers (e.g., Deepgram, AssemblyAI, Google STT).
Serve as the primary technical point of contact for B2B customers on API capabilities; scope, commit to, and deliver new feature requests on defined timelines.
Collaborate directly with the algorithm team on ASR model integration and with product on the developer platform roadmap; use AI tooling to accelerate development, documentation, and observability workflows.
Minimum Qualifications
5+ years of backend engineering experience with strong Go (Golang) proficiency; shipped and operated production services in Go.
Hands-on production experience building and operating AI/ML-serving backend services - inference API design, model service integration, on-call and incident response - not just internal tooling or prototypes.
Solid architecture experience with high-concurrency, low-latency platform services: multi-region deployment, streaming/real-time protocols (WebSocket and/or gRPC), async/queue-based job pipelines, and autoscaling.
Experience designing and shipping public-facing APIs or SDKs - authentication, rate limiting, versioning, and developer-facing documentation.
Professional English proficiency (written and spoken) for developer documentation and direct B2B customer engagement.
What We Offer
Founding Team: Opportunity to join the founding team of this new initiative, with meaningful ownership and impact on a fast-growing startup.
Competitive Compensation: $159K-$219K base salary+performance bonus+Equity.
Comprehensive Benefits: Top-tier healthcare for employees and dependents, including dental and vision, and a generous employer subsidy.
Retirement Planning: 401(k) plan for full time employees with company matching.
Paid Time Off: Unlimited PTO, plus 13 paid holidays.
New Parent Leave: 12 weeks of paid time off to spend time with your new family, regardless of gender.
Hybrid Office: Minimum of 3x in office per week.
Gear: New hires are equipped with their choice of new top-of-the-line laptops and workstation setups.
Perks: Best office equipment. Annual offsites. Free office drinks and snacks.