Agentic Systems Engineer

Kepler Group

$150K — $180K *
Information Technology
5 - 7 years of experience
Job Overview by Ladders

Qualifications

  • 7+ years of software engineering experience shipping production systems at scale
  • Proficient in backend technologies like Python or Node.js, distributed systems, PostgreSQL, Redis, and AWS
  • Strong architectural experience designing scalable systems for complex workflows
  • Experience in AI/ML systems and working with LLMs or ML infrastructure
  • Familiarity with large datasets, ETL pipelines, and semantic systems is a plus
  • Skillful in Git workflows, CI/CD practices, automated testing, and observability
  • Effective communicator who clearly articulates technical trade-offs

Responsibilities

  • Build distributed systems that reliably coordinate and execute a large number of agents
  • Create evaluation frameworks to monitor agent performance and catch regressions
  • Optimize agent performance through context compression and prompt optimization
  • Develop ontology and provenance systems to map concepts and trace outputs to sources
  • Integrate language models into production for intelligent research workflows
  • Take ownership of systems from design through to production, ensuring reliability and performance
  • Implement comprehensive testing and monitoring for newly developed systems

Benefits

  • Comprehensive medical, dental, vision, and 401k for employees and dependents
  • Automatic coverage for basic life, AD&D, and disability insurance
  • Daily lunch provided in-office
  • Development environment budget for necessary tools and setups
  • Unlimited PTO policy
  • Dedicated funds for technical tools and infrastructure without questions
  • Learning budget for attending conferences and courses to enhance skills
Full Job Description
The Role
What You'll Own

You'll build the agentic infrastructure that powers Kepler's AI research platform. You'll work on the foundational systems that make autonomous AI agents reliable at scale: distributed execution frameworks that run thousands of agents in parallel, evaluation systems that ensure agent quality, context management that maximizes agent performance, and the ontology and provenance systems that let us trace every number back to its source.

This role is for engineers who want to work at the frontier of AI systems, building the infrastructure that makes agents trustworthy for enterprise-critical decisions. This is our most senior individual contributor role: you build the runtime our agents execute on, the layer the models, workflows, and product all depend on.

Within your first 90 days, you will:
  • Own a production agent system end-to-end, from architecture to deployment
  • Build and deploy infrastructure that powers real financial research workflows
  • See your code enable agents to conduct research at top financial institutions


What You'll Do
  • Build agent execution infrastructure: Distributed systems that orchestrate and run massive numbers of agents in parallel with reliability, retry logic, and graceful degradation.
  • Build evaluation systems: Frameworks that measure agent quality, catch regressions, and ensure agents perform reliably across diverse research tasks.
  • Optimize agent performance: Context compression, prompt optimization, model routing, and latency reduction. Make agents faster and smarter.
  • Build ontology and provenance systems: The semantic layer that maps concepts to precise definitions and traces every output back to authoritative sources. This is what makes our platform trustworthy.
  • Integrate AI into production: Language models powering intelligent research workflows with robust error handling, fallback mechanisms, and cost optimization.
  • Own systems end-to-end: Design to production. Services, database optimization, deployment, monitoring.
  • Ship with production excellence: Comprehensive testing, monitoring, deployment pipelines. You own reliability for what you build.


Who You Are
  • 7+ years building production software. No upper limit, comp scales with experience.
  • Backend: deep distributed-systems experience. Our backend is Rust, but we don't require Rust experience. We believe strong engineering fundamentals and experience in other languages is what matters.
  • Architecture: Experience designing systems that scale and handle complex workflows
  • AI/ML systems: Experience building with LLMs, agent frameworks, or ML infrastructure
  • Data: Large datasets, ETL pipelines, knowledge graphs or semantic systems a plus
  • Practices: Git workflows, CI/CD, automated testing, observability
  • Strong communicator who can discuss technical trade-offs clearly
  • Curious about the frontier of AI agents and eager to push what's possible
  • Thrives in fast-paced environments with high ownership

Don't check every box? Apply anyway. We prioritize problem-solving ability, systems thinking, and drive to build transformative agentic infrastructure.

Our Technical Stack
  • Backend: Rust - agent orchestration, data extraction, computation pipelines.
  • Frontend: TypeScript, React - the analyst workspace and verification interfaces.
  • Data: PostgreSQL, plus direct integrations with official data sources.
  • Infra: AWS.
  • AI: Model-agnostic by design. We currently use Claude and GPT. The model is the replaceable part.


Mentorship & Growth

You'll be directly mentored by engineers who built Palantir's core systems. Expect:
  • Weekly 1:1s with senior engineers who've built systems at Palantir and Meta scale.
  • Deep architectural reviews and guidance on system design.
  • Clear growth path toward technical leadership and system ownership.
  • Learn by building production systems that power real financial research.


Working at Kepler
Our Benefits
  • 100% covered top-of-the-line medical, dental, and vision insurance for employees and their families. HSA maxed by the company to the IRS limit.
  • Automatic coverage for life, AD&D, and disability insurance.
  • Daily lunch in office.
  • Unlimited PTO policy.
  • Development environment budget - latest MacBook Pro, multiple monitors, ergonomic setup, and any development tools you need.
  • "Build anything" budget - dedicated funding for whatever tools, libraries, datasets, or infrastructure you need to solve technical challenges, no questions asked.
  • Learning budget - attend any conference, course, or program that makes you better at what we're building.


Our Operating Principles
  • Trust as the Default: People do their best work when confidence is mutual. We show our work, keep our promises, and flag risks before they bite. Trust isn't an aspiration - it's the baseline.
  • Forward-Deployed with Product DNA: We own customer outcomes while building a product company. We don't win if they don't win.
  • Extreme Ownership: If you notice a problem, you own it by making sure it doesn't fall through the cracks. Authority comes from initiative, not job titles. Once you step up, you're accountable for the outcome.
  • Production-First Engineering: We design for critical workloads from day one. Durable execution, blue/green deploys, automated rollbacks, continuous delivery with end-to-end observability.
  • Communicate with Intent: Great work disappears without great communication. We push information to the people who need it, when they need it. Silence is never the safe choice.
  • Earn it Every Day: Your work speaks for itself. We create an environment where the best idea wins, the strongest work gets recognized, and everyone is held to the same high standard.
  • Keep Raising the Bar: Great teams compound. Every hire raises the bar, every win gets named, every person gets the tools and runway to grow.

Similar Jobs

More Jobs at Kepler Group

More Information Technology Jobs

Find similar Agentic Systems Engineer jobs: