← WORK 목록으로
GLOBALRemoteOK서버/인프라REMOTE

Harper - Engineering Manager TLM Platform

미국 상업용 보험 유통을 AI 중심으로 자동화하는 Harper의 플랫폼 엔지니어링 팀을 이끌 엔지니어링 매니저(또는 TLM)를 채용하는 정규직 공고입니다. Railway와 AWS 위 200개 이상 서비스와 병렬 에이전트 워크로드를 운영하며, 인프라·개발자 생산성·SRE/프로덕션 지원 세 영역을 총괄하고 2명 팀을 3개월 내 5~6명으로 키우는 역할입니다. 초기에는 직접 손으로 코드를 작성하는 핸즈온 리더십이 요구되어, SRE 운영 경험과 6~10명 팀 관리 경험을 갖춘 플랫폼/인프라 리더에게 적합합니다.

2026.09.12VIEW 69RemoteOK에서 수집
Budget협의
Difficulty전문가
Duration정규직(장기)
Work style원격 가능
Required stack

필요 기술

AWSRailwayDockerCloud InfraAI/LLMObservabilitySRE Tooling
Project brief

프로젝트 내용

The Problem
36 million businesses in America need insurance—it's not optional. 77% are underinsured. 40% have no coverage at all. The distribution system failed them: too slow, too opaque, too confusing.
Over 90% of commercial insurance is still human-led. We're building the inverse: 90%+ AI-led, pushing toward the higher 90s. Not by patching legacy workflows—by building AI that makes humans more effective, improves the customer experience, and eliminates friction at every step.
We're adding ~1,000 customers per month. We've grown 100x since last year. We're looking to do even more this year—and that's why we're hiring.
You'll build the team and the systems the rest of engineering depends on. When they compound, everyone gets faster.

The Thesis
Harper runs 200+ services across Railway and AWS. We orchestrate N parallel agentic workloads. Our AI systems make thousands of decisions a day, and every decision needs to be traceable, evaluated, and cheaper to run tomorrow than it is today. The infrastructure underneath all of that is the difference between a company that scales and one that stalls.
Great platform engineering here isn't invisible—it's the reason product ships fast. When observability catches a silent failure before a customer does. When a pooling layer makes connection exhaustion a non-issue. When developer velocity doubles because someone built the tool nobody knew they needed. That leverage compounds across every other engineer on the team.

The Role
This role is open both to an experienced EM and to someone on the path to becoming an EM who doesn't yet have significant management experience—we're happy to bring that person on and support the transition.
This is a team that still needs to be built. We have 2 people in it today, and we believe this area needs significant investment at Harper. The role owns three things: infrastructure (cloud, third-party vendors, and everything surrounding it), developer productivity (making sure we're genuinely an AI-first shop, with the right AI tooling and harnesses in place for every other engineering team to be successful), and SRE and production support (production discipline is top of mind as we head into a high-growth, high-scale phase—we need someone who has run an SRE function before). Each of these three areas can grow into a larger team of its own over time.

The team is 2 people today, growing to five to six over the next three months. Because the team is small, whoever takes this role—EM or TLM—will need to be hands-on at the start, working directly with the team and doing the work themselves.

What You'll Do
- Own infrastructure — cloud, third-party vendors, and everything surrounding it

- Build developer productivity — the AI tooling and harnesses that make every other engineering team genuinely AI-first

- Own SRE and production support — bring production discipline as we head into a high-growth, high-scale phase

- Build out the team — grow from 2 people to five to six over the next three months

- Stay hands-on — with a team this small, work directly alongside your engineers and do the work yourself, not just direct it

- Grow each function over time — infrastructure, developer productivity, and SRE/production support can each become a larger team of its own

You Might Be a Fit If...
- Have at least 2 years as a line manager directly managing teams of six to ten people—or clear readiness to grow into that role

- You've been a line manager for at least two years, directly managing teams of six to ten people—or you're on a clear path there and want to grow into it

- You have hands-on experience with AWS or GCP and strong database troubleshooting skills

- You've seen real scale—on the order of at least a million requests per minute

- You can troubleshoot across databases, networking, compute, Docker, Kubernetes, and other aspects of infrastructure

- You care as much about developer productivity and AI tooling as you do about raw infrastructure

- You know on-call practices, SLAs, SLOs, and SRE practices generally

- You want to stay hands-on and grow into a fuller managerial role over time

- You move fast, with the same appetite as the rest of Harper

Nice to Have
- Your own network to draw on for hiring—otherwise, comfort working closely with recruiters

- Experience building developer productivity tooling and AI harnesses for other engineering teams

- Experience building a team from a very small base

- Prior startup experience—especially at a company scaling through hypergrowth

Compensation & Logistics
- Salary: $225,000–$275,000

- Location: San Francisco, in-office. Based in SF or willing to relocate.

- Schedule: Monday–Friday, in-office five days a week.

- Benefits: Uber commuter benefits; breakfast, lunch, and dinner provided; snacks, drinks, and coffee daily; free gym membership; health, dental, and vision insurance.

The Process
- Hiring Manager screen - 30 min video chat.

- Technical screen — 60 min remote: project deep dive + system design

- Super Day on-site — meet the team, sit in on the operation, do real work alongside us.

Please mention the word **DAUNTLESS** and tag RMTUuMTY1LjI0MC4xNDk= when applying to show you read the job post completely (#RMTUuMTY1LjI0MC4xNDk=). This is a beta feature to avoid spam applicants. Companies can search these words to find applicants that read this and see they're human.
지원 기회 분석

지원 전에 볼 것

핵심 요구사항

  • 6~10명 규모 팀을 직접 관리한 2년 이상의 라인 매니저 경험(또는 성장 가능성)
  • SRE 기능을 운영해 본 경험과 프로덕션 운영 규율
  • AWS·Railway 기반 클라우드 인프라 및 서드파티 벤더 관리
  • AI-first 개발 생산성 도구/하네스 구축
  • 고성장·고확장 단계에서의 관측성 및 안정성 확보
  • 초기 핸즈온 실무 수행 능력

예상 산출물

  • 인프라·개발자 생산성·SRE를 아우르는 플랫폼 팀 구축
  • AI 워크로드 관측성 및 프로덕션 운영 체계
  • 개발 생산성 향상을 위한 AI 툴링/하네스
  • 3개월 내 5~6명 규모 팀 채용 및 온보딩

매력 포인트

  • 100x 성장한 고성장 스타트업
  • AI-first 대규모 인프라 운영 경험
  • 완전 원격 국제 채용

확인할 점

  • 예산·연봉 정보 미기재
  • 외주 프로젝트가 아닌 정규직 채용 공고(프리랜서 부적합)
Client signal

클라이언트 정보

회사 · Harper / 지역 · San Francisco
START THE LOOP · CHOOSE

시장과 사람의 답을 봤다면,
다음 결과물의 구조를 고릅니다.

한 번의 결과에 기대지 않고 다시 만들 수 있도록, 문제 발견부터 제작·배포·수익화까지 이어지는 전체 흐름을 익혀보세요.

TTJ CLASS에서 다음 구조 고르기 →
처리 중...