- Location
- New York, NY, United States
- Employment
- Full time
- Level
- Mid level
- Posted
- 3 months ago
About this role
About Traversal
Traversal is the AI Site Reliability Engineer (SRE) for the enterprise - already trusted by some of the largest companies in the world to troubleshoot, remediate, and even prevent the most complex production incidents. Our mission is to free engineers from endless firefighting and enable them to focus on creative, high-impact work.
Our roots remain deeply embedded in AI research, and we’re channeling that scientific rigor and creativity into building the premier AI agent lab for the enterprise. Hence, what we’re proudest of is assembling the most talented yet nicest group of individuals, including researchers from MIT, Harvard, and Berkeley, to world-class engineers from industry: Citadel Securities, Cockroach Labs, Datadog, DE Shaw, ServiceNow, Glean, Perplexity, Pinecone, and more, to take on one of the hardest problems for AI to solve. Without the entire team, none of this would be possible.
The Role
As a Forward Deployed Engineer at Traversal, you will own the last-mile delivery of the Traversal product by deploying frontier LLMs in production, optimizing multi-step reasoning workflows, and improving the performance and accuracy of our agentic systems. You will be responsible for the end-to-end deployment of the agent into the largest enterprises and ensuring it delivers accurate and actionable insights.
This is a highly technical role at the intersection of AI, systems engineering, and customer impact. You’ll work across teams to ensure our AI agents interact seamlessly with large volumes of observability data and scale reliably across environments. The insights you generate through your customer-driven work will directly impact roadmaps for all other engineering teams at the company.
Responsibilities
AI Agent Architecture: Design and build multi-step reasoning workflows that use LLMs and other components to analyze large-scale observability data.
Prompting & Tooling: Develop tooling for prompt engineering, function calling, and agentic orchestration that optimizes latency, reliability, and performance.
Production Deployment: Own technical delivery across multiple deployments from first prototype to robust, scalable production environments with low latency and high uptime requirements.
Last-Mile Delivery: Build the custom workflows, services, and integrations required for last-mile delivery. This work is deeply technical, and you should be comfortable diving into any layer of the stack including frontend, backend, data pipelines, or infrastructure to execute whatever changes or additions a customer engagement requires.
Engineering Strategy: Work closely with Platform, Infra, Product, and GTM teams to ensure our AI agents are well-integrated into the broader system and are delivering customer value.
Customer Communication: Lead technical conversations with enterprise customers, asking the right questions to unlock relevant data sources, shape agent discovery approaches, and drive successful deployment of the agent into their environment. You will help build out their roadmaps based on needs you identify in collaboration with customer counterparties.
Requirements
3+ years of software engineering experience.
Strong Python skills, including experience with Asyncio and Pandas.
Strong communicator who is comfortable translating technical details into clear, actionable insights for both engineering and customer stakeholders.
Experience building and deploying distributed systems and/or large-scale data processing pipelines.
Familiarity with LLMs, prompt engineering, or agentic frameworks.
Proven ability to deliver projects end-to-end: from experimentation to production deployment.
Strong debugging skills and ability to work across layers (data, infra, model, and application).
Nice to Have
Experience with observability data and production infrastructure (logs, metrics, traces).
Background in AI agent research or open-source contributions to agent frameworks.
Experience working with Terraform, Kubernetes, or ML orchestration platforms.
Compensation
We offer competitive compensation, startup equity, health insurance, and additional benefits. The U.S. base salary range for this full-time, in-person role in New York is $150,000 - $300,000, plus equity and benefits. Our salary ranges are based on location, level, and role. Individual compensation is determined by experience, skills, and job-related knowledge.
Why You Should Join Us
We’ll make sure you’re fully supported with health insurance, a great tech setup, flexible time off, and plenty of in-office snacks. We offer competitive salary and equity packages, and take thoughtful consideration with every hire on our small, high-impact team.
Traversal is fully in-office, 5 days a week, based in New York near Madison Square Park. We have a collaborative, hard-working culture and are energized by building the future of AI-powered software maintenance.
Working here means owning meaningful parts of the product, having the flexibility to move fast, and learning constantly. This is a place to grow your career, make a real impact, and help define a new category of infrastructure software.
As published by Traversal. Applications are handled on their site.
Skills this posting mentions
About Traversal
Traversal is building an AI site reliability engineer that troubleshoots, remediates, and even prevents production issues in complex software systems - always on call, so engineers don’t have to be. Already deployed in some of the world’s largest enterprises, Traversal improves the resilience of mission-critical systems - reducing MTTD and MTTR by up to 90% and supporting services that reach millions globally. We’re a team of top-tier AI researchers and engineers. To learn more about Traversal or apply for open roles, visit: https://job-boards.greenhouse.io/traversal
All 15 openings at TraversalOne click, then it is written
Apply to Traversal with a resume written for this role.
Queue AI Engineer - Forward Deployed Engineer and I read the posting, rewrite your resume against it, draft the cover letter, and score the fit. Then you press send, or press one button and I fill in Traversal’s form for you.
01Drop your resume
A PDF or a LinkedIn URL. About a minute, once.
02I rank the openings
Every weekday morning, the live catalog scored against your history. Up to 20 worth your time, not two hundred links.
03Each one is written up
Resume rewritten for the posting, a cover letter, a fit score. Press send, or let me fill in the form.
- 25 sent a week, free
- No card
- Nothing sent until you say so
More roles at Traversal
See all- 5 weeks ago
- 5 weeks ago
- 7 weeks ago
- 8 weeks ago
- 2 months ago
- 2 months ago
Similar roles elsewhere
See more- Today
- Today
Senior Software Engineer, Backend (Product Engineering)
BrexNew York, New York
$192k - $240k/yrSeniorEngineering
Put this to work
Paste your career in once. Every application after that is written for you.
Drop a resume or a LinkedIn URL. I rank the live openings against it, rewrite the resume and write a cover letter for the best of them, and fill in the employer's form when you press the button. You read, you decide what goes out.
01Drop your resume
A PDF or a LinkedIn URL. About a minute, once.
02I rank the openings
Every weekday morning, the live catalog scored against your history. Up to 20 worth your time, not two hundred links.
03Each one is written up
Resume rewritten for the posting, a cover letter, a fit score. Press send, or let me fill in the form.
- New matches ranked and written before you are up.
- Every bullet stays inside what your history supports. Nothing invented.
- Queued, submitted, interviewing, offer: one screen, not a spreadsheet.
500 free credits on sign-up. No card. Nothing is sent until you say so.
Listed from the job board Traversal publishes. Refolk is not the employer and does not handle their hiring. Applications go to Traversal directly.