- Location
- New York, NY, United States
- Employment
- Full time
- Level
- Mid level
- Posted
- 15 months ago
About this role
About Traversal
Traversal is the AI Site Reliability Engineer (SRE) for the enterprise - already trusted by some of the largest companies in the world to troubleshoot, remediate, and even prevent the most complex production incidents. Our mission is to free engineers from endless firefighting and enable them to focus on creative, high-impact work.
Our roots remain deeply embedded in AI research, and we’re channeling that scientific rigor and creativity into building the premier AI agent lab for the enterprise. Hence, what we’re proudest of is assembling the most talented yet nicest group of individuals, including researchers from MIT, Harvard, and Berkeley, to world-class engineers from industry: Citadel Securities, Cockroach Labs, Datadog, DE Shaw, ServiceNow, Glean, Perplexity, Pinecone, and more, to take on one of the hardest problems for AI to solve. Without the entire team, none of this would be possible.
The Role
As an AI Researcher at Traversal, you'll work on improving the accuracy and speed of our agents - systems that autonomously diagnose and resolve production incidents for some of the world’s largest enterprises.
This is a hands-on, production-oriented research role. You will design experiments, run them on real data, and ship improvements. Not flag them - ship them.
You'll work end-to-end: identify a failure mode in agent reasoning, design an intervention, evaluate it against real customer traces, and get it into production. The tooling and infrastructure are in place. The research problem is hard. The feedback loop is fast. This is not a publish-papers role. It is a make-the-agent-work role - build-and-ship cutting edge AI.
Responsibilities
LLM & Agent Research: Prototype and evaluate prompting strategies, reasoning workflows, and tool-use policies for agents operating on large-scale observability data and complex troubleshooting workflows. Ship improvements to production.
Evaluation Design: Build and maintain eval harnesses that measure real accuracy improvements on actual customer incident types - not just benchmark scores. Own the loop from hypothesis to production measurement.
Cross-Team Collaboration: Work closely with AI engineers, infrastructure teams, and product leads to bring research into production and close the loop between experimentation and impact.
Stay on the Frontier: Track developments in LLMs, agent architectures, and AI alignment, translating insights into actionable improvements for Traversal’s domain.
Training & Alignment: Apply fine-tuning, reinforcement learning, and reward modeling techniques to align AI behavior with real-world SRE workflows.
Synthetic Data & Experimentation: Design pipelines to generate synthetic incidents and observability signals, enabling scalable training and testing in data-scarce environments.
Requirements
PhD in Computer Science, Electrical Engineering, Statistics, or a related technical field; demonstrated depth in LLMs, agents, or applied machine learning
Deep applied AI expertise, including strong working knowledge of LLMs, transformers, reinforcement learning, or neural networks in agentic systems
Strong judgment in model evaluation and experimental iteration to improve product accuracy and behavior
Strong software engineering depth, with the ability to work effectively in a complex production codebase and ship production-quality code
Some experience shipping AI or ML systems to production
Ability to run rigorous experiments, interpret results, and quickly translate learnings into product improvements
Startup or early-team experience, with comfort operating in ambiguous environments and building without mature infrastructure
Nice to Have
Experience in SRE, observability, or backend systems, especially when paired with strong AI/ML depth
Experience with RLHF, synthetic data pipelines, or LLM evaluation tooling
Contributions to open-source agent frameworks such as LangGraph, DSPy, or similar
Research experience in LLMs, agents, or reinforcement learning, including publications in venues such as NeurIPS, ICML, or ICLR; top-tier conference publications are a plus
Compensation
We offer competitive compensation, startup equity, health insurance, and additional benefits. The U.S. base salary range for this full-time, in-person role in New York is $160,000 - $300,000, plus equity and benefits. Our salary ranges are based on location, level, and role. Individual compensation is determined by experience, skills, and job-related knowledge.
Why You Should Join Us
We’ll make sure you’re fully supported with health insurance, a great tech setup, flexible time off, and plenty of in-office snacks. We offer competitive salary and equity packages, and take thoughtful consideration with every hire on our small, high-impact team.
Traversal is fully in-office, 5 days a week, based in New York near Madison Square Park. We have a collaborative, hard-working culture and are energized by building the future of AI-powered software maintenance.
Working here means owning meaningful parts of the product, having the flexibility to move fast, and learning constantly. This is a place to grow your career, make a real impact, and help define a new category of infrastructure software.
As published by Traversal. Applications are handled on their site.
Skills this posting mentions
About Traversal
Traversal is building an AI site reliability engineer that troubleshoots, remediates, and even prevents production issues in complex software systems - always on call, so engineers don’t have to be. Already deployed in some of the world’s largest enterprises, Traversal improves the resilience of mission-critical systems - reducing MTTD and MTTR by up to 90% and supporting services that reach millions globally. We’re a team of top-tier AI researchers and engineers. To learn more about Traversal or apply for open roles, visit: https://job-boards.greenhouse.io/traversal
All 15 openings at TraversalOne click, then it is written
Apply to Traversal with a resume written for this role.
Queue AI Researcher and I read the posting, rewrite your resume against it, draft the cover letter, and score the fit. Then you press send, or press one button and I fill in Traversal’s form for you.
01Drop your resume
A PDF or a LinkedIn URL. About a minute, once.
02I rank the openings
Every weekday morning, the live catalog scored against your history. Up to 20 worth your time, not two hundred links.
03Each one is written up
Resume rewritten for the posting, a cover letter, a fit score. Press send, or let me fill in the form.
- 25 sent a week, free
- No card
- Nothing sent until you say so
More roles at Traversal
See all- 5 weeks ago
- 5 weeks ago
- 7 weeks ago
- 8 weeks ago
- 2 months ago
- 2 months ago
Similar roles elsewhere
See more- 2 days ago
Research Science Intern (PhD)
DatadogNew York, New York +1
$110k - $140k/yrInternshipScience and research - 2 days ago
- 2 days ago
Staff Applied Scientist - Agentic Interfaces
DatadogNew York, New York
$276k - $345k/yrStaffScience and research - 2 days ago
Staff Applied Scientist - Dashboards
DatadogNew York, New York
$276k - $345k/yrStaffScience and research - 2 days ago
AI Research Scientist - Datadog AI Research (DAIR)
DatadogNew York, New York +1
$320k - $400k/yrScience and research - 5 days ago
Staff Applied Scientist, LTV Modeling
FaireNew York City, NY +1
$247k - $339k/yrStaffScience and research
Put this to work
Paste your career in once. Every application after that is written for you.
Drop a resume or a LinkedIn URL. I rank the live openings against it, rewrite the resume and write a cover letter for the best of them, and fill in the employer's form when you press the button. You read, you decide what goes out.
01Drop your resume
A PDF or a LinkedIn URL. About a minute, once.
02I rank the openings
Every weekday morning, the live catalog scored against your history. Up to 20 worth your time, not two hundred links.
03Each one is written up
Resume rewritten for the posting, a cover letter, a fit score. Press send, or let me fill in the form.
- New matches ranked and written before you are up.
- Every bullet stays inside what your history supports. Nothing invented.
- Queued, submitted, interviewing, offer: one screen, not a spreadsheet.
500 free credits on sign-up. No card. Nothing is sent until you say so.
Listed from the job board Traversal publishes. Refolk is not the employer and does not handle their hiring. Applications go to Traversal directly.