Traversal

AI Research Engineer

New York · Posted 1h ago

$180-300Kpermanentonsite
LLM

Job Description

About Traversal

Traversal is the AI Site Reliability Engineer (SRE) for the enterprise—already trusted by some of the largest companies in the world to troubleshoot, remediate, and even prevent the most complex production incidents. Our mission is to free engineers from endless firefighting and enable them to focus on creative, high-impact work.

Our roots remain deeply embedded in AI research, and we’re channeling that scientific rigor and creativity into building the premier AI agent lab for the enterprise. Hence, what we’re proudest of is assembling the most talented yet nicest group of individuals, including researchers from MIT, Harvard, and Berkeley, to world-class engineers from industry: Citadel Securities, Cockroach Labs, Datadog, DE Shaw, ServiceNow, Glean, Perplexity, Pinecone, and more, to take on one of the hardest problems for AI to solve. Without the entire team, none of this would be possible.

The Role

As an AI Research Engineer at Traversal, you'll build the systems our research team runs on. Our agents autonomously diagnose and resolve production incidents for some of the world's largest enterprises, and improving them is an experimental problem: form a hypothesis about agent behavior, test it against real incident data, and ship what works. The bottleneck is rarely ideas. It's the infrastructure to run the experiment, trust the result, and scale the prototype.

That infrastructure is what you own. You'll build the harnesses that let researchers evaluate agents against real customer incidents, the tooling that makes agent behavior legible enough to debug, and the systems that take a promising prototype from one run to thousands. You'll design experiments alongside researchers rather than only implementing them, and you'll carry the results across the line into production.

This role sits on the research team, and it is an engineering role. We are looking for a strong backend engineer with real experimentation instincts — someone who has always run experiments on the side and wants that to be the job. No PhD required. If you want to work at the frontier of LLM agents while still building serious systems, this is the seat.

Responsibilities

  • Research Infrastructure: Build and scale the infrastructure behind the team's research — experiment harnesses, evaluation pipelines, and trajectory tooling — so that ideas can be tested quickly and results can be trusted.

  • Scaling Research Prototypes: Take promising research prototypes and make them fast, reliable, and production-ready, closing the gap between a one-off experiment and a capability that ships to enterprise customers.

  • Experiment Design: Partner with researchers to design experiments — framing the hypothesis, choosing the measurement, and making sure the result is real before it drives a decision.

  • Cross-Team Collaboration: Work closely with AI engineers, infrastructure teams, and product leads to bring research into production and close the loop between experimentation and impact.

  • Stay on the Frontier: Track developments in LLMs, agent architectures, and evaluation methodology, translating insights into actionable improvements for Traversal's domain.

  • LLM & Agent Research: Prototype and evaluate prompting strategies, reasoning workflows, and tool-use policies for agents operating on large-scale observability data and complex troubleshooting workflows. Ship improvements to production.

Requirements

  • 5+ years of software engineering experience, with a strong focus on backend development and experience building scaled systems.

  • Hands-on experimentation experience — running experiments, evaluating results, and iterating on them. Side projects and personal work count.

  • Proven ability to lead in fast-paced startup environments, with limited resources, shifting priorities, and minimal structure.

  • Strong experience with Python and web frameworks such as FastAPI.

  • Experience deploying applications on AWS, working with ECS or Kubernetes, Postgres for data storage, and S3 for large-scale object storage.

  • Familiarity with observability and monitoring tools to ensure the system is running smoothly and efficiently.

Nice to Have

  • A research engineer background, or prior experience as the engineer embedded in a research team.

  • Experience building or operating agentic AI systems.

  • Background in large-scale, complex, data-driven applications.

  • Familiarity with AI or LLM-powered products, and with evaluation or benchmarking tooling.

Compensation

We offer competitive compensation, startup equity, health insurance, and additional benefits. The U.S. base salary range for this full-time, in-person role in New York is $180,000–$300,000, plus equity and benefits. Our salary ranges are based on location, level, and role. Individual compensation is determined by experience, skills, and job-related knowledge.

Why You Should Join Us

We’ll make sure you’re fully supported with health insurance, a great tech setup, flexible time off, and plenty of in-office snacks. We offer competitive salary and equity packages, and take thoughtful consideration with every hire on our small, high-impact team.

Traversal is fully in-office, 5 days a week, based in New York near Madison Square Park. We have a collaborative, hard-working culture and are energized by building the future of AI-powered software maintenance.

Working here means owning meaningful parts of the product, having the flexibility to move fast, and learning constantly. This is a place to grow your career, make a real impact, and help define a new category of infrastructure software.