- Location
- USA
- Workplace
- Remote
- Employment
- Full time
- Level
- Staff
- Posted
- Yesterday
About this role
About LILT
AI is changing how the world communicates - and LILT is leading that transformation.
We're on a mission to make the world's information accessible to everyone, regardless of the language they speak. We use cutting-edge AI, machine translation, and human-in-the-loop expertise to translate content faster, more accurately, and more cost-effectively without compromising on brand, voice, or quality.
At LILT, we empower our teammates with leading tools, global collaboration, and growth opportunities to do their best work. Our company virtues - Work together, win together; Find a way or make one; Dance in the customer's shoes; Quicker than they expect; Quality is Job 1 - guide everything we do. We are trusted by Intel Corporation, Canva, the United States Department of Defense, the United States Air Force, ASICS, and hundreds of global Enterprises. Backed by Sequoia, Intel Capital, and Redpoint, we’re building a category-defining company in a $50B+ global translation market being redefined by AI.
As part of LILT’s Internal Tools group, you will build and run the tools that LILT's own delivery, contributor-ops, and engineering teams rely on to operate LILT's Applied AI (AAI) benchmarking business - which builds and delivers multilingual benchmarks and evaluation data to frontier AI labs using LILT's global network of subject-matter experts. As the AAI business grows, expect the scope of these platforms to grow in parallel. This platform sits directly on the critical path of paid customer deliverables with real SLAs, maintained by a small, high-leverage engineering team that produces reliable, production-ready tooling. You will own significant surface area across both end-to-end: architecture, data-pipeline design, and the hands-on engineering that keeps this infrastructure reliable. This is a high-impact, high-visibility role: the reliability and craft you bring directly enables on-time, high-quality delivery for some of the most prominent AI labs in the world.
In this role, you will drive the long-term technical strategy for internal platforms while working directly with the delivery, ops, and engineering teams across the business who depend on them daily. This team ships cross-service automation in careful stages - advisory first, then assistive, only later decision-relevant, always with a human fallback - and you'll be expected to hold that same bar as you extend these systems. You will partner with a small existing team to raise the engineering bar across two very different runtimes, set technical direction, and make the calls that determine how this infrastructure scales as the business grows.
What You'll Do
Partner with the benchmarking business's researchers and TPMs to translate new benchmark and data-quality requirements into scoped technical designs
Own the internal workforce-management tools on our internal platform for AI delivery for hundreds of external contributors: vetting flows, candidate assessment, QC, payment/delivery tracking, and roster/reporting exports
Own architecture and long-term technical direction across multiple services and the platform
Extend IAA and audio-QA pipelines: annotator outlier detection, ASR sidecar enhancements, LLM-based QC, and DNSMOS/librosa audio-quality scoring
Design and ship new modules on the platform that plug new benchmark and vetting workflows into the existing multi-stage review lifecycle
Own and extend the platform's API key provisioning and budget-governance system
Build self-serve ChatOps-style automation for the internal engineering org and contributor base to enable accelerated annotator workflows and query resolution
Harden background worker and job-processing infrastructure
Set and enforce testing, CI/CD, and deployment practices across both codebases - from unit/integration testing through infrastructure-as-code and release automation
Raise the technical bar for a small, high-leverage team through code review, design docs, and mentoring as the surface area grows
What We're Looking For
Required
5+ years of professional full-stack software engineering experience, with a track record of owning production systems end-to-end across more than one runtime/language
Deep experience with a Python backend framework (FastAPI or comparable) plus async SQLAlchemy/Postgres, alongside production experience in at least one statically-typed backend language (Go, Java, or similar)
Strong React/TypeScript frontend experience - component architecture, state management (Zustand, Redux, or similar), and a modern data-fetching layer (TanStack Query or comparable)
Experience building and operating background job/worker systems (queue-driven or polling-based) with failure tolerance and idempotency in mind
Experience integrating with third-party and platform APIs - including the GitHub API, OAuth/OIDC SSO, and at least one LLM API (Gemini, OpenAI, or similar) - handling auth, rate limits, and webhook-driven sync
Experience building Slack (or comparable chat-platform) bot integrations that automate internal workflows - resource provisioning, approvals, budget/TTL enforcement - with real operational guardrails, not just CRUD features
Comfort reading and extending applied-statistics or ML-adjacent code (agreement metrics, audio-quality scoring, or comparable data-quality tooling)
Solid grasp of CI/CD, containerized deployment (Docker, Helm, ArgoCD/GitOps or comparable), and infrastructure-as-code (Terraform or comparable)
Experience debugging and optimizing native-library (numpy/scipy/onnxruntime-class) memory growth in long-running Python worker processes via safe, boundary-aware process recycling - not just raising memory limits
Experience building and owning internal platforms/tools that increase leverage for a non-engineering team (research, operations, support, data/workforce management, or similar) - not solely external-customer-facing product work
Strong Plus
Experience with audio/speech pipelines: ASR (Whisper or similar) or audio-quality metrics (DNSMOS, librosa)
Experience building internal tools for managing a data-labeling, annotation, or crowdsourced-contributor workforce (vetting, QC, payments)
Experience with inter-annotator agreement or statistical agreement metrics
Notification and delivery systems experience - Slack bot integrations, transactional email, and idempotent delivery guarantees
Experience designing abstractions over heterogeneous data sources with different consistency guarantees - e.g. a fully-replayable event history vs. an observe-only current-state API requiring synthesized diffing - behind one common interface
Experience building ChatOps-style automation - Slack or GitHub PR-comment bot commands that trigger backend workflows or CI/CD runs
Experience implementing short-lived, rotatable service-to-service JWT auth (key-ID-based rotation, replay-protected tokens, fail-fast config validation) alongside a separate human-facing SSO flow in a paired service
Comfort owning both sides of a system with genuinely different runtimes without a large team to lean on
Prior experience as the primary or sole engineer on a small, high-leverage internal platform
Bonus
Experience with LLM-as-judge or LLM-based QA/review pipelines
Familiarity with OpenTelemetry or comparable observability instrumentation in Go services
Familiarity with LLM provider gateway/routing services (OpenRouter or comparable) - model aliasing, rate-limit and timeout handling, and budget enforcement
Experience with data export/reporting tools (Excel generation, CSV pipelines, or BI-style dashboards)
Our Story
Our founders, Spence and John met at Google working on Google Translate. As researchers at Stanford and Berkeley, they both worked on language technology to make information accessible to everyone. While together at Google, they were amazed to learn that Google Translate wasn’t used for enterprise products and services inside the company.The quality just wasn’t there. So they set out to build something better. LILT was born.
LILT has been a machine learning company since its founding in 2015. At the time, machine translation didn’t meet the quality standard for enterprise translations, so LILT assembled a cutting-edge research team tasked with closing that gap. While meeting customer demand for translation services, LILT has prioritized investments in Large Language Models, human-in-the-loop systems, and now agentic AI.
With AI innovation accelerating and enterprise demand growing, the next phase of LILT’s journey is just beginning.
Our Tech
What sets our platform apart:
Brand-aware AI that learns your voice, tone, and terminology to ensure every translation is accurate and consistent
Agentic AI workflows that automate the entire translation process from content ingestion to quality review to publishing
100+ native integrations with systems like Adobe Experience Manager, Webflow, Salesforce, GitHub, and Google Drive to simplify content translation
Human-in-the-loop reviews via our global network of professional linguists, for high-impact content that requires expert review
LILT in the News
Featured in The Software Report’s Top 100 Software Companies!
LILT makes it onto the Inc. 5000 List.
LILT’s continues to be an intellectual powerhouse, holding numerous patents that help power the most efficient and sophisticated AI and language models in the industry.
Check out all our news on our website.
Information collected and processed as part of your application process, including any job applications you choose to submit, is subject to LILT's Privacy Policy at https://lilt.com/legal/privacy.
At LILT, we are committed to a fair, inclusive, and transparent hiring process. As part of our recruitment efforts, we may use artificial intelligence (AI) and automated tools to assist in the evaluation of applications, including résumé screening, assessment scoring, and interview analysis. These tools are designed to support human decision-making and help us identify qualified candidates efficiently and objectively. All final hiring decisions are made by people. If you have any concerns, require accommodations, or would like to opt-out of the use of AI in our hiring process, please let us know at recruiting@lilt.com.
LILT is an equal opportunity employer. We extend equal opportunity to all individuals without regard to an individual’s race, religion, color, national origin, ancestry, sex, sexual orientation, gender identity, age, physical or mental disability, medical condition, genetic characteristics, veteran or marital status, pregnancy, or any other classification protected by applicable local, state or federal laws. We are committed to the principles of fair employment and the elimination of all discriminatory practices.
As published by Lilt. Applications are handled on their site.
Skills this posting mentions
About Lilt
Lilt powers the global experience across every step of the customer journey. We bring human-powered, technology-assisted translations to global enterprises. We give organizations everything they need to scale their translation programs, go-to-market faster than ever, and improve the global customer experience. Lilt’s translation services are powered by our translation technology, which improves translator speeds by 3-5x and reduces localization costs by 50% or more. We're based in San Francisco with global offices in Berlin, Dublin, Indianapolis, Washington, D.C., and London, and are backed by Intel Capital, Sequoia, Redpoint, Zetta, and XSeed.
All 16 openings at LiltOne click, then it is written
Apply to Lilt with a resume written for this role.
Queue Staff Fullstack Engineer - Internal Tools and I read the posting, rewrite your resume against it, draft the cover letter, and score the fit. Then you press send, or press one button and I fill in Lilt’s form for you.
01Drop your resume
A PDF or a LinkedIn URL. About a minute, once.
02I rank the openings
Every weekday morning, the live catalog scored against your history. Up to 20 worth your time, not two hundred links.
03Each one is written up
Resume rewritten for the posting, a cover letter, a fit score. Press send, or let me fill in the form.
- 25 sent a week, free
- No card
- Nothing sent until you say so
More roles at Lilt
See all- Yesterday
- 5 days ago
Senior Full Stack Engineer
Washington D.C., District of ColumbiaHybrid
$125k - $166k/yrSeniorEngineering - 5 days ago
- 5 days ago
Machine Learning Engineer (Real-Time Speech Translation)
Washington D.C., District of ColumbiaHybrid
$120k - $161k/yrMid levelData and ML - 5 days ago
- Last week
Similar roles elsewhere
See more- Today
Staff Software Engineer, Payments (AirCover Insurance Platform)
AirbnbRemote - USARemote
$212k - $265k/yrStaffEngineering - Today
Engineering Manager, Payments (AirCover Insurance Platform)
AirbnbRemote - USARemote
$212k - $265k/yrManagerEngineering - Yesterday
Senior ServiceNow Platform Engineer
Abnormal SecurityRemote - USARemote
$142k - $204k/yrSeniorEngineering
Put this to work
Paste your career in once. Every application after that is written for you.
Drop a resume or a LinkedIn URL. I rank the live openings against it, rewrite the resume and write a cover letter for the best of them, and fill in the employer's form when you press the button. You read, you decide what goes out.
01Drop your resume
A PDF or a LinkedIn URL. About a minute, once.
02I rank the openings
Every weekday morning, the live catalog scored against your history. Up to 20 worth your time, not two hundred links.
03Each one is written up
Resume rewritten for the posting, a cover letter, a fit score. Press send, or let me fill in the form.
- New matches ranked and written before you are up.
- Every bullet stays inside what your history supports. Nothing invented.
- Queued, submitted, interviewing, offer: one screen, not a spreadsheet.
500 free credits on sign-up. No card. Nothing is sent until you say so.
Listed from the job board Lilt publishes. Refolk is not the employer and does not handle their hiring. Applications go to Lilt directly.