- Location
- Sofia, Sofia City, Bulgaria
- Workplace
- Hybrid
- Employment
- Full time
- Level
- Mid level
- Posted
- Last week
About this role
The job
We're looking for a TypeScript Engineer with Site Reliability Engineering (SRE) mentality to join our Webservices Platform team. You'll help build and operate a central cloud service that handles customer authentication and authorization, storage, and filespace provisioning - critical infrastructure that other engineering teams and functions like Legal and Finance, depend on for uptime and SLA commitments.
We follow an SRE model where the engineers who write a service also monitor and operate it in production. You'll split your time between writing backend code in Node.js/TypeScript, and applying reliability engineering practices - observability, incident response, SLI/SLO - to keep that code healthy in production. Over time, you'll take ownership of initiatives across the platform, including its next major architecture generations (Web Service 2x/3x), as we modernize toward high availability and disaster recovery.
What you'll do:
Write and maintain backend services in Node.js and TypeScript, and manage the cloud infrastructure (AWS) they run on.
Set up and maintain observability - metrics, logging, distributed tracing, and alerting - using tools such as Cloudwatch, Elastic, Prometheus, InfluxDB, and Grafana.
Participate in on-call rotations: triage incidents, drive rapid mitigation, lead root-cause analysis, and write postmortems. (You'll ramp into on-call gradually - first learning the service, its alerts and runbooks)
Define and track SLIs, SLOs, and SLAs for the services you own, and use them to prioritize reliability work.
Reduce operational toil by building internal tooling, CI/CD pipelines, and infrastructure-as-code.
Contribute to modernization efforts aimed at high availability, disaster recovery, and high uptime, and help remove architectural obstacles blocking that work.
Support provisioning and platform work for new customers alongside reliability and feature work.
Collaborate closely with the rest of the Web Service Platform team, Platform Engineering, and other engineering teams whose services depend on this platform.
Key responsibilities:
Design and develop reliable, scalable TypeScript/Node.js services that power authentication, storage, filesystem/filespace provisioning, billing and other services for our customers.
Own what you build in production - set up monitoring, logging, and alerting, and take part in on-call rotations.
Challenge technical decisions to drive the platform toward high availability, disaster recovery, and a high uptime SLA.
Define and track SLIs/SLOs, and reduce operational toil by building automation, CI/CD pipelines, and infrastructure-as-code.
Triage production incidents, drive rapid mitigation, lead root-cause analysis, and write postmortems.
Reproduce and investigate complex production scenarios - performance issues, failure modes, edge cases - not just surface-level bugs.
Job requirements:
3-5+ years of backend software development experience, ideally with TypeScript/Node.js.
Strong backend architecture skills, with experience designing for scale and fault tolerance.
Deep technical understanding of Linux/systems fundamentals
SRE experience or strong SRE practices - observability, incident management, SLO-driven work.
Experience operating and monitoring cloud services in production (AWS or similar); familiarity with tools like Cloudwatch, Elastic, Prometheus, InfluxDB, and Grafana.
Experience with RDBMS and NoSQL systems like MariaDB, Mongo, Redis or others.
Experience with message broker software, like RabbitMQ, Redis, Amazon SQS.
Comfortable working independently, communicating blockers openly, and collaborating across teams.
Interview process:
Here’s what you can expect when you apply:
Initial interview - a chat with our People & Culture team to get to know you and your background.
Technical interview - a conversation with some of our engineers to dive into your skills and experience.
Take-home task - an assignment to work on in your own time, showcasing your problem-solving approach.
Task review - a follow-up interview where you’ll walk us through your solution and thought process.
Final interview - a discussion with our VP of Engineering, Engineering Manager, or both, to make sure we’re the right fit for each other.
Reasons to join LucidLink:
Tackle big challenges: You’ll have the chance to solve complex, high-stakes problems that redefine how teams collaborate globally. By starting with the Media & Entertainment industry and expanding into data-intensive sectors, you’ll gain deep insight into cutting-edge technologies and play a role in shaping the future of global workflows.
Values-led culture: Our values don’t just exist on paper - they guide every decision and interaction. You’ll thrive in an environment where integrity, innovation, and empathy are at the core of how we operate, empowering you to grow personally and professionally.
Hypergrowth journey: Joining a company with rapid growth means unparalleled opportunities for advancement, learning, and being part of an exciting journey toward unicorn status. You’ll experience the adrenaline of startup speed combined with the satisfaction of building something truly impactful.
Immediate impact: At LucidLink, your work will matter - immediately. You’ll be part of a tight-knit team of 230+ builders working at startup speed, where your ideas and actions will create tangible, exponential results that contribute to our collective success.
Comprehensive benefits: We believe in investing in our people. With flexible PTO, a competitive salary, stock options, and full health coverage, you’ll feel supported both professionally and personally while enjoying a strong work-life balance.
At LucidLink, AI fluency is a core professional competency. As part of our interview process, we'll explore how you currently use AI tools in your work and how you think about using them responsibly.
As published by LucidLink. Applications are handled on their site.
Skills this posting mentions
About LucidLink
On-demand cloud object storage optimazation software
All 7 openings at LucidLinkOne click, then it is written
Apply to LucidLink with a resume written for this role.
Queue Software Engineer, Webservice Platform and I read the posting, rewrite your resume against it, draft the cover letter, and score the fit. Then you press send, or press one button and I fill in LucidLink’s form for you.
01Drop your resume
A PDF or a LinkedIn URL. About a minute, once.
02I rank the openings
Every weekday morning, the live catalog scored against your history. Up to 20 worth your time, not two hundred links.
03Each one is written up
Resume rewritten for the posting, a cover letter, a fit score. Press send, or let me fill in the form.
- 25 sent a week, free
- No card
- Nothing sent until you say so
More roles at LucidLink
See all- 5 weeks ago
- 5 weeks ago
- 10 months ago
- 16 months ago
- 18 months ago
- 20 months ago
Similar roles elsewhere
See morePut this to work
Paste your career in once. Every application after that is written for you.
Drop a resume or a LinkedIn URL. I rank the live openings against it, rewrite the resume and write a cover letter for the best of them, and fill in the employer's form when you press the button. You read, you decide what goes out.
01Drop your resume
A PDF or a LinkedIn URL. About a minute, once.
02I rank the openings
Every weekday morning, the live catalog scored against your history. Up to 20 worth your time, not two hundred links.
03Each one is written up
Resume rewritten for the posting, a cover letter, a fit score. Press send, or let me fill in the form.
- New matches ranked and written before you are up.
- Every bullet stays inside what your history supports. Nothing invented.
- Queued, submitted, interviewing, offer: one screen, not a spreadsheet.
500 free credits on sign-up. No card. Nothing is sent until you say so.
Listed from the job board LucidLink publishes. Refolk is not the employer and does not handle their hiring. Applications go to LucidLink directly.