Senior Software Engineer - BQL Reliability Engineering
AI in this role
What You’ll Do:
As part of the BQL (Bloomberg Query Language) Reliability Engineering team, you will build software and platform capabilities that improve the reliability, resilience, and transparency of BQL and the services it depends on. You’ll work on engineering problems at significant scale, using software and automation to make reliability a built-in property of the platform rather than a purely operational concern.
You’ll be trusted to:
Design, build, and maintain software and self-service platform capabilities that enable engineering teams to understand, operate, and improve the reliability of BQL at scale.
Build tools and automated diagnostic capabilities that analyze telemetry and system behavior, helping engineers rapidly identify failures, regressions, and their root causes.
Develop software that improves incident detection and diagnosis, reducing Mean Time to Detect (MTTD) and Mean Time to Resolve (MTTR) for high-severity incidents.
Engineer observability capabilities that turn metrics, logs, traces, and other system signals into actionable insights across BQL’s distributed architecture.
Partner with BQL engineering teams on system design, instrumentation, SLIs and SLOs, ensuring reliability is built into services throughout the Software Development Lifecycle.
Improve platform resilience through engineering and experimentation, including load and stress testing, canary releases, controlled experiments, and failure testing.
Identify recurring operational problems and eliminate them through software, automation, and improvements to platform architecture.
You’ll Need to Have:
4+ years of experience in Software Engineering, Reliability Engineering, Platform Engineering, or a related technical role.
Experience working with an object-oriented programming language (C/C++, Python, Java, etc.)
Strong knowledge of Linux/UNIX systems and experience developing or operating distributed applications in production.
Demonstrated experience improving the performance, availability, resilience, or scalability of mid- to large-scale systems.
Experience with production software delivery, including deployment, release management, testing, and safely introducing changes into distributed systems.
Strong analytical and problem-solving skills, including the ability to use production data and system telemetry to understand complex system behavior.
BA, BS, MS, or PhD in Computer Science, Engineering, or a related technical field.
We’d Love to See Experience with:
Building developer platforms, reliability tooling, observability systems, or other infrastructure used by engineering teams.
Working with metrics, logs, distributed traces, time-series data, and other forms of production telemetry.
Defining and applying SLIs, SLOs, and other quantitative measures of system reliability.
Designing and executing load tests, stress tests, failure tests, canary releases, or other techniques for validating system resilience.
Applying statistical methods to understand system behavior and solve real-world engineering problems.
Querying and analyzing large-scale datasets in enterprise data environments.
Designing, executing, and analyzing A/B tests and other controlled experiments.
Operating in regulated or highly controlled environments.
Salary Range = 160,000 - 240,000 USD Annual + Benefits + Bonus
The referenced salary range is based on the Company's good faith belief at the time of posting. Actual compensation may vary based on factors such as geographic location, work experience, market conditions, education/training and skill level.
We offer one of the most comprehensive and generous benefits plans available and offer a range of total rewards that may include merit increases, incentive compensation (exempt roles only), paid holidays, paid time off, medical, dental, vision, short and long term disability benefits, 401(k) +match, life insurance, and various wellness programs, among others. The Company does not provide benefits directly to contingent workers/contractors and interns.
Discover what makes Bloomberg unique - watch our podcast series for an inside look at our culture, values, and the people behind our success.
How we rate this
Senior Software Engineer - BQL Reliability Engineering at Bloomberg rates 27 out of 100 for how much of the daily work is AI. That makes it Little AI (AI Level 1 of 4). The level is about AI in the job, not seniority.
Little AI. AI is not part of the work.
- ●●●● Builds AI80 to 100
- ●●●○ Works on AI60 to 79
- ●●○○ Uses AI40 to 59
- ●○○○ Little AI0 to 39
Levels come from how often the tools, models and workflows of the role are named in the posting itself. Open the description and count.
Get new software engineer jobs by email
One email a week with the new software engineer jobs, each rated for how much AI is in the work. No recruiter spam, unsubscribe in one click.
Free. One email a week. Unsubscribe in one click.
Similar roles
Software Engineering roles that involve little AI, at other companies.
What kind of AI work fits you?
Answer 12 practical questions in about three minutes. Get a simple profile, the work it points to, and live roles to explore next.
Find my next step