Create an account for powerful AI tools, award-winning courses, and access to our vibrant community.
Already have an account?
Join 250,000+ professionals and teams at Microsoft, Shopify, and even NASA. 🚀
Already have an account? Login
Find the best remote jobs. Answer a few questions and we'll deploy a powerful assistant to help you search, create alerts, and more.
1 What roles are you open to?
2 Experience level
3 Work style
Did you know? If memory is enabled, Writing.io can remember your job search preferences and help you to improve your resume, craft customized outreach and more.
Category
Leads technical execution of data infrastructure and pipelines for government services, ensuring data reliability and usability at scale across multiple systems.
Code for America believes government can work for the people, by the people, in the new digital age, and that government at all levels can and should work well for all people. For more than a decade, we’ve worked to show that with the mindful use of technology, we can break down barriers, meet community needs, and find real solutions.
Our employees build and transform government and community tools and services, making them so good they inspire change. We merge the best parts of technology, nonprofit, and government to help support the people who need it most.
With a focus on transparency and fairness, and deep empathy for partners in government and community organizations and the people that our partners serve, we’re building a movement of motivated change agents driven by meaningful results and lasting impact.
At Code for America, you contribute to exciting work while learning and developing in a supportive and flexible environment. Our compensation and benefits are holistic and thoughtfully curated to represent our employees and our mission. Help us drive real generational change that lasts.
Code for America is looking for a talented Staff Data Engineer who will lead the technical execution of our data infrastructure behind high-impact government services. In this role you will make data reliable and usable at scale — improving how our teams collect, move, model, and analyze the data that tells us whether the services we build actually work for the people who need them most.
The pace of change in the policy and government landscape is accelerating at the same time that technology possibilities are shifting rapidly. Together, these shifts create new windows of opportunity to develop technical solutions that make critical public services more effective for the people who need them. Staff data engineers keep delivery moving through those moments, supporting staff and holding high delivery standards so that data are trustworthy, and reachable by the people who need it.
Code for America’s Staff Data Engineers build and operate the infrastructure that moves data through the products we ship, between our systems and our partners’ systems, and into the analysis that drives meaningful improvements in outcomes across the tax, criminal justice, and safety net delivery areas. Staff Data Engineers lead the technical execution of data infrastructure projects and solutions while guiding their teams’ technical direction, setting a high standard for technical judgment, and enabling other data scientists and engineers through collaboration and mentorship.
This role will report to a Data Science Manager and is expected to travel no more than 10 % of the time.
Code for America is based in California and can employ those who reside full-time within the United States. This is a remote position.
Data Infrastructure & Pipeline Engineering
Technical Leadership & Enablement
Other duties as assigned
Code for America’s salary bands are transparent as a part of our commitment to transparency and fairness. As part of our hiring practices, we aim to target the midpoint of the 2nd quartile of the range for all new hires.
Offer targets vary based on market / geographic location. The offer targets for this role range from $128,945 to $157,850,annually. Your Recruiter will discuss compensation in more detail & answer any related questions, during the initial Recruiter phone screen.
This role includes a comprehensive benefits package including the following:
From the time you are contacted by a Recruiter, you can expect a 1-2 month, or longer, hiring process. The specific hiring process and timeline goals for this role will be covered during the initial Recruiter phone screen. Code for America’s standard interview process details are also available on our careers page.
Code for America is an equal opportunity employer. Applicants will not be discriminated against because of race, color, creed, sex, sexual orientation, gender identity or expression, age, religion, national origin, citizenship status, disability, ancestry, marital status, veteran status, medical condition or any protected category prohibited by local, state or federal laws.
This position is covered by a Collective Bargaining Agreement between Code for America and Code for America Workers United, affiliated with OPEIU, Local 1010. The agreement was ratified on January 13, 2026, and is currently in effect.
#LI-MD1
#LI-Remote
Lead ML engineer develops and validates AI biomarkers for cancer care, owning models from conception through regulatory submission and production deployment.
About Us: Artera is an artificial intelligence company dedicated to transforming cancer care. We’ve developed foundation models that analyze clinical and pathology data, generating actionable insights that guide therapy selection and improve outcomes for cancer patients. By continuously improving these models, we aim to uncover the biological mechanisms driving cancer progression.
We’re looking for an experienced machine learning engineer to own AI biomarker development end to end — from problem framing with clinical and biostatistics partners, through model development and validation, to regulatory submission and production deployment. Beyond owning a biomarker program, you’ll take on the hardest cross-cutting problems in our field: robustness across scanners and sites, mechanistic interpretability of model decisions, and the next generation of our pathology foundation models.
Lead the technical effort and define the strategic vision for patient-facing products, in partnership with product, biostatistics, clinical development, and regulatory/quality.
Design and build AI-based biomarkers on multimodal data — including whole-slide images, clinical variables, and molecular data — to predict patient outcomes, treatment benefit, and molecular traits.
Advance our core self-supervised foundation models and the downstream architectures built on them (multiple-instance learning, time-to-event / hazard models, segmentation and classification components), with generalization as a first-order objective.
Own score reproducibility across scanners, institutions, staining protocols, and patient populations.
Develop and integrate mechanistic interpretability methods to explain model decisions, build clinician trust, and drive actionable model improvements.
Architect tools and processes that streamline the end-to-end model development lifecycle — from prototyping through production deployment and monitoring — ensuring efficiency, reproducibility, regulatory compliance, and scale.
Author and defend regulatory and quality documentation, and represent AI in design and development reviews.
Plan and manage delivery: break multi-quarter programs into milestones, manage dependencies across AI, platform, biostatistics, and clinical teams, surface risk early, and hold submission and launch dates.
Publish in peer-reviewed journals and present at clinical and ML venues; support external academic and industry collaborations.
Mentor and coach machine-learning scientists and engineers, fostering their technical growth and collaboration skills, and raise the bar on scientific rigor, code quality, and written communication across the team.
5+ years of industry experience building deep learning systems in PyTorch (or TensorFlow).
2+ years of experience as a technical lead, launching and monitoring machine-learning products in production environments.
Demonstrated depth in oncology and biomarker development: familiarity with cancer biology and treatment pathways, clinical endpoints, risk stratification, and what makes a biomarker clinically actionable.
Demonstrated project management ability — scoping, sequencing, and managing dependencies and risk across multiple teams on dated deliverables.
Proven ability to communicate complex ML concepts effectively to cross-functional, non-ML collaborators.
Experience mentoring or managing ML scientists and engineers.
Experience building ML on complex clinical data — medical imaging, multi-omics, or longitudinal patient records — including weakly supervised learning and handling variation across sites, devices, and protocols.
Experience developing ML in a regulated environment — FDA 510(k)/De Novo, CE/UKCA, SaMD, design controls, or CLIA/LDT validation.
Experience with self-supervised representation learning (e.g., DINOv3) and adapting medical foundation models to downstream clinical tasks.
Experience with data from randomized controlled trials and multi-institutional clinical cohorts.
Peer-reviewed publications and conference presentations; history of external academic or industry collaborations.
Experience with cloud-scale training and workflow orchestration (e.g., Flyte / Union, Kubernetes, AWS), experiment tracking, and reproducible ML pipelines.
$180,000 - $240,000 a year
In addition to base salary, equity is a core component of our compensation. We also offer 401k matching, unlimited paid time off (PTO), and more.
The base salary is competitive and commensurate with experience, qualifications, and other factors to be discussed during the interview process.
Equal Employee Opportunity:At Artera, we value bringing together individuals from diverse backgrounds to develop new andinnovative solutions for patients and physicians. As an equal opportunity employer, we do notdiscriminate on the basis of race, color, religion, national origin, age, sex (including pregnancy),physical or mental disability, medical condition, genetic information gender identity orexpression, sexual orientation, marital status, protected veteran status, or any other legallyprotected characteristic.
Lead technical architecture and product development for code infrastructure platform serving AI agents and engineering teams at scale.
Our mission is to bring clarity and control to the world’s most complex codebases. AI is accelerating code creation, but the infrastructure to understand, oversee, and evolve that code hasn’t kept pace. Sourcegraph gives engineering organizations full visibility across their systems, precise context for their agents, and the ability to execute coordinated code changes at scale. As agentic development becomes the dominant engineering paradigm, we provide the context layer teams need to take control of their codebase.
With Code Search, Deep Search, MCP, and Agentic Batch Changes, we deliver on that mission today - giving engineering teams and their AI tools the cross-repo context to navigate massive codebases with confidence, and the ability to make changes across hundreds of repositories at once.
Companies like Stripe, Reddit, and Leidos rely on Sourcegraph to ship faster and with higher quality. We’re backed by a16z, Sequoia, and Redpoint, and proud to operate as a globally distributed team that values high agency, direct communication, and customer love.
If you want to build the infrastructure that lets every engineering team - and every agent they deploy - operate on their codebase with confidence, join us.
🌎 While we hire almost anywhere in the world, we have a preference for someone to reside in the following locations for this role. However, if you feel qualified, we welcome you to apply regardless of location. No matter what, working hours must overlap with CEST for at least 20 hours/week.
Preferred locations:
The Code Plane team owns the Sourcegraph products that help developers - and the agents working on their behalf - take action on code, at scale, across the world’s largest codebases. Think “data plane” and “control plane” for enterprise code: the surfaces where intent becomes code changes across hundreds (or more!) of repositories.
This is a team that builds products that use AI and products for AI. The roadmap, the architecture, and the day-to-day technical decisions all hinge on a solid experience with and a clear-eyed view of what AI models and agents are good at, where they fail, and how to design products around them. A strong handle on and opinion of AI, formed from real-world, hands-on use, not just observation, is a hard requirement for this role. If you’re excited to ship for the AI agent ecosystem rather than simply watch from the sidelines, you’ll find a lot to love here.
We’re hiring you to be the technical leader around whom this team rallies. Our product surfaces are evolving drastically as agents reshape how software is built, and we are adjusting our engineering teams to match. It is an exciting time, and a rare chance to be a strong leader shaping dev tools for the agentic age of coding. Code Plane owns high-stakes, fast-moving products at the center of that shift and needs a tech lead who will set technical and product direction, drive the roadmap from issue to shipped, and keep the team unblocked and moving. You are a generalist by choice. You go where the problem is, backend, AI agent, or frontend, and you are who people come to when it crosses a boundary.
The team’s surface areas include:
The features your team ships are used directly by developers and by agents working on their behalf, and you’ll partner closely with Product, Design, Customer Engineering, and adjacent engineering teams to turn customer pain into shipped product.
📅 Within one month, you will…
📅 Within three months, you will…
📅 Within six months , you will…
You are a technical leader and software engineer whom people want to follow. You set technical direction, make the hard architecture and tradeoff calls, and keep the team unblocked and moving. You think in terms of customers and outcomes, not tickets. You bring a technical perspective to what the team should and should not take on, and you make the quality bar real: review culture, testing standards, and release practices. You’re scoping with Product and Design and making calls on what to cut so the team can ship. You’re tackling our hardest technical problems hands-on, and also guiding your team to grow.
Agents are one of the most important parts of this job. You have shipped for the agent ecosystem: something that agents or other programs call, with evals behind it, so you know whether a prompt, tool, or model change made things better or worse. You can say precisely where agents fail, because you have encountered it yourself and measured it.
You can take a position, argue it persuasively, and bring people with you, including the ones who started out disagreeing, and you say plainly what would change your mind. You contribute to a collaborative, respectful, async-first culture, and you keep stakeholders informed without being asked.
Strongly preferred
📊 This job is an IC5. You can read more about our job leveling philosophy in our Handbook.
💸 We pay above-market salaries because we want to hire exceptional people who can focus on building great products, not worrying about paying bills. As an open and transparent company, our compensation philosophy and pay bands are visible to every Sourcegraph teammate, and we strive to make our approach equitable, explainable, and competitive.
Your base salary is determined by the IC5 pay band for your location zone (1-4). Our pay bands are informed by market data and designed to ensure competitive compensation wherever you live. During the recruiting process, we’ll discuss the range applicable to you based on job level, relevant skills, experience, qualifications, and location zone.
💰 The starting salary for the IC5 pay band in each zone is:
📈 In addition to competitive cash compensation, we offer meaningful equity (because when Sourcegraph succeeds, we want you to succeed, too) and generous perks & benefits.
Below is the interview process you can expect for this role (you can read more about the types of interviews in our Handbook). It may look like a lot of steps, but rest assured that we move quickly and the steps are designed to help you get the information needed to determine if we’re the right fit for you… Interviewing is a two-way street, after all!
We expect the interview process to take 4.75 hours in total.
👋 Introduction Stage - we have initial conversations to get to know you better…
🧑💻 Team Interview Stage - we then delve into your experience in more depth and introduce you to members of the team, including cross-functional partners…
🎉 Final Interview Stage- we move you to our final round, where you gain a better understanding of our business and values holistically…
Please note - you are welcome to request additional conversations with anyone you would like to meet, but didn’t get to meet during the interview process.
You can learn more about what it is like to work at Sourcegraph by reading our handbook.
We are an ambitious team who are collectively working hard to build the most influential company in the world. You can read more about our culture, competitive compensation and benefits here.
Sourcegraph is an equal opportunity workplace; we welcome people from all backgrounds.
Sourcegraph participates in E-Verify for U.S. Employees.
Leads technical execution of data infrastructure projects, ensuring data reliability and usability across government services and systems.
Code for America believes government can work for the people, by the people, in the new digital age, and that government at all levels can and should work well for all people. For more than a decade, we’ve worked to show that with the mindful use of technology, we can break down barriers, meet community needs, and find real solutions.
Our employees build and transform government and community tools and services, making them so good they inspire change. We merge the best parts of technology, nonprofit, and government to help support the people who need it most.
With a focus on transparency and fairness, and deep empathy for partners in government and community organizations and the people that our partners serve, we’re building a movement of motivated change agents driven by meaningful results and lasting impact.
At Code for America, you contribute to exciting work while learning and developing in a supportive and flexible environment. Our compensation and benefits are holistic and thoughtfully curated to represent our employees and our mission. Help us drive real generational change that lasts.
Code for America is looking for a talented Staff Data Engineer who will lead the technical execution of our data infrastructure behind high-impact government services. In this role you will make data reliable and usable at scale — improving how our teams collect, move, model, and analyze the data that tells us whether the services we build actually work for the people who need them most.
The pace of change in the policy and government landscape is accelerating at the same time that technology possibilities are shifting rapidly. Together, these shifts create new windows of opportunity to develop technical solutions that make critical public services more effective for the people who need them. Staff data engineers keep delivery moving through those moments, supporting staff and holding high delivery standards so that data are trustworthy, and reachable by the people who need it.
Code for America’s Staff Data Engineers build and operate the infrastructure that moves data through the products we ship, between our systems and our partners’ systems, and into the analysis that drives meaningful improvements in outcomes across the tax, criminal justice, and safety net delivery areas. Staff Data Engineers lead the technical execution of data infrastructure projects and solutions while guiding their teams’ technical direction, setting a high standard for technical judgment, and enabling other data scientists and engineers through collaboration and mentorship.
This role will report to a Data Science Manager and is expected to travel no more than 10 % of the time.
Code for America is based in California and can employ those who reside full-time within the United States. This is a remote position.
Data Infrastructure & Pipeline Engineering
Technical Leadership & Enablement
Other duties as assigned
Code for America’s salary bands are transparent as a part of our commitment to transparency and fairness. As part of our hiring practices, we aim to target the midpoint of the 2nd quartile of the range for all new hires.
Offer targets vary based on market / geographic location. The offer targets for this role range from $128,945 to $157,850,annually. Your Recruiter will discuss compensation in more detail & answer any related questions, during the initial Recruiter phone screen.
This role includes a comprehensive benefits package including the following:
From the time you are contacted by a Recruiter, you can expect a 1-2 month, or longer, hiring process. The specific hiring process and timeline goals for this role will be covered during the initial Recruiter phone screen. Code for America’s standard interview process details are also available on our careers page.
Code for America is an equal opportunity employer. Applicants will not be discriminated against because of race, color, creed, sex, sexual orientation, gender identity or expression, age, religion, national origin, citizenship status, disability, ancestry, marital status, veteran status, medical condition or any protected category prohibited by local, state or federal laws.
This position is covered by a Collective Bargaining Agreement between Code for America and Code for America Workers United, affiliated with OPEIU, Local 1010. The agreement was ratified on January 13, 2026, and is currently in effect.
#LI-MD1
#LI-Remote
Tech lead architecting code infrastructure products that enable developers and AI agents to navigate and modify large codebases at scale.
Our mission is to bring clarity and control to the world’s most complex codebases. AI is accelerating code creation, but the infrastructure to understand, oversee, and evolve that code hasn’t kept pace. Sourcegraph gives engineering organizations full visibility across their systems, precise context for their agents, and the ability to execute coordinated code changes at scale. As agentic development becomes the dominant engineering paradigm, we provide the context layer teams need to take control of their codebase.
With Code Search, Deep Search, MCP, and Agentic Batch Changes, we deliver on that mission today - giving engineering teams and their AI tools the cross-repo context to navigate massive codebases with confidence, and the ability to make changes across hundreds of repositories at once.
Companies like Stripe, Reddit, and Leidos rely on Sourcegraph to ship faster and with higher quality. We’re backed by a16z, Sequoia, and Redpoint, and proud to operate as a globally distributed team that values high agency, direct communication, and customer love.
If you want to build the infrastructure that lets every engineering team - and every agent they deploy - operate on their codebase with confidence, join us.
🌎 While we hire almost anywhere in the world, we have a preference for someone to reside in the following locations for this role. However, if you feel qualified, we welcome you to apply regardless of location. No matter what, working hours must overlap with CEST for at least 20 hours/week.
Preferred locations:
The Code Plane team owns the Sourcegraph products that help developers - and the agents working on their behalf - take action on code, at scale, across the world’s largest codebases. Think “data plane” and “control plane” for enterprise code: the surfaces where intent becomes code changes across hundreds (or more!) of repositories.
This is a team that builds products that use AI and products for AI. The roadmap, the architecture, and the day-to-day technical decisions all hinge on a solid experience with and a clear-eyed view of what AI models and agents are good at, where they fail, and how to design products around them. A strong handle on and opinion of AI, formed from real-world, hands-on use, not just observation, is a hard requirement for this role. If you’re excited to ship for the AI agent ecosystem rather than simply watch from the sidelines, you’ll find a lot to love here.
We’re hiring you to be the technical leader around whom this team rallies. Our product surfaces are evolving drastically as agents reshape how software is built, and we are adjusting our engineering teams to match. It is an exciting time, and a rare chance to be a strong leader shaping dev tools for the agentic age of coding. Code Plane owns high-stakes, fast-moving products at the center of that shift and needs a tech lead who will set technical and product direction, drive the roadmap from issue to shipped, and keep the team unblocked and moving. You are a generalist by choice. You go where the problem is, backend, AI agent, or frontend, and you are who people come to when it crosses a boundary.
The team’s surface areas include:
The features your team ships are used directly by developers and by agents working on their behalf, and you’ll partner closely with Product, Design, Customer Engineering, and adjacent engineering teams to turn customer pain into shipped product.
📅 Within one month, you will…
📅 Within three months, you will…
📅 Within six months , you will…
You are a technical leader and software engineer whom people want to follow. You set technical direction, make the hard architecture and tradeoff calls, and keep the team unblocked and moving. You think in terms of customers and outcomes, not tickets. You bring a technical perspective to what the team should and should not take on, and you make the quality bar real: review culture, testing standards, and release practices. You’re scoping with Product and Design and making calls on what to cut so the team can ship. You’re tackling our hardest technical problems hands-on, and also guiding your team to grow.
Agents are one of the most important parts of this job. You have shipped for the agent ecosystem: something that agents or other programs call, with evals behind it, so you know whether a prompt, tool, or model change made things better or worse. You can say precisely where agents fail, because you have encountered it yourself and measured it.
You can take a position, argue it persuasively, and bring people with you, including the ones who started out disagreeing, and you say plainly what would change your mind. You contribute to a collaborative, respectful, async-first culture, and you keep stakeholders informed without being asked.
Strongly preferred
📊 This job is an IC5. You can read more about our job leveling philosophy in our Handbook.
💸 We pay above-market salaries because we want to hire exceptional people who can focus on building great products, not worrying about paying bills. As an open and transparent company, our compensation philosophy and pay bands are visible to every Sourcegraph teammate, and we strive to make our approach equitable, explainable, and competitive.
Your base salary is determined by the IC5 pay band for your location zone (1-4). Our pay bands are informed by market data and designed to ensure competitive compensation wherever you live. During the recruiting process, we’ll discuss the range applicable to you based on job level, relevant skills, experience, qualifications, and location zone.
💰 The starting salary for the IC5 pay band in each zone is:
📈 In addition to competitive cash compensation, we offer meaningful equity (because when Sourcegraph succeeds, we want you to succeed, too) and generous perks & benefits.
Below is the interview process you can expect for this role (you can read more about the types of interviews in our Handbook). It may look like a lot of steps, but rest assured that we move quickly and the steps are designed to help you get the information needed to determine if we’re the right fit for you… Interviewing is a two-way street, after all!
We expect the interview process to take 4.75 hours in total.
👋 Introduction Stage - we have initial conversations to get to know you better…
🧑💻 Team Interview Stage - we then delve into your experience in more depth and introduce you to members of the team, including cross-functional partners…
🎉 Final Interview Stage- we move you to our final round, where you gain a better understanding of our business and values holistically…
Please note - you are welcome to request additional conversations with anyone you would like to meet, but didn’t get to meet during the interview process.
You can learn more about what it is like to work at Sourcegraph by reading our handbook.
We are an ambitious team who are collectively working hard to build the most influential company in the world. You can read more about our culture, competitive compensation and benefits here.
Sourcegraph is an equal opportunity workplace; we welcome people from all backgrounds.
Sourcegraph participates in E-Verify for U.S. Employees.
Designs and develops flight-critical autonomy algorithms using model-based design tools, Simulink, and DO-178C compliant processes for aerospace systems.
About Merlin:
Merlin (NASDAQ: MRLN) is a publicly traded aerospace and defense company building a non-human pilot to deliver full-stack autonomy for any aircraft from takeoff to touchdown. The Merlin Pilot autonomy system powers a growing range of aircraft and mission profiles and has been proven through hundreds of autonomous flights from Merlin’s global flight test facilities, including Kerikeri, New Zealand; Quonset Point, Rhode Island; and soon, Bedford, Massachusetts. Headquartered in Boston, Merlin is expanding its organization to accelerate the development and deployment of its autonomy platform, helping customers solve some of aviation’s most pressing challenges, from pilot shortages to improving flight safety. Backed by some of the world’s leading investors prior to its public listing, Merlin continues to advance the certification and commercialization of autonomous flight across commercial and defense aviation.
We are seeking a Staff Software Engineer to design, implement, test, and certify flight-critical autonomy algorithms for next-generation aerospace systems. In this role, you will develop model-based flight software using MathWorks tools and support the full lifecycle of DO-178C compliant development.
$200,000 - $265,000 a year
The compensation range provided is reflective of base salary only, and is a good-faith estimate based on a wide range of factors. The actual offer will be determined by a variety of factors including the candidate’s qualifications, skills, experience, education, and training. Highly competitive equity grants are considered part of Merlin’s total compensation package. Additionally, Merlin offers top-tier benefits for full-time employees.
This position is based on-site at Merlin HQ in Boston, MA.
Once you’re here, you’ll enjoy a variety of on-site perks designed to make your workday enjoyable and convenient. These include catered lunches featuring a rotating menu of delicious options, an assortment of snacks to keep you fueled throughout the day, and a selection of beverages, including coffee, tea, and other drinks, to keep you refreshed.
Our goal is to create an environment where you can thrive both professionally and personally
Merlin Labs offers an innovative, entrepreneurial, and team-focused startup environment. We also offer a top-notch benefits package (health, dental, life, unlimited vacation, and 401k with match) and work/life integration. Being part of the Merlin team allows you to become part of a small team that supports professional development while working together to achieve our mission.
Merlin Labs is an equal opportunity employer and values diversity. We do not discriminate on the basis of race, religion, color, national origin, genetic information, sex (including pregnancy), gender, gender identity and expression, sexual orientation, age, marital status, military service or obligation or disability status, or any other characteristic protected by law. All job offers are contingent upon the candidate passing background and reference checks.
At this time, we are unable to provide visa sponsorship or consider candidates who require visa transfers. Applicants must be authorized to work in the United States without the need for visa sponsorship now or in the future.
In compliance with federal law, all persons hired will be required to verify identity and eligibility to work in the United States and to complete the required employment eligibility verification form upon hire.
If you require reasonable accommodation in completing an application, interviewing, completing any pre-employment testing, or otherwise participating in the employee selection process, please direct your inquiries to: [email protected]
Merlin Labs does not accept unsolicited resumes from any source other than directly from candidates.
Tech Lead directs engineering strategy and architecture for Sourcegraph's Code Plane products that enable developers and AI agents to navigate and modify large codebases at scale.
Our mission is to bring clarity and control to the world’s most complex codebases. AI is accelerating code creation, but the infrastructure to understand, oversee, and evolve that code hasn’t kept pace. Sourcegraph gives engineering organizations full visibility across their systems, precise context for their agents, and the ability to execute coordinated code changes at scale. As agentic development becomes the dominant engineering paradigm, we provide the context layer teams need to take control of their codebase.
With Code Search, Deep Search, MCP, and Agentic Batch Changes, we deliver on that mission today - giving engineering teams and their AI tools the cross-repo context to navigate massive codebases with confidence, and the ability to make changes across hundreds of repositories at once.
Companies like Stripe, Reddit, and Leidos rely on Sourcegraph to ship faster and with higher quality. We’re backed by a16z, Sequoia, and Redpoint, and proud to operate as a globally distributed team that values high agency, direct communication, and customer love.
If you want to build the infrastructure that lets every engineering team - and every agent they deploy - operate on their codebase with confidence, join us.
🌎 While we hire almost anywhere in the world, we have a preference for someone to reside in the following locations for this role. However, if you feel qualified, we welcome you to apply regardless of location. No matter what, working hours must overlap with CEST for at least 20 hours/week.
Preferred locations:
The Code Plane team owns the Sourcegraph products that help developers - and the agents working on their behalf - take action on code, at scale, across the world’s largest codebases. Think “data plane” and “control plane” for enterprise code: the surfaces where intent becomes code changes across hundreds (or more!) of repositories.
This is a team that builds products that use AI and products for AI. The roadmap, the architecture, and the day-to-day technical decisions all hinge on a solid experience with and a clear-eyed view of what AI models and agents are good at, where they fail, and how to design products around them. A strong handle on and opinion of AI, formed from real-world, hands-on use, not just observation, is a hard requirement for this role. If you’re excited to ship for the AI agent ecosystem rather than simply watch from the sidelines, you’ll find a lot to love here.
We’re hiring you to be the technical leader around whom this team rallies. Our product surfaces are evolving drastically as agents reshape how software is built, and we are adjusting our engineering teams to match. It is an exciting time, and a rare chance to be a strong leader shaping dev tools for the agentic age of coding. Code Plane owns high-stakes, fast-moving products at the center of that shift and needs a tech lead who will set technical and product direction, drive the roadmap from issue to shipped, and keep the team unblocked and moving. You are a generalist by choice. You go where the problem is, backend, AI agent, or frontend, and you are who people come to when it crosses a boundary.
The team’s surface areas include:
The features your team ships are used directly by developers and by agents working on their behalf, and you’ll partner closely with Product, Design, Customer Engineering, and adjacent engineering teams to turn customer pain into shipped product.
📅 Within one month, you will…
📅 Within three months, you will…
📅 Within six months , you will…
You are a technical leader and software engineer whom people want to follow. You set technical direction, make the hard architecture and tradeoff calls, and keep the team unblocked and moving. You think in terms of customers and outcomes, not tickets. You bring a technical perspective to what the team should and should not take on, and you make the quality bar real: review culture, testing standards, and release practices. You’re scoping with Product and Design and making calls on what to cut so the team can ship. You’re tackling our hardest technical problems hands-on, and also guiding your team to grow.
Agents are one of the most important parts of this job. You have shipped for the agent ecosystem: something that agents or other programs call, with evals behind it, so you know whether a prompt, tool, or model change made things better or worse. You can say precisely where agents fail, because you have encountered it yourself and measured it.
You can take a position, argue it persuasively, and bring people with you, including the ones who started out disagreeing, and you say plainly what would change your mind. You contribute to a collaborative, respectful, async-first culture, and you keep stakeholders informed without being asked.
Strongly preferred
📊 This job is an IC5. You can read more about our job leveling philosophy in our Handbook.
💸 We pay above-market salaries because we want to hire exceptional people who can focus on building great products, not worrying about paying bills. As an open and transparent company, our compensation philosophy and pay bands are visible to every Sourcegraph teammate, and we strive to make our approach equitable, explainable, and competitive.
Your base salary is determined by the IC5 pay band for your location zone (1-4). Our pay bands are informed by market data and designed to ensure competitive compensation wherever you live. During the recruiting process, we’ll discuss the range applicable to you based on job level, relevant skills, experience, qualifications, and location zone.
💰 The starting salary for the IC5 pay band in each zone is:
📈 In addition to competitive cash compensation, we offer meaningful equity (because when Sourcegraph succeeds, we want you to succeed, too) and generous perks & benefits.
Below is the interview process you can expect for this role (you can read more about the types of interviews in our Handbook). It may look like a lot of steps, but rest assured that we move quickly and the steps are designed to help you get the information needed to determine if we’re the right fit for you… Interviewing is a two-way street, after all!
We expect the interview process to take 4.75 hours in total.
👋 Introduction Stage - we have initial conversations to get to know you better…
🧑💻 Team Interview Stage - we then delve into your experience in more depth and introduce you to members of the team, including cross-functional partners…
🎉 Final Interview Stage- we move you to our final round, where you gain a better understanding of our business and values holistically…
Please note - you are welcome to request additional conversations with anyone you would like to meet, but didn’t get to meet during the interview process.
You can learn more about what it is like to work at Sourcegraph by reading our handbook.
We are an ambitious team who are collectively working hard to build the most influential company in the world. You can read more about our culture, competitive compensation and benefits here.
Sourcegraph is an equal opportunity workplace; we welcome people from all backgrounds.
Sourcegraph participates in E-Verify for U.S. Employees.
Leads DevOps and SRE teams to build reliable, scalable, and secure AWS-based production infrastructure while spending 20-30% time on hands-on architecture and tooling work.
Headquarters: Remote, United States
About this Position
Are you passionate about building the reliability, automation, and security foundations that let engineering teams move fast with confidence? At Legion, we are seeking a Director of Engineering, DevOps & SRE to lead the teams responsible for the availability, scalability, and security of our production environment. Our production infrastructure runs on AWS, leveraging services such as EKS, RDS, and a broad set of AWS-native technologies. You will partner closely with engineering and IT to build resilient systems, drive operational excellence, and ensure our platform meets the highest standards of security and compliance.
This is a hands-on leadership role where you'll spend ~20-30% of your time contributing directly to architecture, tooling, and incident response, and the rest driving vision, roadmap, and cross-team execution.
Responsibilities
Required Qualifications
Preferred Qualifications
COMPENSATION & BENEFITS
Salary Range: Base Salary Range $220,000 - $265,000 + Bonus + Stock Equity
At Legion, we offer competitive compensation and benefits packages to all employees. As a fully remote employer, pay for positions is determined using local, national, and industry-specific survey data.
Our posted salary range is done so in good faith based on national data and may be refined for a candidate's region/town/cost of living. We strive to make competitive offers that allow employees room for future growth. Salaries will be based on the applicant’s location, level of experience, education, and specialized knowledge and skills. Additionally, we consider the external market rate, the amount we have budgeted internally, and the internal equity for the same position within the company.
Benefits include, but are not limited to:
ABOUT LEGION
Join Legion's mission to turn hourly jobs into good jobs. We're a remote, mission-driven team seeking exceptional talent to propel this vision. Embrace a culture that's collaborative, fast-paced, and entrepreneurial. With us, you'll grow your skills, work closely with experienced executives, and contribute significantly to our mission.
Legion Technologies delivers the industry’s most innovative workforce management platform. It enables businesses to maximize labor efficiency and employee engagement simultaneously. The award-winning, AI-driven Legion WFM platform is intelligent, automated, and employee-centric. It’s proven to deliver 13x ROI through schedule optimization, reduced attrition, increased productivity, and increased operational efficiency. Legion delivers cutting-edge technology in an easy-to-use platform and mobile app that employees love.
If you're ready to make an impact and grow your career, Legion is where you belong. Join us in making hourly work rewarding and fulfilling.
BACKGROUND AND OPPORTUNITY
There are almost 75 million hourly workers in the United States, representing more than half of the entire workforce. Historically, managing hourly employees has been difficult due to high attrition (average of 60%) and high replacement costs (average of $3,200 per employee in retail). The ongoing labor shortage and competition from the gig economy make it more difficult to attract and retain hourly employees. The top reasons hourly employees leave their jobs are a lack of schedule empowerment, poor communication with employers, and an inability to get paid early. Gen Z and the millennial workforce demand gig-like flexibility, modern technology, and compelling work options. Legion’s mission is to turn hourly jobs into good jobs, serving the hourly workers who make up the majority of the US workforce. We believe in empowering employees and helping employers be efficient and innovative by enabling intelligent automation powered by Legion’s Workforce Management platform to optimize labor efficiency and enhance the employee experience simultaneously. Legion WFM was built for the cloud with AI at the core and designed to handle the complexity of modern businesses and meet the needs of today’s hourly employees. Our team is comprised of dedicated individuals from all backgrounds and experiences, globally distributed across all time zones.
For more information, visit https://legion.co
EQUAL EMPLOYMENT OPPORTUNITY
Legion Technologies is proud to be an equal-opportunity employer and is committed to maintaining a diverse and inclusive work environment. All qualified applicants will be considered for employment without regard to race, color, religion, sex, age, disability, marital status, familial status, sexual orientation, pregnancy, genetic information, gender identity, gender expression, national origin, ancestry, citizenship status, veteran status, and any other legally protected status under federal, state, or local anti-discrimination laws.
DISABILITY ACCOMMODATION
For individuals with disabilities who need additional assistance at any point in the application and interview process, please email recruiting@legion.co
We have noticed a rise in recruiting impersonations across the industry, where scammers attempt to access candidates' personal and financial information through fake interviews and offers. All Legion recruiting email communications will always come from the @legion.co domain. Any outreach claiming to be from Legion via other sources should be ignored. If you are uncertain whether you have been contacted by an official Legion employee, reach out to recruiting@legion.co
Legion is an equal opportunity employer. All applicants will be considered for employment without attention to race, religion, color, sex, sexual orientation, gender identity, age, national origin, veteran, disability status, or any other basis covered by appropriate law.
As a global employer, Legion determines pay for positions using local, national, and industry-specific survey data. We evaluate external equity and the cost of labor/prevailing wage index in the relative marketplace for jobs directly comparable to jobs within our company. Our posted salary range is based on national data and may be refined for a candidate's region/town/cost of living. For new hires, we strive to make competitive offers allowing the new employee room for future growth. Salaries will be based on the applicant’s location, level of experience, education, and specialized knowledge and skills. Additionally, we consider the external market rate, the amount we have budgeted internally, and internal equity within the company for the same position. An employee/candidate with a stronger skill set will receive higher pay.
This Job Applicant Privacy Policy (“Policy”) describes how Legion Technologies, Inc. (“Legion”, “we”, “us” and “our”) collects, uses, and discloses “personal information” as defined under California law from and about job applicants who are residents of California.
This Policy does not apply to our handling of data gathered about you in your role as a user of our consumer-facing services. When you interact with us as in that role, the Legion Privacy Policy applies.
Types of Personal Information We Handle
We collect, store, and use various types of personal information through the application and recruitment process. We collect such information either directly from you or (where applicable) from another person or entity, such as an employment agency or consultancy, background check provider, or other referral sources. This information includes:
How We Use Personal Information
We collect, use, share, and store personal information from job applicants for our and our service providers’ business and operational purposes in the recruitment process such as: processing your application, tracking your application through the recruitment process, contacting references with your authorization, conducting background checks you authorize, and making hiring decisions. We will also use job applicant information for internal analysis purposes to understand the applicants who apply and to improve our recruitment process. We may sometimes need to use applicant information for legal purposes, such as in connection with any challenges made to our hiring decisions.
With Whom We Share Personal Information
We will disclose job applicant personal information to the following types of entities or in the following circumstances (where applicable):
To apply: https://weworkremotely.com/remote-jobs/legion-director-of-production-engineering
Designs and builds HubSpot's observability platform for distributed systems and AI agents, setting architecture patterns and telemetry standards across hundreds of microservices.
POS-5690
The Observability team owns the internal platform that gives every HubSpot engineer real visibility into how their systems behave in production. We build and operate the distributed tracing, metrics, alerting, and logging infrastructure that spans hundreds of microservices, billions of daily events, and thousands of engineers who depend on that signal to ship reliably.
We are now investing in the next generation of this platform. As HubSpot deploys AI agents and ML-powered features across the product, the team is building the tracing and telemetry primitives that make it possible to understand, debug, and trust what those systems are doing in production. This is greenfield, technically interesting work at a scale few companies operate at and we are looking for a Principal Engineer to help lead it.
We are seeking a Principal Software Engineer to be the technical anchor for HubSpot’s Observability platform. This role sits at the intersection of large-scale distributed systems, developer platform design, and AI observability. A big part of this role is working horizontally across a large engineering org and setting patterns and standards that make it easier for teams to instrument, alert on, and reason about their services. You will also shape how we trace and understand our growing fleet of AI agents and ML systems in production: a technically distinct and increasingly critical problem.
HubSpot is scaling fast — more engineers, more microservices, more AI systems running in production — and the Observability platform is at an inflection point. The foundations are solid, but the next phase requires a different kind of investment: rethinking cardinality economics, making OpenTelemetry the default across a large polyglot org, and building an entirely new layer of AI tracing that doesn’t yet exist.
This Principal Engineer will have a direct line of sight from the architecture they design to the engineering outcomes we measure: incident response times, adoption rates, developer satisfaction, and the trust that product teams place in their production signal. The AI observability layer in particular is greenfield — there is no playbook to follow, which is exactly what makes this the right moment for the right person.
If you want to build the platform that helps thousands of engineers understand what their systems are doing — including systems powered by AI that are genuinely hard to see inside — this is the role.
Pay & Benefits
The cash compensation below includes base salary, on-target commission for employees in eligible roles, and annual bonus targets under HubSpot’s bonus plan for eligible roles. In addition to cash compensation, some roles are eligible to participate in HubSpot’s equity plan to receive restricted stock units (RSUs). Some roles may also be eligible for overtime pay. Individual compensation packages are tailored to your skills, experience, qualifications, and other job-related reasons.
This resource will help guide how we recommend thinking about the range you see. Learn more about HubSpot’s compensation philosophy.
Benefits are also an important piece of your total compensation package. Explore the benefits and perks HubSpot offers to help employees grow better.
At HubSpot, fair compensation practices aren’t just about checking off the box for legal compliance. It’s about living out our value of transparency with our employees, candidates, and community.
Annual Cash Compensation Range:
$313,800—$502,080 USD
At HubSpot, we value both flexibility and connection. Whether you’re a Remote employee or work from the Office, we want you to start your journey here by building strong connections with your team and peers. If you are joining our Engineering team, you will be required to attend a regional HubSpot office for in-person onboarding. If you join our broader Product team, you’ll also attend other in-person events, such as your Product Group Summit and other gatherings, to continue building on those connections.
If you require an accommodation due to travel limitations or other reasons, please inform your recruiter during the hiring process. We are committed to supporting candidates who may need alternative arrangements
About HubSpot
HubSpot (NYSE: HUBS) is an AI-powered customer platform with all the software, integrations, and resources customers need to connect marketing, sales, and service. HubSpot’s connected platform enables businesses to grow faster by focusing on what matters most: customers.
At HubSpot, bold is our baseline. Our employees around the globe move fast, stay customer-obsessed, and win together. Our culture is grounded in four commitments: Solve for the Customer, Be Bold, Learn Fast, Align, Adapt & Go!, and Deliver with HEART. These commitments shape how we work, lead, and grow.
We’re building a company where people can do their best work. We focus on brilliant work, not badge swipes. By combining clarity, ownership, and trust, we create space for big thinking and meaningful progress. And we know that when our employees grow, our customers do too.
Recognized globally for our award-winning culture by Comparably, Glassdoor, Fortune, and more, HubSpot is headquartered in Cambridge, MA, with employees and offices around the world.
Explore more:
If you need accommodations or assistance due to a disability, please reach out to us using this form.
Massachusetts Applicants: It is unlawful in Massachusetts to require or administer a lie detector test as a condition of employment or continued employment. An employer who violates this law shall be subject to criminal penalties and civil liability.
Germany Applicants: (m/f/d) - link to HubSpot’s Career Diversity page here.
India Applicants: link to HubSpot India’s equal opportunity policy here.
HubSpot may use AI to help screen or assess candidates, but all hiring decisions are always human. More information can be found here. By submitting your application, you agree that HubSpot may collect your personal data for recruiting, global organization planning, and related purposes. We may use CLEAR ID Verification during the hiring process to confirm your identity and help maintain a safe, secure, and trusted experience for all candidates. Refer to HubSpot’s Recruiting Privacy Notice for details on data processing and your rights.
Staff SRE architecting observability solutions, defining reliability standards, leading incident response, and automating infrastructure at scale.
Replit is the agentic software creation platform that enables anyone to build applications using natural language. With millions of users worldwide, Replit is democratizing software development by removing traditional barriers to application creation.
Join our Site Reliability Engineering (SRE) team and help ensure the reliability, scalability, and performance of Replit’s infrastructure that serves millions of developers worldwide. As a Staff Site Reliability Engineer, you will bridge the gap between development and operations, implementing automation and establishing best practices that enable our platform to scale efficiently while maintaining high availability.
We are seeking Staff SREs who are passionate about building and maintaining resilient systems at scale. Your mission will be to proactively find and analyze reliability problems across our stack, then design and implement software and systems to create step-function improvements. You will design robust observability solutions, lead incident response, automate operational tasks, and continuously improve our infrastructure’s reliability, all while mentoring and educating the broader engineering team to make reliability a core value at Replit.
Architect and Implement Observability: Design, build, and lead the implementation of comprehensive monitoring, logging, and tracing solutions. Create dashboards and metrics that provide real-time visibility into system health and performance, enabling proactive issue detection.
Define and Drive Reliability Standards: Work with product and engineering teams to define, implement, and track Service Level Objectives (SLOs) and Service Level Indicators (SLIs). Build systems to monitor and report on these metrics, holding teams accountable and ensuring we maintain high reliability standards while balancing innovation speed.
Lead Incident Management and Response: Act as a senior leader during high-impact incidents, guiding the team to rapid resolution. Conduct thorough, blameless post-mortems and drive the implementation of preventative measures. Develop and refine runbooks and build automation to reduce Mean Time To Recovery (MTTR).
Drive Automation and Infrastructure as Code: Architect, build, and improve automation to eliminate toil and operational work. Design and maintain CI/CD pipelines and infrastructure automation using tools like Terraform or Pulumi. Create self-healing systems that can automatically respond to common failure scenarios.
Optimize Performance on Kubernetes: Collaborate with core infrastructure and product teams to performance-tune and optimize our large-scale cloud deployments, with a deep focus on Kubernetes, Docker, and GCP. Identify and resolve performance bottlenecks, implement capacity planning strategies, and reduce latency across global regions.
Debug and Harden Distributed Systems: Dive deep into debugging extremely difficult technical problems across the stack. Use your findings to design and implement long-term fixes that make our systems and products more robust, operable, and easier to diagnose.
Provide Staff-Level Guidance: Review feature and system designs from across the company, acting as a key owner for the reliability, scalability, security, and operational integrity of those designs.
Educate and Mentor: Educate, mentor, and hold accountable the broader engineering team to improve the reliability of our systems, making reliability a core value of the Replit engineering culture.
Build and Integrate: Write high-quality, well-tested code in Python or Go to meet the needs of your customers, whether it’s building new internal tools or integrating with third-party vendors.
8-10 years of experience in Site Reliability Engineering or similar roles (e.g., DevOps, Systems Engineering, Infrastructure Engineering).
Strong programming skills in languages like Python or Go. You write high-quality, well-tested code.
Deep understanding of distributed systems. You’ve designed, built, scaled, and maintained production services and know how to compose a service-oriented architecture.
Deep experience with container orchestration platforms, specifically Kubernetes, and cloud-native technologies.
Proven track record of designing, implementing, and maintaining sophisticated monitoring and observability solutions (e.g., metrics, logging, tracing).
Strong incident management skills with extensive experience leading incident response for complex systems and demonstrated critical thinking under pressure.
Experience with infrastructure as code (e.g., Terraform, Pulumi) and configuration management tools.
Excellent written and verbal communication skills, with an ability to explain complex technical concepts clearly and simply and a bias toward open, transparent cultural practices.
Strong interpersonal skills, with experience working with and mentoring engineers from junior to principal levels.
A willingness to dive into understanding, debugging, and improving any layer of the stack.
You’re passionate about making software creation accessible and empowering the next generation of builders.
Deep experience with Google Cloud Platform (GCP) services and tools.
Expert-level knowledge of modern observability platforms (e.g., Prometheus, Grafana, Datadog, OpenTelemetry).
Experience designing and building reliable systems capable of handling high throughput and low latency.
Significant experience with Go and Terraform.
Familiarity with working in rapid-growth, startup environments.
Experience writing company-facing blog posts and training materials.
Full-Time Employee Benefits Include:
💰 Competitive Salary & Equity
💹 401(k) Program with a 4% match ( US Only)
⚕️ Health, Dental, Vision and Life Insurance
🩼 Short Term and Long Term Disability
🚼 Paid Parental, Medical, Caregiver Leave
🏝 Flexible Time Off (FTO) + Holidays
🚗 Commuter Benefits ( In-Office & US Only)
📱 Monthly Wellness Stipend
🧑💻 Autonomous Work Environment
🖥 In Office Set-Up Reimbursement ( In-Office Only)
🚀 Quarterly Team Gatherings
☕ In Office Amenities ( In-Office Only)
Want to learn more about what we are up to?
Self-driving Company
Replit Agent at Scale
AI Adoption
Build Open-Source Apps
Interviewing + Culture at Replit
Operating Principles
Reasons not to work at Replit
To achieve our mission of making programming more accessible around the world, we need our team to be representative of the world. We welcome your unique perspective and experiences in shaping this product. We encourage people from all kinds of backgrounds to apply, including and especially candidates from underrepresented and non-traditional backgrounds.
Designs and leads HubSpot's observability platform architecture, including distributed tracing, metrics, logging, and AI/ML system telemetry at scale across hundreds of microservices.
POS-5690
The Observability team owns the internal platform that gives every HubSpot engineer real visibility into how their systems behave in production. We build and operate the distributed tracing, metrics, alerting, and logging infrastructure that spans hundreds of microservices, billions of daily events, and thousands of engineers who depend on that signal to ship reliably.
We are now investing in the next generation of this platform. As HubSpot deploys AI agents and ML-powered features across the product, the team is building the tracing and telemetry primitives that make it possible to understand, debug, and trust what those systems are doing in production. This is greenfield, technically interesting work at a scale few companies operate at and we are looking for a Principal Engineer to help lead it.
We are seeking a Principal Software Engineer to be the technical anchor for HubSpot’s Observability platform. This role sits at the intersection of large-scale distributed systems, developer platform design, and AI observability. A big part of this role is working horizontally across a large engineering org and setting patterns and standards that make it easier for teams to instrument, alert on, and reason about their services. You will also shape how we trace and understand our growing fleet of AI agents and ML systems in production: a technically distinct and increasingly critical problem.
HubSpot is scaling fast — more engineers, more microservices, more AI systems running in production — and the Observability platform is at an inflection point. The foundations are solid, but the next phase requires a different kind of investment: rethinking cardinality economics, making OpenTelemetry the default across a large polyglot org, and building an entirely new layer of AI tracing that doesn’t yet exist.
This Principal Engineer will have a direct line of sight from the architecture they design to the engineering outcomes we measure: incident response times, adoption rates, developer satisfaction, and the trust that product teams place in their production signal. The AI observability layer in particular is greenfield — there is no playbook to follow, which is exactly what makes this the right moment for the right person.
If you want to build the platform that helps thousands of engineers understand what their systems are doing — including systems powered by AI that are genuinely hard to see inside — this is the role.
Pay & Benefits
The cash compensation below includes base salary, on-target commission for employees in eligible roles, and annual bonus targets under HubSpot’s bonus plan for eligible roles. In addition to cash compensation, some roles are eligible to participate in HubSpot’s equity plan to receive restricted stock units (RSUs). Some roles may also be eligible for overtime pay. Individual compensation packages are tailored to your skills, experience, qualifications, and other job-related reasons.
This resource will help guide how we recommend thinking about the range you see. Learn more about HubSpot’s compensation philosophy.
Benefits are also an important piece of your total compensation package. Explore the benefits and perks HubSpot offers to help employees grow better.
At HubSpot, fair compensation practices aren’t just about checking off the box for legal compliance. It’s about living out our value of transparency with our employees, candidates, and community.
Annual Cash Compensation Range:
$313,800—$502,080 USD
At HubSpot, we value both flexibility and connection. Whether you’re a Remote employee or work from the Office, we want you to start your journey here by building strong connections with your team and peers. If you are joining our Engineering team, you will be required to attend a regional HubSpot office for in-person onboarding. If you join our broader Product team, you’ll also attend other in-person events, such as your Product Group Summit and other gatherings, to continue building on those connections.
If you require an accommodation due to travel limitations or other reasons, please inform your recruiter during the hiring process. We are committed to supporting candidates who may need alternative arrangements
About HubSpot
HubSpot (NYSE: HUBS) is an AI-powered customer platform with all the software, integrations, and resources customers need to connect marketing, sales, and service. HubSpot’s connected platform enables businesses to grow faster by focusing on what matters most: customers.
At HubSpot, bold is our baseline. Our employees around the globe move fast, stay customer-obsessed, and win together. Our culture is grounded in four commitments: Solve for the Customer, Be Bold, Learn Fast, Align, Adapt & Go!, and Deliver with HEART. These commitments shape how we work, lead, and grow.
We’re building a company where people can do their best work. We focus on brilliant work, not badge swipes. By combining clarity, ownership, and trust, we create space for big thinking and meaningful progress. And we know that when our employees grow, our customers do too.
Recognized globally for our award-winning culture by Comparably, Glassdoor, Fortune, and more, HubSpot is headquartered in Cambridge, MA, with employees and offices around the world.
Explore more:
If you need accommodations or assistance due to a disability, please reach out to us using this form.
Massachusetts Applicants: It is unlawful in Massachusetts to require or administer a lie detector test as a condition of employment or continued employment. An employer who violates this law shall be subject to criminal penalties and civil liability.
Germany Applicants: (m/f/d) - link to HubSpot’s Career Diversity page here.
India Applicants: link to HubSpot India’s equal opportunity policy here.
HubSpot may use AI to help screen or assess candidates, but all hiring decisions are always human. More information can be found here. By submitting your application, you agree that HubSpot may collect your personal data for recruiting, global organization planning, and related purposes. We may use CLEAR ID Verification during the hiring process to confirm your identity and help maintain a safe, secure, and trusted experience for all candidates. Refer to HubSpot’s Recruiting Privacy Notice for details on data processing and your rights.
Senior Staff Software Engineer oversees technical roadmap for partner integrations, building APIs and SDKs that enable external partners to connect to Flex's rent payment platform.
Flex is a growth-stage, NYC headquartered FinTech company that is creating the best rent payment experience. It’s hard to believe that it’s 2026 and paying rent on time is expensive, inflexible, and difficult. We’re here to change that! Flex enables our users to pay rent throughout the month on a schedule that better fits their finances and budget. Our mission is to empower as many renters as possible with flexibility over their most significant recurring expense. After deliberately keeping a stealth profile as we built up unprecedented investor support and an enthusiastic user base, we are looking for motivated individuals to help us keep our mission growing. Will you be a part of the team?
About Our Opportunity
Flex exists to make paying rent work the way real life actually works — smoothing out cash flow, helping renters avoid late fees, and letting them build credit instead of falling behind. As a Senior Staff Software Engineer, you’ll oversee the technical roadmap for the team, working across APIs, SDKs, and Web experiences. You will work with teams across the organization including Engineering, Product, Design, Infrastructure, Sales, Partner and Customer Success to ensure that the technical strategy meets our goals. We expect you to be hands-on and execute work as an individual, and build products that allow for flexibility as we evolve our product offerings.
About Our Team!
We have some really exciting teams who are looking for an amazing engineer like you! Check them out!
Who thrives here?
Qualifications:
Flex takes a market-based approach to pay, and compensation may vary depending on your primary work location. Work locations are categorized into one of three tiers based on a cost of labor index for that geographic area. The successful candidate’s starting pay will be commensurate with their experience, qualifications, and Flex’s internal leveling guidelines and benchmarks.
Tier 1 (NYC/Bay Area, Los Angeles, Seattle)
$240,000—$300,000 USD
Tier 2 (Austin, Washington D.C. Philadelphia, San Diego, Chicago, Atlanta)
$216,000—$270,000 USD
Tier 3 (Salt Lake City, all other USA cities)
$204,000—$255,000 USD
We understand that it takes a diverse team of highly intelligent, curious, determined, empathetic, and self aware people to grow a successful company. Our HQ is located in New York City, but we have employees located throughout the US, Australia, Canada and South America. We are growing quickly, but deliberately, with a focus on building an inclusive culture. Our dynamic team has incredible perspectives to share, just as we know you do, and we take great pride in being an equal opportunity workplace.
Offices
Roles posted in New York, San Francisco, and Salt Lake City are hybrid positions with on-site expectations of 2-3 days per week in our local offices. For candidates outside of these areas, you may be eligible for our relocation assistance program.
Benefits
For full-time U.S. employees we offer:
For full-time non-U.S. employees, we offer:
Technically leads the Foundations platform team, building shared backend infrastructure and primitives that other engineering teams depend on.
Leads a team of forward-deployed engineers building full-stack digital solutions across manufacturing sites, managing technical strategy, delivery, and people development.
At Re:Build, our mission is to ensure the next generation of important products are made, at scale, in America. We are laying the foundation for a better future for our customers, employees, and communities by revitalizing America’s manufacturing base and creating meaningful jobs across the country, including in historically deindustrialized regions.
We operate an advanced, end-to-end manufacturing platform that partners with industrial companies and innovators to take products from first concept to full-scale production in critical verticals including aerospace and defense, electrification, medical, energy and environment, and robotics and automation.
The way we operate is as important as the work we do. It’s guided by The Re:Build Way, 16 principles that shape how we collaborate with each other, partner with our customers and vendors, and contribute to the communities where we operate. (link to The Re:Build Way principles )
The Director, Forward Deployed Engineering is a hands-on leadership role in Re:Build Manufacturing’s Digital Innovation Group (DIG). This role reflects the organization’s belief that digital innovation is key to scaling as a premier industrial company. In this position, the Director is responsible for hiring, leading, and developing a team of Forward Deployed Engineers (FDEs) who are embedded across Re:Build sites and Resource Center functions, implementing full-stack digital solutions that solve real-world business problems and drive measurable business impact. Success in the role depends on the ability to build trust, ensure alignment, and partner effectively with internal stakeholders throughout the entire product lifecycle. The ideal candidate combines a strong software engineering background with proven experience managing software delivery and solution engineering that directly serves customers. They bring deep technical credibility, strong people leadership skills, and the judgment needed to keep engagements well prioritized and delivered quickly.
What you get to do
What you bring to the Team
The BIG payoff
We are a company who is going to make a difference in the industries and the communities in which we choose to operate. Every employee of Re:Build will share ownership in the company and will share in the financial rewards of the success we achieve together, at all levels of the company!
We want to work with people that reflect the communities in which we operate
Re:Build Manufacturing is proud to be an Equal Employment Opportunity and Affirmative Action employer. We do not discriminate based upon race, color, religion, gender, gender identity or expression, sexual orientation, national origin, genetics, disability, age, veteran status, marital status, parental status, cultural background, organizational level, work styles, tenure and life experiences. Or for any other reason.
Re:Build is committed to providing reasonable accommodations for qualified individuals with disabilities in our job application procedures. If you need assistance or an accommodation due to a disability, you may contact us at accommodations.ta@ReBuildmanufacturing.com or you may call us at 617.909.6275.
Designs and manages enterprise data platforms, lakehouse architectures, and data pipelines while leading teams on data governance, quality, and compliance.
\*\*\* This is where your organization can create a consistent intro to all of your jobs, creating consistency in voice and messaging across all job posts
\*\*\* C’est ici que votre organisation peut créer une introduction cohérente à tous vos emplois, en créant une cohérence dans la voix et la messagerie dans tous les postes.
Marlabs, a global AI and Digital Solutions Consulting firm, delivers intelligent solutions across AI, data, analytics, and product engineering. Since 2000, we have partnered with some of the largest healthcare, life sciences, financial services, and government organizations worldwide. As we continue to expand our global footprint, we have an exciting opportunity for a highly skilled Lead Data Engineer to join our innovative and dynamic team.
Lead Data Engineer | About You
As a Lead Data Engineer, you will be responsible for designing, building, and managing the organization’s modern data platform, ensuring reliable, secure, and scalable data products that support analytics, reporting, and AI-driven business initiatives. You will lead the development of enterprise data pipelines, lakehouse architecture, and governance frameworks while partnering closely with AI/ML, platform engineering, and security teams. The ideal candidate combines deep expertise in data engineering, data modeling, cloud-based architectures, and data governance with a strong focus on reliability, observability, and regulatory compliance.
Lead Data Engineer | Day-to-Day
Lead Data Engineer | Skills & Experience
\*\*\* Similar to the introduction that can precede all job descriptions, an outro can also be formatted for consistency on all posts
\*\*\* Semblable à l’introduction qui peut précéder toutes les descriptions de poste, une outro peut également être formatée pour la cohérence sur tous les messages
Staff-level engineer who writes production code, leads technical projects, mentors engineers, and shapes architectural decisions across multiple technology stacks.
At Fluxon, we believe that how you build matters as much as what you build. We help businesses navigate their most important technology decisions with confidence, and take responsibility for seeing them through. Founded by ex-Googlers and startup veterans, we’re proud to partner with teams behind some of the most ambitious products, including Google, OpenAI, Anthropic, Walmart and Stripe.
Our work spans strategy, design, and engineering — often in complex, AI-driven environments — where clarity, speed and quality are the standard. We use AI intentionally, applying it only where it adds real value and expands what’s possible. Care shapes everything we do.
Inside Fluxon, you’ll find a global, remote-first team of experienced builders, who are curious, kind and serious about their craft. We’re building a place where people can take ownership, solve problems that matter and do work they’re proud to stand behind. If you want to do your best work alongside people who care as much as you do, you’ll feel at home here.
This role is fully remote, with candidates based in Buenos Aires, Argentina.
As a Staff Software Engineer at Fluxon, you’ll play a key role in shaping the technical direction of our engineering organization. This is a highly senior, hands-on leadership position where you’ll partner closely with company and engineering leadership to influence strategy, guide architectural decisions, and elevate our overall engineering practice. All Staff Engineers write production code, and everyone joins Fluxon as an individual contributor before stepping into project leadership or management.
You’ll be responsible for:
You’ll work with a diversity of technologies, including:
Core Languages & Runtimes
Frameworks & Ecosystems
Cloud & Infrastructure
Data & Messaging Services
Data Stores
Advanced Technologies & Architecture
We believe diverse teams perform better, and an inclusive environment is essential to building a successful organization. We welcome applicants from all backgrounds, experiences, and perspectives. We are an equal opportunity employer and are committed to providing accommodations throughout the hiring process.
This role uses AI-assisted tools to support initial screening. All assessments and decisions are made by a human reviewer.
Technical Lead designs and manages GPU infrastructure, Kubernetes deployment, and managed inference platform architecture for Tether's AI compute services.
Join Tether and Shape the Future of Digital Finance
At Tether, we’re not just building products, we’re pioneering a global financial revolution. Our cutting-edge solutions empower businesses—from exchanges and wallets to payment processors and ATMs—to seamlessly integrate reserve-backed tokens across blockchains. By harnessing the power of blockchain technology, Tether enables you to store, send, and receive digital tokens instantly, securely, and globally, all at a fraction of the cost. Transparency is the bedrock of everything we do, ensuring trust in every transaction.
Innovate with Tether
Tether Finance: Our innovative product suite features the world’s most trusted stablecoin, USDT, relied upon by hundreds of millions worldwide, alongside pioneering digital asset tokenization services.
But that’s just the beginning:
Tether Power: Driving sustainable growth, our energy solutions optimize excess power for Bitcoin mining using eco-friendly practices in state-of-the-art, geo-diverse facilities.
Tether Data: Fueling breakthroughs in AI and peer-to-peer technology, we reduce infrastructure costs and enhance global communications with cutting-edge solutions like KEET, our flagship app that redefines secure and private data sharing.
Tether Education: Democratizing access to top-tier digital learning, we empower individuals to thrive in the digital and gig economies, driving global growth and opportunity.
Tether Evolution: At the intersection of technology and human potential, we are pushing the boundaries of what is possible, crafting a future where innovation and human capabilities merge in powerful, unprecedented ways.
Why Join Us?
Our team is a global talent powerhouse, working remotely from every corner of the world. If you’re passionate about making a mark in the fintech space, this is your opportunity to collaborate with some of the brightest minds, pushing boundaries and setting new standards. We’ve grown fast, stayed lean, and secured our place as a leader in the industry.
If you have excellent English communication skills and are ready to contribute to the most innovative platform on the planet, Tether is the place for you.
Are you ready to be part of the future?
About the job
Cosmic AC is Tether Data’s GPU compute and managed inference platform: GPU containers, managed inference endpoints and platform observability, delivered as a self-hosted package on Kubernetes, with a control plane written in JavaScript. The platform is expanding from orchestrating workloads on a managed cluster to owning the full stack on bare-metal GPU infrastructure: a managed Slurm scheduling layer for internal research and model-training teams first, and our own Kubernetes control plane for inference tenancy after that.
The Technical Lead owns the architecture and delivery of that stack and leads the engineering team building it: about twelve engineers across backend, frontend, DevOps, QA and documentation, distributed across Europe and India. The role reports to the Senior Technical Product Manager for Cosmic AC, who owns scope, sequencing and partner commitments; the Technical Lead owns architecture, implementation and delivery plans, line-manages the engineers, and is the primary technical interface to our infrastructure partners.
This is a hands-on infrastructure leadership role with a fixed delivery window in its first six months. It is not a research role, not a pure Kubernetes SRE role, and not a management-only role.
Responsibilities
Architecture. Own the platform architecture end to end: architecture proposals, high-level and low-level designs, driven through review and kept current as the baseline.
Team leadership. Lead and line-manage a distributed team across backend (Node.js), frontend (React), DevOps, QA and documentation: engineering standards, code and design review, release gates, one-to-ones, growth and performance input.
Bare-metal GPU scheduling layer. Design, build and operate a managed Slurm service for research users: controller and accounting, partitions and login nodes, node onboarding and acceptance, driver and CUDA baseline and upgrades, stalled-job and node-health detection, drain and autohealing, storage visibility, identity and isolation.
Kubernetes control plane and GPU enablement. Own cluster bootstrap and lifecycle on partner-provided bare metal, NVIDIA GPU Operator and Network Operator, VM-based GPU isolation (KubeVirt and VFIO), and day-2 operations: upgrades, backup and recovery, node replacement.
Managed inference at scale. Serving architecture, multi-GPU and multi-node parallelism, autoscaling, request routing and endpoint reliability; confidential-compute-capable capacity for sensitive workloads.
Observability and operations. Metrics, logging, alerting and SLOs across control plane, GPU fleet and application tiers; incident response and post-incident review; an on-call model a small team can sustain.
Partners and vendors. Primary technical interface to infrastructure partners and vendors: turning requirements into written specifications and acceptance tests, running escalations to closure, and providing technical input to capacity planning and hardware sourcing.
Internal consumers. Work directly with research, model-training and product teams to translate their workloads into platform requirements, and broker capacity when it is short.
Hiring. Complete the platform team and set the technical bar for the engineers who join it.
Must have
Experience. Eight or more years of hands-on engineering, including at least three leading teams that build and operate infrastructure platforms other teams depend on. Bachelor’s or Master’s degree in computer science or engineering, or equivalent practical experience.
Slurm at scale, hands on. Has run slurmctld and slurmdbd for real users: partitions, QoS and priority, accounting, prolog and epilog, node health scripting, upgrades with jobs on the system. Ideally has operated an HPC or GPU training cluster for a research population.
GPU fleet operation on bare metal. NVIDIA driver and CUDA lifecycle, Fabric Manager and NVSwitch behaviour on SXM systems, DCGM-based health and utilisation, MIG, node burn-in and acceptance.
High-performance interconnects. InfiniBand fabric and subnet configuration, RDMA, SR-IOV, and diagnosing multi-node NCCL performance problems.
Linux systems depth. Kernel modules and drivers, PCIe passthrough and vfio-pci, cgroups and namespaces, performance tuning for compute-heavy workloads.
Production Kubernetes operation, not just deployment: control plane, upgrades, CNI and CSI, operators and custom controllers, multi-tenancy design.
HPC storage and data movement. Shared filesystems (VAST, Lustre, NFS), node-local NVMe caching, distributing large model weights and datasets across many nodes.
Observability and operations. Prometheus, Grafana and Loki or equivalents, SLOs, incident response and post-incident review.
Working fluency in JavaScript and Node.js sufficient to review a control plane, CLI and worker services with authority and to make architecture decisions on them. Not a feature-development requirement.
A shipped platform with real users. A multi-tenant IaaS or PaaS, or a research computing service: resource isolation, quotas, usage metering, and user-facing API and CLI surfaces.
Leadership that stays in the code. People management across time zones, cross-track review, written architecture decisions with alternatives recorded, and the ability to tell a partner or an executive no with reasons.
Excellent written and spoken English. Most partner and leadership work happens in writing.
Location. Fully remote, based between UTC and UTC+5:30 so the working day overlaps both Europe and India, where the team and its partners work. Occasional travel to partner sites and team events.
Desirable
Slurm operators on Kubernetes (Soperator, Slinky) or Kubernetes-native schedulers (Kueue, Volcano, KAI, Kubeflow Trainer).
Modern serving stacks (vLLM, SGLang, TensorRT-LLM): parallelism strategies, quantisation trade-offs, GPU memory planning.
VM and container isolation for multi-tenant GPU compute (KubeVirt, Kata Containers, QEMU and KVM, Firecracker); confidential computing (Intel TDX, AMD SEV-SNP, NVIDIA confidential-compute mode).
Cluster API and kubeadm, Cilium, NVSentinel-class autohealing, infrastructure as code and GitOps.
Time on the operator side of a GPU cloud, a national or university HPC centre, or an AI lab’s platform team.
Peer-to-peer or distributed-systems background.
Experience with a hardware provider who provisions but does not operate, and turning that relationship into a written contract with acceptance tests.
Important information for candidates
Recruitment scams have become increasingly common. To protect yourself, please keep the following in mind when applying for roles:
Apply only through our official channels. We do not use third-party platforms or agencies for recruitment unless clearly stated. All open roles are listed on our official careers page: https://tether.recruitee.com/
Verify the recruiter’s identity. All our recruiters have verified LinkedIn profiles. If you’re unsure, you can confirm their identity by checking their profile or contacting us through our website.
Be cautious of unusual communication methods. We do not conduct interviews over WhatsApp, Telegram, or SMS. All communication is done through official company emails and platforms.
Double-check email addresses. All communication from us will come from emails ending in @ tether.to or @ tether.io
We will never request payment or financial details. If someone asks for personal financial information or payment at any point during the hiring process, it is a scam. Please report it immediately.
When in doubt, feel free to reach out through our official website.
Technical lead who designs and manages GPU infrastructure and Kubernetes-based compute platforms for AI inference and containerized workloads.
Join Tether and Shape the Future of Digital Finance
At Tether, we’re not just building products, we’re pioneering a global financial revolution. Our cutting-edge solutions empower businesses—from exchanges and wallets to payment processors and ATMs—to seamlessly integrate reserve-backed tokens across blockchains. By harnessing the power of blockchain technology, Tether enables you to store, send, and receive digital tokens instantly, securely, and globally, all at a fraction of the cost. Transparency is the bedrock of everything we do, ensuring trust in every transaction.
Innovate with Tether
Tether Finance: Our innovative product suite features the world’s most trusted stablecoin, USDT, relied upon by hundreds of millions worldwide, alongside pioneering digital asset tokenization services.
But that’s just the beginning:
Tether Power: Driving sustainable growth, our energy solutions optimize excess power for Bitcoin mining using eco-friendly practices in state-of-the-art, geo-diverse facilities.
Tether Data: Fueling breakthroughs in AI and peer-to-peer technology, we reduce infrastructure costs and enhance global communications with cutting-edge solutions like KEET, our flagship app that redefines secure and private data sharing.
Tether Education: Democratizing access to top-tier digital learning, we empower individuals to thrive in the digital and gig economies, driving global growth and opportunity.
Tether Evolution: At the intersection of technology and human potential, we are pushing the boundaries of what is possible, crafting a future where innovation and human capabilities merge in powerful, unprecedented ways.
Why Join Us?
Our team is a global talent powerhouse, working remotely from every corner of the world. If you’re passionate about making a mark in the fintech space, this is your opportunity to collaborate with some of the brightest minds, pushing boundaries and setting new standards. We’ve grown fast, stayed lean, and secured our place as a leader in the industry.
If you have excellent English communication skills and are ready to contribute to the most innovative platform on the planet, Tether is the place for you.
Are you ready to be part of the future?
About the job
Cosmic AC is Tether Data’s GPU compute and managed inference platform: GPU containers, managed inference endpoints and platform observability, delivered as a self-hosted package on Kubernetes, with a control plane written in JavaScript. The platform is expanding from orchestrating workloads on a managed cluster to owning the full stack on bare-metal GPU infrastructure: a managed Slurm scheduling layer for internal research and model-training teams first, and our own Kubernetes control plane for inference tenancy after that.
The Technical Lead owns the architecture and delivery of that stack and leads the engineering team building it: about twelve engineers across backend, frontend, DevOps, QA and documentation, distributed across Europe and India. The role reports to the Senior Technical Product Manager for Cosmic AC, who owns scope, sequencing and partner commitments; the Technical Lead owns architecture, implementation and delivery plans, line-manages the engineers, and is the primary technical interface to our infrastructure partners.
This is a hands-on infrastructure leadership role with a fixed delivery window in its first six months. It is not a research role, not a pure Kubernetes SRE role, and not a management-only role.
Responsibilities
Architecture. Own the platform architecture end to end: architecture proposals, high-level and low-level designs, driven through review and kept current as the baseline.
Team leadership. Lead and line-manage a distributed team across backend (Node.js), frontend (React), DevOps, QA and documentation: engineering standards, code and design review, release gates, one-to-ones, growth and performance input.
Bare-metal GPU scheduling layer. Design, build and operate a managed Slurm service for research users: controller and accounting, partitions and login nodes, node onboarding and acceptance, driver and CUDA baseline and upgrades, stalled-job and node-health detection, drain and autohealing, storage visibility, identity and isolation.
Kubernetes control plane and GPU enablement. Own cluster bootstrap and lifecycle on partner-provided bare metal, NVIDIA GPU Operator and Network Operator, VM-based GPU isolation (KubeVirt and VFIO), and day-2 operations: upgrades, backup and recovery, node replacement.
Managed inference at scale. Serving architecture, multi-GPU and multi-node parallelism, autoscaling, request routing and endpoint reliability; confidential-compute-capable capacity for sensitive workloads.
Observability and operations. Metrics, logging, alerting and SLOs across control plane, GPU fleet and application tiers; incident response and post-incident review; an on-call model a small team can sustain.
Partners and vendors. Primary technical interface to infrastructure partners and vendors: turning requirements into written specifications and acceptance tests, running escalations to closure, and providing technical input to capacity planning and hardware sourcing.
Internal consumers. Work directly with research, model-training and product teams to translate their workloads into platform requirements, and broker capacity when it is short.
Hiring. Complete the platform team and set the technical bar for the engineers who join it.
Must have
Experience. Eight or more years of hands-on engineering, including at least three leading teams that build and operate infrastructure platforms other teams depend on. Bachelor’s or Master’s degree in computer science or engineering, or equivalent practical experience.
Slurm at scale, hands on. Has run slurmctld and slurmdbd for real users: partitions, QoS and priority, accounting, prolog and epilog, node health scripting, upgrades with jobs on the system. Ideally has operated an HPC or GPU training cluster for a research population.
GPU fleet operation on bare metal. NVIDIA driver and CUDA lifecycle, Fabric Manager and NVSwitch behaviour on SXM systems, DCGM-based health and utilisation, MIG, node burn-in and acceptance.
High-performance interconnects. InfiniBand fabric and subnet configuration, RDMA, SR-IOV, and diagnosing multi-node NCCL performance problems.
Linux systems depth. Kernel modules and drivers, PCIe passthrough and vfio-pci, cgroups and namespaces, performance tuning for compute-heavy workloads.
Production Kubernetes operation, not just deployment: control plane, upgrades, CNI and CSI, operators and custom controllers, multi-tenancy design.
HPC storage and data movement. Shared filesystems (VAST, Lustre, NFS), node-local NVMe caching, distributing large model weights and datasets across many nodes.
Observability and operations. Prometheus, Grafana and Loki or equivalents, SLOs, incident response and post-incident review.
Working fluency in JavaScript and Node.js sufficient to review a control plane, CLI and worker services with authority and to make architecture decisions on them. Not a feature-development requirement.
A shipped platform with real users. A multi-tenant IaaS or PaaS, or a research computing service: resource isolation, quotas, usage metering, and user-facing API and CLI surfaces.
Leadership that stays in the code. People management across time zones, cross-track review, written architecture decisions with alternatives recorded, and the ability to tell a partner or an executive no with reasons.
Excellent written and spoken English. Most partner and leadership work happens in writing.
Location. Fully remote, based between UTC and UTC+5:30 so the working day overlaps both Europe and India, where the team and its partners work. Occasional travel to partner sites and team events.
Desirable
Slurm operators on Kubernetes (Soperator, Slinky) or Kubernetes-native schedulers (Kueue, Volcano, KAI, Kubeflow Trainer).
Modern serving stacks (vLLM, SGLang, TensorRT-LLM): parallelism strategies, quantisation trade-offs, GPU memory planning.
VM and container isolation for multi-tenant GPU compute (KubeVirt, Kata Containers, QEMU and KVM, Firecracker); confidential computing (Intel TDX, AMD SEV-SNP, NVIDIA confidential-compute mode).
Cluster API and kubeadm, Cilium, NVSentinel-class autohealing, infrastructure as code and GitOps.
Time on the operator side of a GPU cloud, a national or university HPC centre, or an AI lab’s platform team.
Peer-to-peer or distributed-systems background.
Experience with a hardware provider who provisions but does not operate, and turning that relationship into a written contract with acceptance tests.
Important information for candidates
Recruitment scams have become increasingly common. To protect yourself, please keep the following in mind when applying for roles:
Apply only through our official channels. We do not use third-party platforms or agencies for recruitment unless clearly stated. All open roles are listed on our official careers page: https://tether.recruitee.com/
Verify the recruiter’s identity. All our recruiters have verified LinkedIn profiles. If you’re unsure, you can confirm their identity by checking their profile or contacting us through our website.
Be cautious of unusual communication methods. We do not conduct interviews over WhatsApp, Telegram, or SMS. All communication is done through official company emails and platforms.
Double-check email addresses. All communication from us will come from emails ending in @ tether.to or @ tether.io
We will never request payment or financial details. If someone asks for personal financial information or payment at any point during the hiring process, it is a scam. Please report it immediately.
When in doubt, feel free to reach out through our official website.
Technical Lead manages GPU infrastructure, Kubernetes deployment, and platform observability for a distributed compute and inference platform.
Join Tether and Shape the Future of Digital Finance
At Tether, we’re not just building products, we’re pioneering a global financial revolution. Our cutting-edge solutions empower businesses—from exchanges and wallets to payment processors and ATMs—to seamlessly integrate reserve-backed tokens across blockchains. By harnessing the power of blockchain technology, Tether enables you to store, send, and receive digital tokens instantly, securely, and globally, all at a fraction of the cost. Transparency is the bedrock of everything we do, ensuring trust in every transaction.
Innovate with Tether
Tether Finance: Our innovative product suite features the world’s most trusted stablecoin, USDT, relied upon by hundreds of millions worldwide, alongside pioneering digital asset tokenization services.
But that’s just the beginning:
Tether Power: Driving sustainable growth, our energy solutions optimize excess power for Bitcoin mining using eco-friendly practices in state-of-the-art, geo-diverse facilities.
Tether Data: Fueling breakthroughs in AI and peer-to-peer technology, we reduce infrastructure costs and enhance global communications with cutting-edge solutions like KEET, our flagship app that redefines secure and private data sharing.
Tether Education: Democratizing access to top-tier digital learning, we empower individuals to thrive in the digital and gig economies, driving global growth and opportunity.
Tether Evolution: At the intersection of technology and human potential, we are pushing the boundaries of what is possible, crafting a future where innovation and human capabilities merge in powerful, unprecedented ways.
Why Join Us?
Our team is a global talent powerhouse, working remotely from every corner of the world. If you’re passionate about making a mark in the fintech space, this is your opportunity to collaborate with some of the brightest minds, pushing boundaries and setting new standards. We’ve grown fast, stayed lean, and secured our place as a leader in the industry.
If you have excellent English communication skills and are ready to contribute to the most innovative platform on the planet, Tether is the place for you.
Are you ready to be part of the future?
About the job
Cosmic AC is Tether Data’s GPU compute and managed inference platform: GPU containers, managed inference endpoints and platform observability, delivered as a self-hosted package on Kubernetes, with a control plane written in JavaScript. The platform is expanding from orchestrating workloads on a managed cluster to owning the full stack on bare-metal GPU infrastructure: a managed Slurm scheduling layer for internal research and model-training teams first, and our own Kubernetes control plane for inference tenancy after that.
The Technical Lead owns the architecture and delivery of that stack and leads the engineering team building it: about twelve engineers across backend, frontend, DevOps, QA and documentation, distributed across Europe and India. The role reports to the Senior Technical Product Manager for Cosmic AC, who owns scope, sequencing and partner commitments; the Technical Lead owns architecture, implementation and delivery plans, line-manages the engineers, and is the primary technical interface to our infrastructure partners.
This is a hands-on infrastructure leadership role with a fixed delivery window in its first six months. It is not a research role, not a pure Kubernetes SRE role, and not a management-only role.
Responsibilities
Architecture. Own the platform architecture end to end: architecture proposals, high-level and low-level designs, driven through review and kept current as the baseline.
Team leadership. Lead and line-manage a distributed team across backend (Node.js), frontend (React), DevOps, QA and documentation: engineering standards, code and design review, release gates, one-to-ones, growth and performance input.
Bare-metal GPU scheduling layer. Design, build and operate a managed Slurm service for research users: controller and accounting, partitions and login nodes, node onboarding and acceptance, driver and CUDA baseline and upgrades, stalled-job and node-health detection, drain and autohealing, storage visibility, identity and isolation.
Kubernetes control plane and GPU enablement. Own cluster bootstrap and lifecycle on partner-provided bare metal, NVIDIA GPU Operator and Network Operator, VM-based GPU isolation (KubeVirt and VFIO), and day-2 operations: upgrades, backup and recovery, node replacement.
Managed inference at scale. Serving architecture, multi-GPU and multi-node parallelism, autoscaling, request routing and endpoint reliability; confidential-compute-capable capacity for sensitive workloads.
Observability and operations. Metrics, logging, alerting and SLOs across control plane, GPU fleet and application tiers; incident response and post-incident review; an on-call model a small team can sustain.
Partners and vendors. Primary technical interface to infrastructure partners and vendors: turning requirements into written specifications and acceptance tests, running escalations to closure, and providing technical input to capacity planning and hardware sourcing.
Internal consumers. Work directly with research, model-training and product teams to translate their workloads into platform requirements, and broker capacity when it is short.
Hiring. Complete the platform team and set the technical bar for the engineers who join it.
Must have
Experience. Eight or more years of hands-on engineering, including at least three leading teams that build and operate infrastructure platforms other teams depend on. Bachelor’s or Master’s degree in computer science or engineering, or equivalent practical experience.
Slurm at scale, hands on. Has run slurmctld and slurmdbd for real users: partitions, QoS and priority, accounting, prolog and epilog, node health scripting, upgrades with jobs on the system. Ideally has operated an HPC or GPU training cluster for a research population.
GPU fleet operation on bare metal. NVIDIA driver and CUDA lifecycle, Fabric Manager and NVSwitch behaviour on SXM systems, DCGM-based health and utilisation, MIG, node burn-in and acceptance.
High-performance interconnects. InfiniBand fabric and subnet configuration, RDMA, SR-IOV, and diagnosing multi-node NCCL performance problems.
Linux systems depth. Kernel modules and drivers, PCIe passthrough and vfio-pci, cgroups and namespaces, performance tuning for compute-heavy workloads.
Production Kubernetes operation, not just deployment: control plane, upgrades, CNI and CSI, operators and custom controllers, multi-tenancy design.
HPC storage and data movement. Shared filesystems (VAST, Lustre, NFS), node-local NVMe caching, distributing large model weights and datasets across many nodes.
Observability and operations. Prometheus, Grafana and Loki or equivalents, SLOs, incident response and post-incident review.
Working fluency in JavaScript and Node.js sufficient to review a control plane, CLI and worker services with authority and to make architecture decisions on them. Not a feature-development requirement.
A shipped platform with real users. A multi-tenant IaaS or PaaS, or a research computing service: resource isolation, quotas, usage metering, and user-facing API and CLI surfaces.
Leadership that stays in the code. People management across time zones, cross-track review, written architecture decisions with alternatives recorded, and the ability to tell a partner or an executive no with reasons.
Excellent written and spoken English. Most partner and leadership work happens in writing.
Location. Fully remote, based between UTC and UTC+5:30 so the working day overlaps both Europe and India, where the team and its partners work. Occasional travel to partner sites and team events.
Desirable
Slurm operators on Kubernetes (Soperator, Slinky) or Kubernetes-native schedulers (Kueue, Volcano, KAI, Kubeflow Trainer).
Modern serving stacks (vLLM, SGLang, TensorRT-LLM): parallelism strategies, quantisation trade-offs, GPU memory planning.
VM and container isolation for multi-tenant GPU compute (KubeVirt, Kata Containers, QEMU and KVM, Firecracker); confidential computing (Intel TDX, AMD SEV-SNP, NVIDIA confidential-compute mode).
Cluster API and kubeadm, Cilium, NVSentinel-class autohealing, infrastructure as code and GitOps.
Time on the operator side of a GPU cloud, a national or university HPC centre, or an AI lab’s platform team.
Peer-to-peer or distributed-systems background.
Experience with a hardware provider who provisions but does not operate, and turning that relationship into a written contract with acceptance tests.
Important information for candidates
Recruitment scams have become increasingly common. To protect yourself, please keep the following in mind when applying for roles:
Apply only through our official channels. We do not use third-party platforms or agencies for recruitment unless clearly stated. All open roles are listed on our official careers page: https://tether.recruitee.com/
Verify the recruiter’s identity. All our recruiters have verified LinkedIn profiles. If you’re unsure, you can confirm their identity by checking their profile or contacting us through our website.
Be cautious of unusual communication methods. We do not conduct interviews over WhatsApp, Telegram, or SMS. All communication is done through official company emails and platforms.
Double-check email addresses. All communication from us will come from emails ending in @ tether.to or @ tether.io
We will never request payment or financial details. If someone asks for personal financial information or payment at any point during the hiring process, it is a scam. Please report it immediately.
When in doubt, feel free to reach out through our official website.
Senior Staff Software Engineer designs and builds full-stack platform capabilities for benefits applications, focusing on scalable systems and AI-driven interfaces.
About Gusto
At Gusto, we’re on a mission to grow the small business economy. We handle the hard stuff — payroll, health insurance, 401(k)s, and HR — so owners can focus on their craft and their customers. With teams in Denver, San Francisco, and New York, we support more than 500,000 small businesses nationwide and are building a workplace that reflects the people we serve.
All full-time employees receive competitive base pay, benefits, and equity (RSUs) — because everyone who helps build Gusto should share in its success. Offer amounts are determined by role, level, and location. Learn more about our Total Rewards philosophy.
AI is a fundamental part of how work gets done at Gusto. We expect all team members to actively engage with AI tools relevant to their role and grow their fluency as the technology evolves. AI experience requirements vary by role and will be assessed during the interview process.
About the Role
As a Staff Software Engineer on the Benefits Advisory team, you will be directly accountable for key architectural improvements to Gusto’s benefits platform. This is a full-stack role where you will design and build platform capabilities that enable customers to explore, apply for, and maintain their benefits within your product area.
Your work will focus on transforming our benefits opportunity, shopping and renewal flows, creating clear system boundaries that enable efficient reuse and increased scalability, allowing the ability to offer a wider variety of products more quickly. You will design services that continue to scale the organization, with an emphasis on enabling novel AI-driven interfaces to deliver a delightful benefits shopping experience.
You’ll operate at the intersection of marketing, sales, operations and engineering. If you are passionate about building highly scalable systems that can reason, predict, and personalize to unlock Gusto Benefits’ next phase of sustainable growth, we’d love to have you join our team.
About the Team
The Benefits Advisory team is building the next generation of infrastructure that powers benefits applications, making it easier than ever to give customers access to the absolute best benefits for their needs.
Our mission is to build a robust system that informs our customers of the most relevant benefits options, making it easier than ever to confidently fulfill benefits that fit our customers current and future needs. We’re building a world-class technology platform designed to understand customer context in real-time, and continuously improve through data and feedback. This means rapid experimentation, fast learning, and strong cross-functional partnership.
We prioritize quality, observability, and uptime because these intelligent systems are fundamental to Gusto’s growth and brand. We partner closely with Marketing, Sales, and Operations to build and connect the AI-powered tools they use every day.
Here’s what you’ll do day-to-day:
Build innovative AI interfaces to best assist our customers
Here’s what we’re looking for:
Compensation
Our cash compensation amount for this role is targeted at $191,000/yr to $225,000/yr in Denver & most remote locations, and $225,000/yr to $265,000/yr for San Francisco, Seattle & New York. Stock equity is additional. Final offer amounts are determined by multiple factors including candidate experience and expertise and may vary from the amounts listed above.
Gusto has physical office spaces in Denver, San Francisco, and New York City. Employees who are based in those locations will be expected to work from the office on designated days approximately 2-3 days per week (or more depending on role). The same office expectations apply to all Symmetry roles, Gusto’s subsidiary, whose physical office is in Scottsdale.
Note: The San Francisco office expectations encompass both the San Francisco and San Jose metro areas.
When approved to work from a location other than a Gusto office, a secure, reliable, and consistent internet connection is required. This includes non-office days for hybrid employees.
Our customers come from all walks of life and so do we. We hire great people from a wide variety of backgrounds, not just because it’s the right thing to do, but because it makes our company stronger. If you share our values and our enthusiasm for small businesses, you will find a home at Gusto.
Gusto is proud to be an equal opportunity employer. We do not discriminate in hiring or any employment decision based on race, color, religion, national origin, age, sex (including pregnancy, childbirth, or related medical conditions), marital status, ancestry, physical or mental disability, genetic information, veteran status, gender identity or expression, sexual orientation, or other applicable legally protected characteristic. Gusto considers qualified applicants with criminal histories, consistent with applicable federal, state and local law. Gusto is also committed to providing reasonable accommodations for qualified individuals with disabilities and disabled veterans in our job application procedures. We want to see our candidates perform to the best of their ability. If you require a medical or religious accommodation at any time throughout your candidate journey, please fill out this form and a member of our team will get in touch with you.
Gusto takes security and protection of your personal information very seriously. Please review our Fraudulent Activity Disclaimer.
Personal information collected and processed as part of your Gusto application will be subject to Gusto’s Applicant Privacy Notice.