GL

Grafana Labs

20 open positions available

1 location
1 employment type
Actively hiring
Full-time

Latest Positions

Showing 20 most recent jobs
GL

Senior FullStack Engineer - Grafana Cloud Observability| US | Remote

Grafana LabsAnywhereFull-time
View Job
Compensation$90K - 150K a year

Develop and maintain cloud integrations and observability features across frontend and backend using TypeScript, Go, and Jsonnet. | Requires strong backend and frontend development skills with Go and TypeScript, product focus, and good communication in observability ecosystem. | Grafana Labs, the company behind the open observability cloud, is founded on the principles of open source, open standards, open ecosystems, and open culture. Grafana Cloud, our fully managed observability platform, is flexible and built for scale. With Grafana Cloud's actually useful AI, organizations can see, understand, and act on all their disparate data to move at the speed of their ambitions. Today, more than 35 million users and 7,000+ customers – including Anthropic, Bloomberg, NVIDIA, Microsoft, and Salesforce – trust Grafana Labs to ensure reliability of their applications and systems, resolve incidents quickly, and optimize their telemetry to reduce noise and cost. We are a 100% remote company with 1,600+ team members across 40+ countries, and we’re backed by leading investors including Lightspeed Venture Partners, Sequoia Capital, GIC, Coatue, J.P. Morgan, CapitalG, and Lead Edge Capital. Learn more at grafana.com and follow us on LinkedIn and X. We’re scaling fast and staying true to what makes us different: an open-source legacy, a global collaborative culture, and a passion for meaningful work. Our team thrives in an innovation-driven environment where transparency, autonomy, and trust fuel everything we do. You may not meet every requirement, and that’s okay. If this role excites you, we’d love you to raise your hand for what could be a truly career-defining opportunity. This is a full-time remote opportunity. We are considering candidates from US and Canada only (EST time zone). The Opportunity: The Grafana Cloud Observability department is focused on enabling developers to understand the health and performance of their applications and infrastructure in any environment by providing tools to instrument their code, ingest observability data into Grafana Cloud and visualize and explore it. In this role, you will be part of the Cloud Integrations team. We develop our portfolio of integrations that allow customers to collect, visualize and act on telemetry from various systems and applications. The software you will create is a core building block for the users of Grafana Cloud, and has direct impact on Grafana Cloud’s out of the box experience. Our work is built on Open Source technologies such as Prometheus and OpenTelemetry, which you will also have the opportunity to actively contribute to. What You’ll Be Doing: You will bring your software engineering expertise to develop the tooling that powers our portfolio of cloud integrations and observability apps; working to develop, maintain and scale our infrastructure observability features in Grafana Cloud. We’re looking for someone who can work across the full stack: Frontend - written in Typescript / React Backend - written in Golang Monitoring Mixins - written in Jsonnet Day to day, that looks like: Ship features end to end, from the UI a user clicks through to the backend services behind it and the dashboards and alerts that ship with each integration Move between parts of the stack as the work demands. Some weeks are a React panel, some weeks a Go service, some weeks a Jsonnet mixin Own projects from problem statement to production, including talking to users and checking whether what you shipped actually worked Represent the team's work outside the team: write the design proposals, run the demos, and make the case for our roadmap in cross-team discussions Contribute upstream to Prometheus, OpenTelemetry and Grafana Alloy where our work touches them Join the on-call rotation for the services you own We invest heavily in developer productivity. You can use modern AI coding assistants as part of your daily workflow (your choice of tools, within security guidelines), backed by a company-funded usage budget so you can iterate quickly without unnecessary friction. We encourage pragmatic AI-assisted development: faster prototyping, test generation, refactors, documentation, and incident follow-ups—always with strong code review and quality standards. You'll also have access to the latest frontier models from OpenAI, Anthropic, and Google. As we have embraced a remote-first approach and our engineering team is primarily remote, it is essential to possess strong communication skills and the ability to work independently. We provide support and hold regular meetings through video calls to ensure effective collaboration and alignment. What Makes You a Great Fit: Interest in working across the full stack, with a product-focused mindset. You start from the user problem and are comfortable deciding what to build, not only how to build it Experience on both the backend and the frontend. We use Go and TypeScript; solid experience with comparable languages translates well Fluency with AI-assisted development. You use coding agents and LLM tooling as part of how you work, and you have a clear view of where they speed things up and where they get in the way Comfort being visible. You raise concerns early, advocate for your team's work across the org, and turn quiet good work into decisions other people can see and act on Working knowledge of the observability ecosystem: Prometheus, Grafana dashboards, alerting, and what it takes to run these systems in production Strong written communication. Most of our decisions happen in docs, issues and pull requests across several time zones Bonus Points For: Experience contributing to or maintaining Open Source projects, with evidence of successful pull requests and community collaboration Experience working on Developer Tooling with a large user base Experience building Grafana plugins, apps or dashboards Familiarity with Grafana Alloy, the OpenTelemetry Collector, or similar telemetry pipeline tooling Experience writing Jsonnet, or another configuration-as-code language You have been on-call for the systems you instrument, so you know what a useful alert looks like In US, the OTE (On-Target Earnings)/Base compensation range for this role is USD $154,445 - $185.334. Actual compensation may vary based on level, experience, and skillset as assessed throughout the interview process. All of our roles include Restricted Stock Units (RSUs), giving every team member ownership in Grafana Labs' success. We believe in shared outcomes—RSUs help us stay aligned and invested as we scale globally. *Compensation ranges are country specific. If you are applying for this role from a different location than listed above, your recruiter will discuss your specific market’s defined pay range & benefits at the beginning of the process. Why You’ll Thrive at Grafana Labs: 100% Remote, Global Culture - As a remote-only company, we bring together talent from around the world, united by a culture of collaboration and shared purpose. Scaling Organization – Tackle meaningful work in a high-growth, ever-evolving environment. Transparent Communication – Expect open decision-making and regular company-wide updates. Innovation-Driven – Autonomy and support to ship great work and try new things. Open Source Roots – Built on community-driven values that shape how we work. Empowered Teams – High trust, low ego culture that values outcomes over optics. Career Growth Pathways – Defined opportunities to grow and develop your career. Approachable Leadership – Transparent execs who are involved, visible, and human. Passionate People – Join a team of smart, supportive folks who care deeply about what they do. In-Person onboarding - We want you to thrive from day 1 with your fellow new ‘Grafanistas’ to learn all about what we do and how we do it. Balance is Key - We operate a global annual leave policy of 30 days per annum. 3 days of your annual leave entitlement are reserved for Grafana Shutdown Days to allow the team to really disconnect. *We will comply with local legislation where applicable. Equal Opportunity Employer: Grafana Labs is an equal opportunities employer. We welcome applications from everyone regardless of race, colour, nationality, origin, caste, sex, gender reassignment identity or expression, sexual orientation, age, religion or belief, disability, veteran status, genetic information, pregnancy, maternity, marital, family or carer status, or any other characteristic which is protected by local law. We believe that equality and diversity build a strong organisation, and we work hard to ensure that is the foundation of our organisation as we grow. Grafana Labs may utilize AI tools in its recruitment process to assist in matching information provided in CVs to job postings. The recruitment team will continue to review inbound CVs manually to identify alignment with current openings. #LI-Remote For information about how your personal data is used once you’ve applied to a job, check out our privacy policy.

TypeScript
Go
React
API development
Observability
Cloud infrastructure
Direct Apply
Posted 5 days ago
Grafana Labs

Senior AI Engineer - Grafana AI/ML | USA | Remote

Grafana LabsAnywhereFull-time
View Job
Compensation$120K - 180K a year

Develop and scale AI features and agentic workflows for incident detection and resolution using observability data. | Strong software engineering background with experience in production-ready GenAI applications, LLM prompt engineering, cloud-native environments, and observability tools. | Grafana Labs, the company behind the open observability cloud, is founded on the principles of open source, open standards, open ecosystems, and open culture. Grafana Cloud, our fully managed observability platform, is flexible and built for scale. With Grafana Cloud's actually useful AI, organizations can see, understand, and act on all their disparate data to move at the speed of their ambitions. Today, more than 35 million users and 7,000+ customers – including Anthropic, Bloomberg, NVIDIA, Microsoft, and Salesforce – trust Grafana Labs to ensure reliability of their applications and systems, resolve incidents quickly, and optimize their telemetry to reduce noise and cost. We are a 100% remote company with 1,600+ team members across 40+ countries, and we’re backed by leading investors including Lightspeed Venture Partners, Sequoia Capital, GIC, Coatue, J.P. Morgan, CapitalG, and Lead Edge Capital. Learn more at grafana.com and follow us on LinkedIn and X. We’re scaling fast and staying true to what makes us different: an open-source legacy, a global collaborative culture, and a passion for meaningful work. Our team thrives in an innovation-driven environment where transparency, autonomy, and trust fuel everything we do. You may not meet every requirement, and that’s okay. If this role excites you, we’d love you to raise your hand for what could be a truly career-defining opportunity. This is a remote opportunity and we would be interested in applicants from USA time zones only at this time. Senior AI Engineer The Opportunity: At Grafana, we build observability tools that help users understand, respond to, and improve their systems – regardless of scale, complexity, or tech stack. The Grafana AI teams play a key role in this mission by helping users make sense of complex observability data through AI-driven features. These capabilities reduce toil, lower the barrier of domain expertise, and surface meaningful signals from noisy environments. What makes our team different is how we work: we operate with a high degree of autonomy and ownership, both as individuals and as a team. Engineers are empowered to make decisions, move quickly, and validate ideas early – while being supported by a deeply collaborative culture that values curiosity, feedback, and cross-functional partnership. We’re looking for an AI Software Engineer with a strong software engineering background, a quick iteration mindset, and a passion for experimentation – balanced by a focus on shipping and scaling impactful features that deliver value to users. You’ll work closely with cross-functional teams to develop, test, and ship AI-powered features that contribute to improving infrastructure and observability quality through automation, while also expanding the capabilities of AI agents across the observability stack to assist users with incident response. As the team matures, there’s a broad opportunity to expand or redefine this role based on impact and initiative. What You’ll Be Doing: Build and deliver AI solutions: Take ownership of developing high-performance AI features to help users detect, triage, and resolve incidents using observability data and tools. Rapid experimentation and iteration: Implement a highly iterative process where you quickly prototype, test, and validate with real users, including shipping and evolving LLM- or agent-powered workflows for incident lifecycle management and automated analysis tasks. Collaborate cross-functionally: Work with data analysts, product managers, and designers to shape AI-driven product features, including integration of agentic components with internal tools, alerting systems, runbooks, and developer workflows. Utilize AI tools effectively: Use AI and automation tools to enhance both product functionality and your own development workflows. Effective communication: You’ll be working in a highly dynamic and collaborative environment, so we need someone who can communicate effectively and contribute across teams. Ownership and impact: Take full ownership of the AI solutions you develop, ensuring they are not only innovative but also scalable, maintainable, and aligned with real user workflows. We invest heavily in developer productivity. You can use modern AI coding assistants as part of your daily workflow (your choice of tools, within security guidelines), backed by a company-funded usage budget so you can iterate quickly without unnecessary friction. We encourage pragmatic AI-assisted development: faster prototyping, test generation, refactors, documentation, and incident follow-ups—always paired with strong code review and quality standards. You’ll also have access to frontier models (e.g., GPT-Codex 5/3, Claude Opus 4.6, Gemini 3 Pro). What Makes You a Great Fit: Strong engineering skills: Solid experience building production software systems (backend and / or full stack). You’re a self-starter, capable of tackling complex engineering problems with minimal supervision. AI experience with a practical mindset: You’re familiar with AI technologies and frameworks, and you focus on delivering high-quality solutions that work in the real world, not just in theory. Quick iteration and experimentation: You’re comfortable releasing prototypes, collecting feedback, and iterating with a pragmatic mindset. Proven initiative: You take ownership and drive projects forward, pushing boundaries to find the most impactful solutions. You can deal with ambiguity and are able to define scope where things are loosely defined. Collaborative attitude: You communicate effectively with peers, product managers, and designers. You’re open to feedback, and you bring a solutions-oriented mindset to the table. Requirements: Experience with LLMs, prompt engineering, and building applications powered by GenAI. Proven track record of delivering software that made it into production and is actively used by users. Exposure to working in cloud-native environments (e.g., AWS, GCP, Azure). Experience using observability tools to understand and troubleshoot system behavior. Bonus Points For: Experience building or working with agent frameworks or multi‑agent workflows. Experience with infrastructure / devops related tooling: Kubernetes, Docker, Terraform or similar for deployments. Familiarity with model fine-tuning techniques. Experience building observability tooling. Compensation & Rewards: In the United States, the Base compensation range for this role is USD 127,651 - USD 203,867. Actual compensation may vary based on level, experience, and skillset as assessed in the interview process. Benefits include equity, bonus (if applicable) and other benefits listed here. All of our roles include Restricted Stock Units (RSUs), giving every team member ownership in Grafana Labs' success. We believe in shared outcomes—RSUs help us stay aligned and invested as we scale globally. *Compensation ranges are country specific. If you are applying for this role from a different location than listed above, your recruiter will discuss your specific market’s defined pay range & benefits at the beginning of the process. Why You’ll Thrive at Grafana Labs: 100% Remote, Global Culture - As a remote-only company, we bring together talent from around the world, united by a culture of collaboration and shared purpose. Scaling Organization – Tackle meaningful work in a high-growth, ever-evolving environment. Transparent Communication – Expect open decision-making and regular company-wide updates. Innovation-Driven – Autonomy and support to ship great work and try new things. Open Source Roots – Built on community-driven values that shape how we work. Empowered Teams – High trust, low ego culture that values outcomes over optics. Career Growth Pathways – Defined opportunities to grow and develop your career. Approachable Leadership – Transparent execs who are involved, visible, and human. Passionate People – Join a team of smart, supportive folks who care deeply about what they do. In-Person onboarding - We want you to thrive from day 1 with your fellow new ‘Grafanistas’ to learn all about what we do and how we do it. Balance is Key - We operate a global annual leave policy of 30 days per annum. 3 days of your annual leave entitlement are reserved for Grafana Shutdown Days to allow the team to really disconnect. *We will comply with local legislation where applicable. Equal Opportunity Employer: Grafana Labs is an equal opportunities employer. We welcome applications from everyone regardless of race, colour, nationality, origin, caste, sex, gender reassignment identity or expression, sexual orientation, age, religion or belief, disability, veteran status, genetic information, pregnancy, maternity, marital, family or carer status, or any other characteristic which is protected by local law. We believe that equality and diversity build a strong organisation, and we work hard to ensure that is the foundation of our organisation as we grow. Grafana Labs may utilize AI tools in its recruitment process to assist in matching information provided in CVs to job postings. The recruitment team will continue to review inbound CVs manually to identify alignment with current openings. #LI-Remote For information about how your personal data is used once you’ve applied to a job, check out our privacy policy.

Software Architecture
Observability Tools
Backend Development
Direct Apply
Posted 9 days ago
Grafana Labs

Staff Backend Engineer - Second Horizon | US | Remote

Grafana LabsAnywhereFull-time
View Job
Compensation$120K - 180K a year

Design and implement backend architecture for AI-native data intelligence focusing on scalable SaaS and APIs. | Requires production-grade backend software experience with LLMs, GenAI, cloud-native environments, and observability tools. | Grafana Labs, the company behind the open observability cloud, is founded on the principles of open source, open standards, open ecosystems, and open culture. Grafana Cloud, our fully managed observability platform, is flexible and built for scale. With Grafana Cloud's actually useful AI, organizations can see, understand, and act on all their disparate data to move at the speed of their ambitions. Today, more than 35 million users and 7,000+ customers – including Anthropic, Bloomberg, NVIDIA, Microsoft, and Salesforce – trust Grafana Labs to ensure reliability of their applications and systems, resolve incidents quickly, and optimize their telemetry to reduce noise and cost. We are a 100% remote company with 1,600+ team members across 40+ countries, and we’re backed by leading investors including Lightspeed Venture Partners, Sequoia Capital, GIC, Coatue, J.P. Morgan, CapitalG, and Lead Edge Capital. Learn more at grafana.com and follow us on LinkedIn and X. We’re scaling fast and staying true to what makes us different: an open-source legacy, a global collaborative culture, and a passion for meaningful work. Our team thrives in an innovation-driven environment where transparency, autonomy, and trust fuel everything we do. You may not meet every requirement, and that’s okay. If this role excites you, we’d love you to raise your hand for what could be a truly career-defining opportunity. This is a remote opportunity and we are looking for candidates from the U.S. or Canada. Residents of Quebec are not eligible for this role. At Grafana Labs, we build observability tools that help users understand, respond to, and improve their systems – regardless of scale, complexity, or tech stack. We recently started a skunkworks initiative with a mission to bring observability to the rest of the business. Our goal is to make Grafana the single best place where humans and AI agents understand and act on real-time data from across the enterprise. As part of this initiative, we are building an AI-native data intelligence system that gives agents reliable, governed access to enterprise context. This system helps agents retrieve the right data, metadata, definitions, lineage, quality signals, and institutional knowledge without every agent builder having to maintain their own brittle context files. The principle is simple: the context system retrieves, agents decide, and data teams maintain the intelligence. The new team is a mix of seasoned Grafanistas and new hires. We operate with a high degree of autonomy and ownership, both as individuals and as a team. Engineers are empowered to make decisions, move quickly, and validate ideas early – while being supported by a deeply collaborative culture that values curiosity, feedback, and cross-functional partnership. We’re looking for a Staff-level Backend Engineer to help build the first production services for this context layer management system. You will design and ship the core backend architecture that powers ingestion, context storage, retrieval APIs, and agent-facing integrations. This is an early-stage role on a high-autonomy team, so you should be comfortable working through ambiguity, making pragmatic architectural decisions, and building systems that can evolve from internal dogfooding to production-grade SaaS. As the team matures, there’s a broad opportunity to expand or redefine this role based on impact and initiative. Key responsibilities: Build the core backend services: Design, implement, test, and operate the first services for context ingestion, context indexing, retrieval orchestration, API access, source configuration, and system administration. Create a scalable SaaS foundation: Help define and build the architecture for a multi-tenant service, including tenant isolation, usage tracking, quotas, audit logs, background jobs, and reliable service boundaries. Power agent-facing retrieval workflows: Build APIs and service interfaces that allow AI agents, MCP tools, CLIs, and internal applications to retrieve relevant context, provenance, confidence signals, and warnings. Work across product and infrastructure: Partner with the team to make practical tradeoffs between fast experimentation and long-term reliability, especially as the project moves from prototype to production. Operate what you build: Instrument services with metrics, logs, traces, alerts, and dashboards. Use observability tools to understand system behavior and improve reliability. Contribute to technical direction: Help shape the architecture, service boundaries, storage choices, API contracts, deployment patterns, and engineering practices for a new product area. Effective communication: You’ll be working in a highly dynamic and collaborative environment, so we need someone who can communicate effectively and contribute across teams. Ownership and impact: Take full ownership of the AI solutions you develop, ensuring they are not only innovative but also scalable, maintainable, and aligned with real user workflows. What we're looking for: Strong engineering skills: Solid experience building production-grade, user-facing software systems. You’re a self-starter, capable of tackling complex engineering problems and making UI design decisions with minimal supervision. AI experience with a practical mindset: You’re familiar with AI technologies and frameworks, and you focus on delivering high-quality solutions that work in the real world, not just in theory. Quick iteration and experimentation: You’re comfortable releasing prototypes, collecting feedback, and iterating with a pragmatic mindset. Proven initiative: You take ownership and drive projects forward, pushing boundaries to find the most impactful solutions. You can deal with ambiguity and are able to define scope where things are loosely defined. Collaborative attitude: You communicate effectively with your peers. You’re open to feedback, and you bring a solutions-oriented mindset to the table. Requirements: Experience with LLMs, prompt engineering, and building applications powered by GenAI. Proven track record of delivering software that made it into production and is actively used by users. Exposure to working in cloud-native environments (e.g., AWS, GCP, Azure). Experience using observability tools to understand and troubleshoot system behavior. Nice to have: Experience building or working with agent frameworks or multi‑agent workflows. Experience as a data analyst or work with data platforms (e.g., Looker, Tableau, PowerBI, Snowflake, DataBricks) Experience building tools for data engineering. In the US, the Base compensation range for this role is $174,986 - $209,983. Actual compensation may vary based on level, experience, and skillset as assessed in the interview process. Benefits include equity, bonus (if applicable) and other benefits listed here. #LI-Remote *Compensation ranges are country specific. If you are applying for this role from a different location than listed above, your recruiter will discuss your specific market’s defined pay range & benefits at the beginning of the process. Why You’ll Thrive at Grafana Labs: 100% Remote, Global Culture - As a remote-only company, we bring together talent from around the world, united by a culture of collaboration and shared purpose. Scaling Organization – Tackle meaningful work in a high-growth, ever-evolving environment. Transparent Communication – Expect open decision-making and regular company-wide updates. Innovation-Driven – Autonomy and support to ship great work and try new things. Open Source Roots – Built on community-driven values that shape how we work. Empowered Teams – High trust, low ego culture that values outcomes over optics. Career Growth Pathways – Defined opportunities to grow and develop your career. Approachable Leadership – Transparent execs who are involved, visible, and human. Passionate People – Join a team of smart, supportive folks who care deeply about what they do. In-Person onboarding - We want you to thrive from day 1 with your fellow new ‘Grafanistas’ to learn all about what we do and how we do it. Balance is Key - We operate a global annual leave policy of 30 days per annum. 3 days of your annual leave entitlement are reserved for Grafana Shutdown Days to allow the team to really disconnect. *We will comply with local legislation where applicable. Equal Opportunity Employer: Grafana Labs is an equal opportunities employer. We welcome applications from everyone regardless of race, colour, nationality, origin, caste, sex, gender reassignment identity or expression, sexual orientation, age, religion or belief, disability, veteran status, genetic information, pregnancy, maternity, marital, family or carer status, or any other characteristic which is protected by local law. We believe that equality and diversity build a strong organisation, and we work hard to ensure that is the foundation of our organisation as we grow. Grafana Labs may utilize AI tools in its recruitment process to assist in matching information provided in CVs to job postings. The recruitment team will continue to review inbound CVs manually to identify alignment with current openings. #LI-Remote For information about how your personal data is used once you’ve applied to a job, check out our privacy policy.

API Design
System Architecture
Observability & Monitoring
Direct Apply
Posted 23 days ago
Grafana Labs

Senior Solutions Engineer | West Coast | Remote

Grafana LabsAnywhereFull-time
View Job
Compensation$70K - 120K a year

Drive technical engagements as a customer-facing product expert to support sales and provide feedback to product management. | Requires 5+ years of technical pre-sales experience, strong communication, and ability to collaborate remotely in an international environment. | Grafana Labs, the company behind the open observability cloud, is founded on the principles of open source, open standards, open ecosystems, and open culture. Grafana Cloud, our fully managed observability platform, is flexible and built for scale. With Grafana Cloud's actually useful AI, organizations can see, understand, and act on all their disparate data to move at the speed of their ambitions. Today, more than 35 million users and 7,000+ customers – including Anthropic, Bloomberg, NVIDIA, Microsoft, and Salesforce – trust Grafana Labs to ensure reliability of their applications and systems, resolve incidents quickly, and optimize their telemetry to reduce noise and cost. We are a 100% remote company with 1,600+ team members across 40+ countries, and we’re backed by leading investors including Lightspeed Venture Partners, Sequoia Capital, GIC, Coatue, J.P. Morgan, CapitalG, and Lead Edge Capital. Learn more at grafana.com and follow us on LinkedIn and X. We’re scaling fast and staying true to what makes us different: an open-source legacy, a global collaborative culture, and a passion for meaningful work. Our team thrives in an innovation-driven environment where transparency, autonomy, and trust fuel everything we do. You may not meet every requirement, and that’s okay. If this role excites you, we’d love you to raise your hand for what could be a truly career-defining opportunity. The Opportunity: Make a difference. Are you looking for something new in your career? Enjoy a challenge? Love to innovate and learn? This is an exciting opportunity to join a technology startup in a high growth phase. Our Enterprise Solutions Engineering team are our in-house, customer-facing product experts. The Enterprise SE team enables Grafana Labs worldwide growth by educating potential and existing customers to ensure they are happy and successful. We share our technical and product expertise with customers through demos, hands-on enablement, presentations, technical evaluations and ongoing interaction. We also partner closely with our Sales team to help qualify and close opportunities. The Enterprise SE team is also one of the primary routes of communication from our customers through into the Product Management and Engineering teams. What You’ll Be Doing: Join as an early member of a high growth team working together to solve technical and customer problems the right way Partner with our Sales team to articulate the overall Grafana value proposition, vision, and strategy to customers Own the technical engagement with customers and help to close high-velocity opportunities through advanced competitive knowledge, technical skill, and credibility Deliver product and technical presentations to potential and existing customers Proactively communicate with customers and internal teams to provide a feedback loop on our products and the competitive landscape Drive product conversations based on need and problems learned during customer interactions Work with the team to enhance documentation, write blog posts, record videos, contribute knowledge base articles, and create other public and internal enablement material What Makes You a Great Fit: Ideally located on the West Coast of the US 5+ years of technical pre-sales experience, ideally with Open Source technologies, or in the Metrics/Monitoring space while we will be adding more junior folks later on, right now we just need team members that can hit the ground running. We’re a startup so your job duties will be varied and complex and will require strong judgment, collaboration, and leadership We are a remote-first company so you would need to learn to work and collaborate with an international team You will need first class written and oral communication skills to collaborate with our remote-first internal teams and with our worldwide customers. You will need to be able to skillfully articulate our value proposition and the technical advantages of our products You will love solving technical challenges and thrive on bringing creative solutions to our customers You should have a technical mindset and a desire to grow technically. Great candidates may be a recent grad of a coding bootcamp or have previous post-sales or engineering experience–we are excited to see what unique experiences and skill sets you bring to the table! You are self-motivated, initiative, creative, and maybe you’ve already done a brave thing or two in your life! Compensation & Rewards: In the United States the OTE (On-Target Earnings) compensation range for this role is $204,000 - $254,000. Actual compensation may vary based on level, experience, and skillset as assessed throughout the interview process. All of our roles include Restricted Stock Units (RSUs), giving every team member ownership in Grafana Labs' success. We believe in shared outcomes—RSUs help us stay aligned and invested as we scale globally. *Compensation ranges are country specific. If you are applying for this role from a different location than listed above, your recruiter will discuss your specific market’s defined pay range & benefits at the beginning of the process. Why You’ll Thrive at Grafana Labs: 100% Remote, Global Culture - As a remote-only company, we bring together talent from around the world, united by a culture of collaboration and shared purpose. Scaling Organization – Tackle meaningful work in a high-growth, ever-evolving environment. Transparent Communication – Expect open decision-making and regular company-wide updates. Innovation-Driven – Autonomy and support to ship great work and try new things. Open Source Roots – Built on community-driven values that shape how we work. Empowered Teams – High trust, low ego culture that values outcomes over optics. Career Growth Pathways – Defined opportunities to grow and develop your career. Approachable Leadership – Transparent execs who are involved, visible, and human. Passionate People – Join a team of smart, supportive folks who care deeply about what they do. In-Person onboarding - We want you to thrive from day 1 with your fellow new ‘Grafanistas’ to learn all about what we do and how we do it. Balance is Key - We operate a global annual leave policy of 30 days per annum. 3 days of your annual leave entitlement are reserved for Grafana Shutdown Days to allow the team to really disconnect. *We will comply with local legislation where applicable. Equal Opportunity Employer: Grafana Labs is an equal opportunities employer. We welcome applications from everyone regardless of race, colour, nationality, origin, caste, sex, gender reassignment identity or expression, sexual orientation, age, religion or belief, disability, veteran status, genetic information, pregnancy, maternity, marital, family or carer status, or any other characteristic which is protected by local law. We believe that equality and diversity build a strong organisation, and we work hard to ensure that is the foundation of our organisation as we grow. Grafana Labs may utilize AI tools in its recruitment process to assist in matching information provided in CVs to job postings. The recruitment team will continue to review inbound CVs manually to identify alignment with current openings. #LI-Remote For information about how your personal data is used once you’ve applied to a job, check out our privacy policy.

Technical Pre-sales
Open Source Technologies
Metrics and Monitoring
Direct Apply
Posted 27 days ago
GL

Senior Solutions Engineer | West Coast | Remote

Grafana LabsAnywhereFull-time
View Job
Compensation$70K - 120K a year

Lead technical engagements with sales to close opportunities and create enablement materials. | 5+ years technical pre-sales experience with strong communication skills in a remote environment. | Grafana Labs, the company behind the open observability cloud, is founded on the principles of open source, open standards, open ecosystems, and open culture. Grafana Cloud, our fully managed observability platform, is flexible and built for scale. With Grafana Cloud's actually useful AI, organizations can see, understand, and act on all their disparate data to move at the speed of their ambitions. Today, more than 35 million users and 7,000+ customers – including Anthropic, Bloomberg, NVIDIA, Microsoft, and Salesforce – trust Grafana Labs to ensure reliability of their applications and systems, resolve incidents quickly, and optimize their telemetry to reduce noise and cost. We are a 100% remote company with 1,600+ team members across 40+ countries, and we’re backed by leading investors including Lightspeed Venture Partners, Sequoia Capital, GIC, Coatue, J.P. Morgan, CapitalG, and Lead Edge Capital. Learn more at grafana.com and follow us on LinkedIn and X. We’re scaling fast and staying true to what makes us different: an open-source legacy, a global collaborative culture, and a passion for meaningful work. Our team thrives in an innovation-driven environment where transparency, autonomy, and trust fuel everything we do. You may not meet every requirement, and that’s okay. If this role excites you, we’d love you to raise your hand for what could be a truly career-defining opportunity. The Opportunity: Make a difference. Are you looking for something new in your career? Enjoy a challenge? Love to innovate and learn? This is an exciting opportunity to join a technology startup in a high growth phase. Our Enterprise Solutions Engineering team are our in-house, customer-facing product experts. The Enterprise SE team enables Grafana Labs worldwide growth by educating potential and existing customers to ensure they are happy and successful. We share our technical and product expertise with customers through demos, hands-on enablement, presentations, technical evaluations and ongoing interaction. We also partner closely with our Sales team to help qualify and close opportunities. The Commercial SE team is also one of the primary routes of communication from our customers through into the Product Management and Engineering teams. What You’ll Be Doing: Join as an early member of a high growth team working together to solve technical and customer problems the right way Partner with our Sales team to articulate the overall Grafana value proposition, vision, and strategy to customers Own the technical engagement with customers and help to close high-velocity opportunities through advanced competitive knowledge, technical skill, and credibility Deliver product and technical presentations to potential and existing customers Proactively communicate with customers and internal teams to provide a feedback loop on our products and the competitive landscape Drive product conversations based on need and problems learned during customer interactions Work with the team to enhance documentation, write blog posts, record videos, contribute knowledge base articles, and create other public and internal enablement material What Makes You a Great Fit: Ideally located on the West Coast of the US 5+ years of technical pre-sales experience, ideally with Open Source technologies, or in the Metrics/Monitoring space while we will be adding more junior folks later on, right now we just need team members that can hit the ground running. We’re a startup so your job duties will be varied and complex and will require strong judgment, collaboration, and leadership We are a remote-first company so you would need to learn to work and collaborate with an international team You will need first class written and oral communication skills to collaborate with our remote-first internal teams and with our worldwide customers. You will need to be able to skillfully articulate our value proposition and the technical advantages of our products You will love solving technical challenges and thrive on bringing creative solutions to our customers You should have a technical mindset and a desire to grow technically. Great candidates may be a recent grad of a coding bootcamp or have previous post-sales or engineering experience–we are excited to see what unique experiences and skill sets you bring to the table! You are self-motivated, initiative, creative, and maybe you’ve already done a brave thing or two in your life! Compensation & Rewards: In the United States the OTE (On-Target Earnings) compensation range for this role is $204,000 - $254,000. Actual compensation may vary based on level, experience, and skillset as assessed throughout the interview process. All of our roles include Restricted Stock Units (RSUs), giving every team member ownership in Grafana Labs' success. We believe in shared outcomes—RSUs help us stay aligned and invested as we scale globally. *Compensation ranges are country specific. If you are applying for this role from a different location than listed above, your recruiter will discuss your specific market’s defined pay range & benefits at the beginning of the process. Why You’ll Thrive at Grafana Labs: 100% Remote, Global Culture - As a remote-only company, we bring together talent from around the world, united by a culture of collaboration and shared purpose. Scaling Organization – Tackle meaningful work in a high-growth, ever-evolving environment. Transparent Communication – Expect open decision-making and regular company-wide updates. Innovation-Driven – Autonomy and support to ship great work and try new things. Open Source Roots – Built on community-driven values that shape how we work. Empowered Teams – High trust, low ego culture that values outcomes over optics. Career Growth Pathways – Defined opportunities to grow and develop your career. Approachable Leadership – Transparent execs who are involved, visible, and human. Passionate People – Join a team of smart, supportive folks who care deeply about what they do. In-Person onboarding - We want you to thrive from day 1 with your fellow new ‘Grafanistas’ to learn all about what we do and how we do it. Balance is Key - We operate a global annual leave policy of 30 days per annum. 3 days of your annual leave entitlement are reserved for Grafana Shutdown Days to allow the team to really disconnect. *We will comply with local legislation where applicable. Equal Opportunity Employer: Grafana Labs is an equal opportunities employer. We welcome applications from everyone regardless of race, colour, nationality, origin, caste, sex, gender reassignment identity or expression, sexual orientation, age, religion or belief, disability, veteran status, genetic information, pregnancy, maternity, marital, family or carer status, or any other characteristic which is protected by local law. We believe that equality and diversity build a strong organisation, and we work hard to ensure that is the foundation of our organisation as we grow. Grafana Labs may utilize AI tools in its recruitment process to assist in matching information provided in CVs to job postings. The recruitment team will continue to review inbound CVs manually to identify alignment with current openings. #LI-Remote For information about how your personal data is used once you’ve applied to a job, check out our privacy policy.

Technical Pre-sales
Open Source Technologies
Metrics and Monitoring
Direct Apply
Posted about 1 month ago
GL

Senior Solutions Engineer | East Coast | Remote

Grafana LabsAnywhereFull-time
View Job
Compensation$70K - 120K a year

Partner with sales to deliver technical presentations and manage technical engagements to close deals. | 5+ years in technical pre-sales with strong communication and technical skills, ideally in open source or monitoring. | Grafana Labs, the company behind the open observability cloud, is founded on the principles of open source, open standards, open ecosystems, and open culture. Grafana Cloud, our fully managed observability platform, is flexible and built for scale. With Grafana Cloud's actually useful AI, organizations can see, understand, and act on all their disparate data to move at the speed of their ambitions. Today, more than 35 million users and 7,000+ customers – including Anthropic, Bloomberg, NVIDIA, Microsoft, and Salesforce – trust Grafana Labs to ensure reliability of their applications and systems, resolve incidents quickly, and optimize their telemetry to reduce noise and cost. We are a 100% remote company with 1,600+ team members across 40+ countries, and we’re backed by leading investors including Lightspeed Venture Partners, Sequoia Capital, GIC, Coatue, J.P. Morgan, CapitalG, and Lead Edge Capital. Learn more at grafana.com and follow us on LinkedIn and X. We’re scaling fast and staying true to what makes us different: an open-source legacy, a global collaborative culture, and a passion for meaningful work. Our team thrives in an innovation-driven environment where transparency, autonomy, and trust fuel everything we do. You may not meet every requirement, and that’s okay. If this role excites you, we’d love you to raise your hand for what could be a truly career-defining opportunity. The Opportunity: Make a difference. Are you looking for something new in your career? Enjoy a challenge? Love to innovate and learn? This is an exciting opportunity to join a technology startup in a high growth phase. Our Enterprise Solutions Engineering team are our in-house, customer-facing product experts. The Enterprise SE team enables Grafana Labs worldwide growth by educating potential and existing customers to ensure they are happy and successful. We share our technical and product expertise with customers through demos, hands-on enablement, presentations, technical evaluations and ongoing interaction. We also partner closely with our Sales team to help qualify and close opportunities. The Commercial SE team is also one of the primary routes of communication from our customers through into the Product Management and Engineering teams. What You’ll Be Doing: Join as an early member of a high growth team working together to solve technical and customer problems the right way Partner with our Sales team to articulate the overall Grafana value proposition, vision, and strategy to customers Own the technical engagement with customers and help to close high-velocity opportunities through advanced competitive knowledge, technical skill, and credibility Deliver product and technical presentations to potential and existing customers Proactively communicate with customers and internal teams to provide a feedback loop on our products and the competitive landscape Drive product conversations based on need and problems learned during customer interactions Work with the team to enhance documentation, write blog posts, record videos, contribute knowledge base articles, and create other public and internal enablement material What Makes You a Great Fit: Ideally located in Greater NYC or Boston Area 5+ years of technical pre-sales experience, ideally with Open Source technologies, or in the Metrics/Monitoring space while we will be adding more junior folks later on, right now we just need team members that can hit the ground running. We’re a startup so your job duties will be varied and complex and will require strong judgment, collaboration, and leadership We are a remote-first company so you would need to learn to work and collaborate with an international team You will need first class written and oral communication skills to collaborate with our remote-first internal teams and with our worldwide customers. You will need to be able to skillfully articulate our value proposition and the technical advantages of our products You will love solving technical challenges and thrive on bringing creative solutions to our customers You should have a technical mindset and a desire to grow technically. Great candidates may be a recent grad of a coding bootcamp or have previous post-sales or engineering experience–we are excited to see what unique experiences and skill sets you bring to the table! You are self-motivated, initiative, creative, and maybe you’ve already done a brave thing or two in your life! Compensation & Rewards: In the United States the OTE (On-Target Earnings) compensation range for this role is $204,000 - $254,000. Actual compensation may vary based on level, experience, and skillset as assessed throughout the interview process. All of our roles include Restricted Stock Units (RSUs), giving every team member ownership in Grafana Labs' success. We believe in shared outcomes—RSUs help us stay aligned and invested as we scale globally. *Compensation ranges are country specific. If you are applying for this role from a different location than listed above, your recruiter will discuss your specific market’s defined pay range & benefits at the beginning of the process. Why You’ll Thrive at Grafana Labs: 100% Remote, Global Culture - As a remote-only company, we bring together talent from around the world, united by a culture of collaboration and shared purpose. Scaling Organization – Tackle meaningful work in a high-growth, ever-evolving environment. Transparent Communication – Expect open decision-making and regular company-wide updates. Innovation-Driven – Autonomy and support to ship great work and try new things. Open Source Roots – Built on community-driven values that shape how we work. Empowered Teams – High trust, low ego culture that values outcomes over optics. Career Growth Pathways – Defined opportunities to grow and develop your career. Approachable Leadership – Transparent execs who are involved, visible, and human. Passionate People – Join a team of smart, supportive folks who care deeply about what they do. In-Person onboarding - We want you to thrive from day 1 with your fellow new ‘Grafanistas’ to learn all about what we do and how we do it. Balance is Key - We operate a global annual leave policy of 30 days per annum. 3 days of your annual leave entitlement are reserved for Grafana Shutdown Days to allow the team to really disconnect. *We will comply with local legislation where applicable. Equal Opportunity Employer: Grafana Labs is an equal opportunities employer. We welcome applications from everyone regardless of race, colour, nationality, origin, caste, sex, gender reassignment identity or expression, sexual orientation, age, religion or belief, disability, veteran status, genetic information, pregnancy, maternity, marital, family or carer status, or any other characteristic which is protected by local law. We believe that equality and diversity build a strong organisation, and we work hard to ensure that is the foundation of our organisation as we grow. Grafana Labs may utilize AI tools in its recruitment process to assist in matching information provided in CVs to job postings. The recruitment team will continue to review inbound CVs manually to identify alignment with current openings. #LI-Remote For information about how your personal data is used once you’ve applied to a job, check out our privacy policy.

Technical Presentations
Customer Relationship Management
Communication
Direct Apply
Posted about 1 month ago
GL

Senior Solutions Engineer | West Coast | Remote

Grafana LabsAnywhereFull-time
View Job
Compensation$85K - 130K a year

Partner with sales to deliver technical presentations and manage technical engagements to close opportunities. | 5+ years technical pre-sales experience with strong communication skills and ability to work autonomously in a remote environment. | Grafana Labs, the company behind the open observability cloud, is founded on the principles of open source, open standards, open ecosystems, and open culture. Grafana Cloud, our fully managed observability platform, is flexible and built for scale. With Grafana Cloud's actually useful AI, organizations can see, understand, and act on all their disparate data to move at the speed of their ambitions. Today, more than 35 million users and 7,000+ customers – including Anthropic, Bloomberg, NVIDIA, Microsoft, and Salesforce – trust Grafana Labs to ensure reliability of their applications and systems, resolve incidents quickly, and optimize their telemetry to reduce noise and cost. We are a 100% remote company with 1,600+ team members across 40+ countries, and we’re backed by leading investors including Lightspeed Venture Partners, Sequoia Capital, GIC, Coatue, J.P. Morgan, CapitalG, and Lead Edge Capital. Learn more at grafana.com and follow us on LinkedIn and X. We’re scaling fast and staying true to what makes us different: an open-source legacy, a global collaborative culture, and a passion for meaningful work. Our team thrives in an innovation-driven environment where transparency, autonomy, and trust fuel everything we do. You may not meet every requirement, and that’s okay. If this role excites you, we’d love you to raise your hand for what could be a truly career-defining opportunity. The Opportunity: Make a difference. Are you looking for something new in your career? Enjoy a challenge? Love to innovate and learn? This is an exciting opportunity to join a technology startup in a high growth phase. Our Enterprise Solutions Engineering team are our in-house, customer-facing product experts. The Enterprise SE team enables Grafana Labs worldwide growth by educating potential and existing customers to ensure they are happy and successful. We share our technical and product expertise with customers through demos, hands-on enablement, presentations, technical evaluations and ongoing interaction. We also partner closely with our Sales team to help qualify and close opportunities. The Commercial SE team is also one of the primary routes of communication from our customers through into the Product Management and Engineering teams. What You’ll Be Doing: Join as an early member of a high growth team working together to solve technical and customer problems the right way Partner with our Sales team to articulate the overall Grafana value proposition, vision, and strategy to customers Own the technical engagement with customers and help to close high-velocity opportunities through advanced competitive knowledge, technical skill, and credibility Deliver product and technical presentations to potential and existing customers Proactively communicate with customers and internal teams to provide a feedback loop on our products and the competitive landscape Drive product conversations based on need and problems learned during customer interactions Work with the team to enhance documentation, write blog posts, record videos, contribute knowledge base articles, and create other public and internal enablement material What Makes You a Great Fit: Ideally located on the West Coast of the US 5+ years of technical pre-sales experience, ideally with Open Source technologies, or in the Metrics/Monitoring space while we will be adding more junior folks later on, right now we just need team members that can hit the ground running. We’re a startup so your job duties will be varied and complex and will require strong judgment, collaboration, and leadership We are a remote-first company so you would need to learn to work and collaborate with an international team You will need first class written and oral communication skills to collaborate with our remote-first internal teams and with our worldwide customers. You will need to be able to skillfully articulate our value proposition and the technical advantages of our products You will love solving technical challenges and thrive on bringing creative solutions to our customers You should have a technical mindset and a desire to grow technically. Great candidates may be a recent grad of a coding bootcamp or have previous post-sales or engineering experience–we are excited to see what unique experiences and skill sets you bring to the table! You are self-motivated, initiative, creative, and maybe you’ve already done a brave thing or two in your life! Compensation & Rewards: In the United States the OTE (On-Target Earnings) compensation range for this role is $204,000 - $254,000. Actual compensation may vary based on level, experience, and skillset as assessed throughout the interview process. All of our roles include Restricted Stock Units (RSUs), giving every team member ownership in Grafana Labs' success. We believe in shared outcomes—RSUs help us stay aligned and invested as we scale globally. *Compensation ranges are country specific. If you are applying for this role from a different location than listed above, your recruiter will discuss your specific market’s defined pay range & benefits at the beginning of the process. Why You’ll Thrive at Grafana Labs: 100% Remote, Global Culture - As a remote-only company, we bring together talent from around the world, united by a culture of collaboration and shared purpose. Scaling Organization – Tackle meaningful work in a high-growth, ever-evolving environment. Transparent Communication – Expect open decision-making and regular company-wide updates. Innovation-Driven – Autonomy and support to ship great work and try new things. Open Source Roots – Built on community-driven values that shape how we work. Empowered Teams – High trust, low ego culture that values outcomes over optics. Career Growth Pathways – Defined opportunities to grow and develop your career. Approachable Leadership – Transparent execs who are involved, visible, and human. Passionate People – Join a team of smart, supportive folks who care deeply about what they do. In-Person onboarding - We want you to thrive from day 1 with your fellow new ‘Grafanistas’ to learn all about what we do and how we do it. Balance is Key - We operate a global annual leave policy of 30 days per annum. 3 days of your annual leave entitlement are reserved for Grafana Shutdown Days to allow the team to really disconnect. *We will comply with local legislation where applicable. Equal Opportunity Employer: Grafana Labs is an equal opportunities employer. We welcome applications from everyone regardless of race, colour, nationality, origin, caste, sex, gender reassignment identity or expression, sexual orientation, age, religion or belief, disability, veteran status, genetic information, pregnancy, maternity, marital, family or carer status, or any other characteristic which is protected by local law. We believe that equality and diversity build a strong organisation, and we work hard to ensure that is the foundation of our organisation as we grow. Grafana Labs may utilize AI tools in its recruitment process to assist in matching information provided in CVs to job postings. The recruitment team will continue to review inbound CVs manually to identify alignment with current openings. #LI-Remote For information about how your personal data is used once you’ve applied to a job, check out our privacy policy.

Technical Pre-sales
Open Source Technologies
Metrics and Monitoring
Direct Apply
Posted about 1 month ago
GL

Senior AI Engineer | US | Remote

Grafana LabsAnywhereFull-time
View Job
Compensation$90K - 150K a year

Design and deploy scalable backend services and AI agent infrastructure integrating multi-agent architectures and LLMs. | 8+ years software engineering with 2+ years applying LLMs, proficiency in Python, Node.js, GCP, and expertise in agentic systems and RAG. | Grafana Labs, the company behind the open observability cloud, is founded on the principles of open source, open standards, open ecosystems, and open culture. Grafana Cloud, our fully managed observability platform, is flexible and built for scale. With Grafana Cloud's actually useful AI, organizations can see, understand, and act on all their disparate data to move at the speed of their ambitions. Today, more than 35 million users and 7,000+ customers – including Anthropic, Bloomberg, NVIDIA, Microsoft, and Salesforce – trust Grafana Labs to ensure reliability of their applications and systems, resolve incidents quickly, and optimize their telemetry to reduce noise and cost. We are a 100% remote company with 1,600+ team members across 40+ countries, and we’re backed by leading investors including Lightspeed Venture Partners, Sequoia Capital, GIC, Coatue, J.P. Morgan, CapitalG, and Lead Edge Capital. Learn more at grafana.com and follow us on LinkedIn and X. We’re scaling fast and staying true to what makes us different: an open-source legacy, a global collaborative culture, and a passion for meaningful work. Our team thrives in an innovation-driven environment where transparency, autonomy, and trust fuel everything we do. You may not meet every requirement, and that’s okay. If this role excites you, we’d love you to raise your hand for what could be a truly career-defining opportunity. This is a remote opportunity and we are looking for candidates from the U.S. The Opportunity Grafana Labs is seeking a Senior Engineer (AI & Automation) to own the AI agent infrastructure and automation platform that powers our Marketing Operations organization. You’ll build multi-agent architectures, LLM integrations, and backend services that connect AI models to internal and third-party data platforms. You’ll ship production systems that teams depend on daily. This is a high-autonomy role where you own the technical direction. You’ll identify the highest-leverage problems across Marketing, RevOps, and SDR teams, design the solutions, and ship them. You’ll define the technical direction for the automation platform (data models, API contracts, shared libraries, reference architectures) and partner with Data Engineering, GTM Systems, and Field Operations to build scalable, self-service automation that eliminates manual work and drives operational efficiency. What You'll Be Doing Agentic Systems & AI Infrastructure Own end-to-end development of multi-agent AI systems, from architecture and implementation through testing, deployment, and ongoing operation Build modular, composable agentic systems using orchestration frameworks (LangChain, CrewAI, Anthropic MCP, or similar) that operate 24/7 across teams Develop reusable agentic skills that agents invoke across interfaces (Slack, dashboards, internal apps, CLIs) Implement observability and feedback loops including logging, performance metrics, prompt iteration, model evaluation, and cost management Establish governance and compliance standards for AI workflows including access controls, audit trails, PII handling, and human-in-the-loop escalation paths Systems Integration & Backend Services Build MCP servers, APIs, CLIs, and microservices connecting AI models to business systems (BigQuery, Slack, CRMs, email, calendars, analytics tools) Architect data flows for retrieval-augmented generation (RAG), connecting LLMs to internal knowledge bases, customer data, and real-time business context Build serverless or containerized services (GCP Cloud Functions, Cloud Run) that scale with usage and integrate with Grafana's cloud infrastructure Automation & Workflow Enablement Partner with RevOps, Demand Generation, Regional Marketing, and SDR teams to scope high-impact automation problems, identify bottlenecks, and build solutions with measurable business outcomes Design and deploy workflows using orchestration tools (n8n, Workato, or custom platforms) with CI/CD, testing, and production reliability standards Build systems designed for self-service with documentation, playbooks, and enablement materials that let partner teams operate independently We invest heavily in developer productivity. You'll have access to AI coding assistants (Claude Code, Gemini CLI, OpenAI Codex, and others of your choice within security guidelines). We encourage pragmatic AI-assisted development paired with strong code review and quality standards. What Makes You a Great Fit 8+ years of software engineering experience with depth in backend development, systems integration, or data/analytics engineering 2+ years hands-on experience applying LLMs/AI to production workflows, not just prototypes Strong proficiency in Python and JavaScript/Node.js with Git-based workflows, code review practices, and testing discipline Hands-on experience with LLM frameworks and patterns including prompt engineering, RAG, function calling/tool use, structured output parsing, and evaluation Experience building and operating multi-agent systems at scale including agent decomposition, orchestration patterns (sequential chains, router/dispatcher, parallel fan-out), state management, and production monitoring You diagnose business problems before writing code. You think in workflows and outcomes, not just functions. Deep familiarity with Google Cloud Platform, BigQuery, and serverless/containerized services (Cloud Functions, Cloud Run) Understanding of LLM failure modes and production mitigations including confidence thresholds, fallback logic, human escalation, and cost/latency management Proven ability to identify high-leverage problems, push back on low-impact requests, and deliver end-to-end with minimal direction Fluent with AI-assisted development tools (GitHub Copilot, Cursor, Claude Code). You use AI to build AI systems Clear technical communicator who can explain complex systems in simple terms to both engineers and business stakeholders Bonus Points Experience with vector databases or retrieval pipelines (Pinecone, Weaviate, ChromaDB, Qdrant, pgvector) Familiarity with marketing or sales platforms (Salesforce, Customer.io, HubSpot, Marketo, Outreach) Experience with frontend frameworks (React, Slack Block Kit) for building user-facing AI tool interfaces Observability tooling for AI systems (LangSmith, Weights & Biases, custom evaluation frameworks) Experience with workflow orchestration platforms (n8n, Temporal, Prefect, Airflow) Familiarity with Model Context Protocol (MCP) or similar standards for connecting AI systems to data sources Prior work automating marketing, sales, or customer success workflows in a B2B SaaS environment Active in open-source communities. Grafana is built on OSS and we value engineers who share that DNA In the United States, the base compensation range for this role is USD $154,445 - USD $185,334. Actual compensation may vary based on level, experience, and skillset as assessed throughout the interview process. All of our roles include Restricted Stock Units (RSUs), giving every team member ownership in Grafana Labs' success. We believe in shared outcomes—RSUs help us stay aligned and invested as we scale globally. *Compensation ranges are country specific. If you are applying for this role from a different location than listed above, your recruiter will discuss your specific market’s defined pay range & benefits at the beginning of the process. Why You’ll Thrive at Grafana Labs: 100% Remote, Global Culture - As a remote-only company, we bring together talent from around the world, united by a culture of collaboration and shared purpose. Scaling Organization – Tackle meaningful work in a high-growth, ever-evolving environment. Transparent Communication – Expect open decision-making and regular company-wide updates. Innovation-Driven – Autonomy and support to ship great work and try new things. Open Source Roots – Built on community-driven values that shape how we work. Empowered Teams – High trust, low ego culture that values outcomes over optics. Career Growth Pathways – Defined opportunities to grow and develop your career. Approachable Leadership – Transparent execs who are involved, visible, and human. Passionate People – Join a team of smart, supportive folks who care deeply about what they do. In-Person onboarding - We want you to thrive from day 1 with your fellow new ‘Grafanistas’ to learn all about what we do and how we do it. Balance is Key - We operate a global annual leave policy of 30 days per annum. 3 days of your annual leave entitlement are reserved for Grafana Shutdown Days to allow the team to really disconnect. *We will comply with local legislation where applicable. Equal Opportunity Employer: Grafana Labs is an equal opportunities employer. We welcome applications from everyone regardless of race, colour, nationality, origin, caste, sex, gender reassignment identity or expression, sexual orientation, age, religion or belief, disability, veteran status, genetic information, pregnancy, maternity, marital, family or carer status, or any other characteristic which is protected by local law. We believe that equality and diversity build a strong organisation, and we work hard to ensure that is the foundation of our organisation as we grow. Grafana Labs may utilize AI tools in its recruitment process to assist in matching information provided in CVs to job postings. The recruitment team will continue to review inbound CVs manually to identify alignment with current openings. #LI-Remote For information about how your personal data is used once you’ve applied to a job, check out our privacy policy.

API Design
System Architecture
JavaScript
TypeScript
Observability
Direct Apply
Posted about 2 months ago
Grafana Labs

Staff Backend Engineer - Grafana Enterprise | US | Remote

Grafana LabsAnywhereFull-time
View Job
Compensation$175K - 210K a year

Design and build scalable backend systems for observability platform. | Senior-level experience with backend services, distributed systems, API design, and large-scale infrastructure. | This is a remote position. We are looking for candidates in the US and Canada. The Opportunity: Grafana Enterprise is our composable observability platform focused on large-scale operators with security and regulatory requirements who need to self-operate their observability solutions. Grafana Enterprise features are also part of our Grafana Cloud service. Grafana Cloud is an integrated suite of observability applications offered as a service. It allows our customers to leverage the best open source observability software, including Prometheus, Mimir, Loki, and Tempo, without the overhead of installing, maintaining, and scaling their own observability stack. The Grafana Enterprise team is a diverse team of expert technologists working remotely from around the globe. The team focuses on innovative solutions to large-scale customer challenges and developing solutions that improve security, robustness, flexibility, multitenancy, and interoperability on behalf of our customers. We work across the larger Grafana Labs organisation and directly with external customers, including major cloud service providers and global enterprises running observability and build infrastructure at massive scale, to solve their most demanding distributed systems challenges. Our customers include software engineers, site reliability engineers, and platform operators at medium, large, and global enterprises. As a company we are remote-only and global; we embrace people of different experiences and backgrounds to build diverse teams where every person brings a new perspective to the software. Our meetings tend to happen between 14:00 and 17:00 UTC, but we are flexible and responsive to the needs of our engineers and our customers. We measure ourselves by our openness, helpfulness, and our shared successes. We are looking for experienced software engineers who are passionate about distributed systems and building reliable, scalable backend infrastructure to join our growing team. Our backend is Go, and we build tools for people operating some of the world's largest observability and software delivery stacks. Backend Engineers at Grafana also contribute to open source communities. What You'll Be Doing: • Earning the trust of our large-scale operator customers to further Grafana's "big tent" philosophy of data accessibility and to meet clear business objectives • Designing and leading the development of backend services, distributed systems, and enterprise features at scale • Driving continuous improvement of our engineering culture through words and actions • Driving projects from initial ideation through the development lifecycle to production • Contributing to the scalability, reliability, security, and multi-tenancy of the Grafana platform trusted by some of the world's largest operators • Owning the operational health of our platform by participating in weekday 12h x 5d and separate weekend 24h x 2d on-call rotations. (Yes, we prioritize ops load reduction.) • Hiring and developing the best engineers to build the future of Grafana • Developing your skills as a thought leader to drive continuous improvement of engineering and operational practices across Grafana Labs • As we are remote-first, we provide guidance and meet regularly using video calls, so strong teamwork and excellent written and interpersonal communication skills are a must. We invest heavily in developer productivity. You can use modern AI coding assistants as part of your daily workflow (your choice of tools, within security guidelines), backed by a company-funded usage budget so you can iterate quickly without unnecessary friction. We encourage pragmatic AI-assisted development: faster prototyping, test generation, refactors, documentation, and incident follow-ups-always paired with strong code review and quality standards. You'll also have access to frontier models (e.g., GPT-Codex 5/3, Claude Opus 4.7, Gemini 3 Pro). What Makes You a Great Fit: • You work well as a communicative member of a team of engineering professionals. • You earn trust by saying what you mean and doing what you say. • You are customer focused and especially attuned to the needs of large-scale operators who rely on Grafana as critical infrastructure. You start with their needs and work backwards. • You insist on the highest standards and work to develop the skills and knowledge of your fellow team members. • You take on complex distributed systems challenges, break them down into digestible problems, and leverage your team and organization to deliver. • You design modular solutions, deliver minimum loveable products, gather data and feedback, and then progress iteratively. What will you be doing? (Role specifics) As a Staff Backend Engineer, you will design and build the backend systems powering Grafana Enterprise - the platform trusted by the world's largest operators to run their observability and software delivery infrastructure. • Hiring and developing the best engineers we can to deliver the future of Observability. • Architecting and implementing distributed backend services in Go, with a focus on correctness, observability, and performance at scale • Designing APIs and service contracts used by thousands of enterprise operators and cloud service providers • Collaborating with Product and UX to shape features and partnering with frontend engineers to ship complete, end-to-end solutions • Driving scalability and reliability improvements that matter to large-scale operators running Grafana in regulated, high-availability environments • Engaging directly with large enterprise customers and cloud service providers to understand their requirements and translate them into robust engineering solutions • Advocating for our customers at every stage of the development lifecycle What technical competencies are we looking for? • Deep professional experience writing production services, from ideation through to production operations at scale • Strong distributed systems fundamentals: replication, consistency models, partitioning, fault tolerance, and the trade-offs that come with operating at scale • Demonstrated experience designing and operating systems for large-scale, high-traffic, high-availability, or multi-tenant environments, ideally in the context of infrastructure, observability, or software delivery platforms • Professional experience building and consuming gRPC/protobuf APIs and designing clean service contracts across service boundaries • Strong database skills, such as PostgreSQL and/or MySQL; including schema design, query optimisation, and schema migrations at scale • Experience with large-scale CI/CD systems and build tooling, designing, operating, or integrating with continuous delivery pipelines that serve large engineering organisations or external operators at scale • Comfort working with Kubernetes and containerised deployment environments, including patterns for operating stateful workloads and multi-tenant clusters • Experience with observability tooling: OpenTelemetry, Prometheus metrics, structured logging, and distributed tracing • Familiarity with dependency injection patterns (e.g., Google Wire) and clean, testable service architecture What are your "nice to have" technical competencies? • Experience with TypeScript and React for contributing to frontend features and collaborating closely with frontend engineers • Experience with Grafana's LGTM+ observability stack (Loki, Mimir, Tempo, Pyroscope, Alloy) • Prior experience at or building for large-scale cloud service providers, IaaS providers, or global enterprises with demanding SLA requirements • Experience designing or operating large-scale build infrastructure artifact registries, distributed build caches, hermetic build systems (e.g., Bazel), or developer platform tooling In the United States, the compensation range for this role is $174,986 - $209,983 USD. Actual compensation may vary based on level, experience, and skillset as assessed throughout the interview process. All of our roles include Restricted Stock Units (RSUs), giving every team member ownership in Grafana Labs' success. We believe in shared outcomes-RSUs help us stay aligned and invested as we scale globally. *Compensation ranges are country specific. If you are applying for this role from a different location than listed above, your recruiter will discuss your specific market's defined pay range & benefits at the beginning of the process.

API Design
Backend Development
Distributed Systems
Verified Source
Posted 2 months ago
GL

Staff Backend Engineer - Grafana Enterprise | US | Remote

Grafana LabsAnywhereFull-time
View Job
Compensation$120K - 160K a year

Design and build scalable backend services and distributed systems while mentoring engineers and collaborating with Product and UX teams. | Requires deep experience in production services at scale, distributed systems, database design, and proficiency with gRPC, Kubernetes, and observability tooling. | Grafana Labs is a remote-first, open-source powerhouse. There are more than 20M users of Grafana, the open source visualization tool, around the globe, monitoring everything from beehives to climate change in the Alps. The instantly recognizable dashboards have been spotted everywhere from a NASA launch and Minecraft HQ to Wimbledon and the Tour de France. Grafana Labs also helps more than 3,000 companies -- including Bloomberg, JPMorgan Chase, and eBay -- manage their observability strategies with the Grafana LGTM Stack, which can be run fully managed with Grafana Cloud or self-managed with the Grafana Enterprise Stack, both featuring scalable metrics (Grafana Mimir), logs (Grafana Loki), and traces (Grafana Tempo). We’re scaling fast and staying true to what makes us different: an open-source legacy, a global collaborative culture, and a passion for meaningful work. Our team thrives in an innovation-driven environment where transparency, autonomy, and trust fuel everything we do. You may not meet every requirement, and that’s okay. If this role excites you, we’d love you to raise your hand for what could be a truly career-defining opportunity. This is a remote position. We are looking for candidates in the US and Canada. The Opportunity: Grafana Enterprise is our composable observability platform focused on large-scale operators with security and regulatory requirements who need to self-operate their observability solutions. Grafana Enterprise features are also part of our Grafana Cloud service. Grafana Cloud is an integrated suite of observability applications offered as a service. It allows our customers to leverage the best open source observability software, including Prometheus, Mimir, Loki, and Tempo, without the overhead of installing, maintaining, and scaling their own observability stack. The Grafana Enterprise team is a diverse team of expert technologists working remotely from around the globe. The team focuses on innovative solutions to large-scale customer challenges and developing solutions that improve security, robustness, flexibility, multitenancy, and interoperability on behalf of our customers. We work across the larger Grafana Labs organisation and directly with external customers, including major cloud service providers and global enterprises running observability and build infrastructure at massive scale, to solve their most demanding distributed systems challenges. Our customers include software engineers, site reliability engineers, and platform operators at medium, large, and global enterprises. As a company we are remote-only and global; we embrace people of different experiences and backgrounds to build diverse teams where every person brings a new perspective to the software. Our meetings tend to happen between 14:00 and 17:00 UTC, but we are flexible and responsive to the needs of our engineers and our customers. We measure ourselves by our openness, helpfulness, and our shared successes. We are looking for experienced software engineers who are passionate about distributed systems and building reliable, scalable backend infrastructure to join our growing team. Our backend is Go, and we build tools for people operating some of the world's largest observability and software delivery stacks. Backend Engineers at Grafana also contribute to open source communities. What You’ll Be Doing: Earning the trust of our large-scale operator customers to further Grafana's "big tent" philosophy of data accessibility and to meet clear business objectives Designing and leading the development of backend services, distributed systems, and enterprise features at scale Driving continuous improvement of our engineering culture through words and actions Driving projects from initial ideation through the development lifecycle to production Contributing to the scalability, reliability, security, and multi-tenancy of the Grafana platform trusted by some of the world's largest operators Owning the operational health of our platform by participating in weekday 12h x 5d and separate weekend 24h x 2d on-call rotations. (Yes, we prioritize ops load reduction.) Hiring and developing the best engineers to build the future of Grafana Developing your skills as a thought leader to drive continuous improvement of engineering and operational practices across Grafana Labs As we are remote-first, we provide guidance and meet regularly using video calls, so strong teamwork and excellent written and interpersonal communication skills are a must. We invest heavily in developer productivity. You can use modern AI coding assistants as part of your daily workflow (your choice of tools, within security guidelines), backed by a company-funded usage budget so you can iterate quickly without unnecessary friction. We encourage pragmatic AI-assisted development: faster prototyping, test generation, refactors, documentation, and incident follow-ups—always paired with strong code review and quality standards. You’ll also have access to frontier models (e.g., GPT-Codex 5/3, Claude Opus 4.7, Gemini 3 Pro). What Makes You a Great Fit: You work well as a communicative member of a team of engineering professionals. You earn trust by saying what you mean and doing what you say. You are customer focused and especially attuned to the needs of large-scale operators who rely on Grafana as critical infrastructure. You start with their needs and work backwards. You insist on the highest standards and work to develop the skills and knowledge of your fellow team members. You take on complex distributed systems challenges, break them down into digestible problems, and leverage your team and organization to deliver. You design modular solutions, deliver minimum loveable products, gather data and feedback, and then progress iteratively. What will you be doing? (Role specifics) As a Staff Backend Engineer, you will design and build the backend systems powering Grafana Enterprise — the platform trusted by the world's largest operators to run their observability and software delivery infrastructure. Hiring and developing the best engineers we can to deliver the future of Observability. Architecting and implementing distributed backend services in Go, with a focus on correctness, observability, and performance at scale Designing APIs and service contracts used by thousands of enterprise operators and cloud service providers Collaborating with Product and UX to shape features and partnering with frontend engineers to ship complete, end-to-end solutions Driving scalability and reliability improvements that matter to large-scale operators running Grafana in regulated, high-availability environments Engaging directly with large enterprise customers and cloud service providers to understand their requirements and translate them into robust engineering solutions Advocating for our customers at every stage of the development lifecycle What technical competencies are we looking for? Deep professional experience writing production services, from ideation through to production operations at scale Strong distributed systems fundamentals: replication, consistency models, partitioning, fault tolerance, and the trade-offs that come with operating at scale Demonstrated experience designing and operating systems for large-scale, high-traffic, high-availability, or multi-tenant environments, ideally in the context of infrastructure, observability, or software delivery platforms Professional experience building and consuming gRPC/protobuf APIs and designing clean service contracts across service boundaries Strong database skills, such as PostgreSQL and/or MySQL; including schema design, query optimisation, and schema migrations at scale Experience with large-scale CI/CD systems and build tooling, designing, operating, or integrating with continuous delivery pipelines that serve large engineering organisations or external operators at scale Comfort working with Kubernetes and containerised deployment environments, including patterns for operating stateful workloads and multi-tenant clusters Experience with observability tooling: OpenTelemetry, Prometheus metrics, structured logging, and distributed tracing Familiarity with dependency injection patterns (e.g., Google Wire) and clean, testable service architecture What are your "nice to have" technical competencies? Experience with TypeScript and React for contributing to frontend features and collaborating closely with frontend engineers Experience with Grafana's LGTM+ observability stack (Loki, Mimir, Tempo, Pyroscope, Alloy) Prior experience at or building for large-scale cloud service providers, IaaS providers, or global enterprises with demanding SLA requirements Experience designing or operating large-scale build infrastructure artifact registries, distributed build caches, hermetic build systems (e.g., Bazel), or developer platform tooling In the United States, the compensation range for this role is $174,986 - $209,983 USD. Actual compensation may vary based on level, experience, and skillset as assessed throughout the interview process. All of our roles include Restricted Stock Units (RSUs), giving every team member ownership in Grafana Labs' success. We believe in shared outcomes—RSUs help us stay aligned and invested as we scale globally. *Compensation ranges are country specific. If you are applying for this role from a different location than listed above, your recruiter will discuss your specific market’s defined pay range & benefits at the beginning of the process. Why You’ll Thrive at Grafana Labs: 100% Remote, Global Culture - As a remote-only company, we bring together talent from around the world, united by a culture of collaboration and shared purpose. Scaling Organization – Tackle meaningful work in a high-growth, ever-evolving environment. Transparent Communication – Expect open decision-making and regular company-wide updates. Innovation-Driven – Autonomy and support to ship great work and try new things. Open Source Roots – Built on community-driven values that shape how we work. Empowered Teams – High trust, low ego culture that values outcomes over optics. Career Growth Pathways – Defined opportunities to grow and develop your career. Approachable Leadership – Transparent execs who are involved, visible, and human. Passionate People – Join a team of smart, supportive folks who care deeply about what they do. In-Person onboarding - We want you to thrive from day 1 with your fellow new ‘Grafanistas’ to learn all about what we do and how we do it. Balance is Key - We operate a global annual leave policy of 30 days per annum. 3 days of your annual leave entitlement are reserved for Grafana Shutdown Days to allow the team to really disconnect. *We will comply with local legislation where applicable. Equal Opportunity Employer: We will recruit, train, compensate and promote regardless of race, religion, color, national origin, gender, disability, age, veteran status, and all the other fascinating characteristics that make us different and unique. We believe that equality and diversity builds a strong organization and we’re working hard to make sure that’s the foundation of our organization as we grow. Grafana Labs may utilize AI tools in its recruitment process to assist in matching information provided in CVs to job postings. The recruitment team will continue to review inbound CVs manually to identify alignment with current openings. #LI-Remote For information about how your personal data is used once you’ve applied to a job, check out our privacy policy.

API Design
System Architecture
Backend Engineering
Direct Apply
Posted 2 months ago
Grafana Labs

Staff Backend Engineer - Grafana App Platform| US| Remote

Grafana LabsAnywhereFull-time
View Job
Compensation$175K - 210K a year

Develop and operate scalable SaaS backend features, contribute to design and roadmap, mentor team members, and participate in on-call rotations. | Strong coding and operational experience on SaaS platforms, excellent communication, teamwork, pragmatic approach, and willingness to learn Golang. | Grafana Labs is a remote-first, open-source powerhouse. There are more than 20M users of Grafana, the open source visualization tool, around the globe, monitoring everything from beehives to climate change in the Alps. The instantly recognizable dashboards have been spotted everywhere from a NASA launch and Minecraft HQ to Wimbledon and the Tour de France. Grafana Labs also helps more than 3,000 companies -- including Bloomberg, JPMorgan Chase, and eBay -- manage their observability strategies with the Grafana LGTM Stack, which can be run fully managed with Grafana Cloud or self-managed with the Grafana Enterprise Stack, both featuring scalable metrics (Grafana Mimir), logs (Grafana Loki), and traces (Grafana Tempo). We’re scaling fast and staying true to what makes us different: an open-source legacy, a global collaborative culture, and a passion for meaningful work. Our team thrives in an innovation-driven environment where transparency, autonomy, and trust fuel everything we do. You may not meet every requirement, and that’s okay. If this role excites you, we’d love you to raise your hand for what could be a truly career-defining opportunity. This is a remote position. We are looking for candidates in the US and Canada. The Opportunity: As our Cloud / Grafana-as-a-service business continues to grow, we've started to change Grafana's core architecture to be fully multi-tenant and scalable, as well as a solid platform for our opinionated Cloud apps. We're looking for someone with a lot of experience with SaaS platforms at scale. We are turning Grafana into a proper observability app platform where OSS and proprietary apps can directly tap into dashboards, alerts, incidents, and telemetry and deliver even more integrated experiences. To get there, we need to refactor a big part of Grafana so that it’s simpler and standardized. Grafana is used by countless OSS and Cloud users across different platforms, so planning and rolling out changes safely to avoid service disruptions is crucial; we are looking for someone who is excited about this sort of work. What You’ll Be Doing: • Coding new features, enhancing the operational experience of the systems you develop, and iteratively improving them based on production insights • Authoring, contributing to and reviewing design documents • Taking an active role in shaping the roadmap • Mentoring and supporting other team members and collaborating with different teams across the organization • Owning the experience of our customers by participating in weekday 12h x 5d and a separate weekend 24h x 2d on-call rotations. We invest heavily in developer productivity. You can use modern AI coding assistants as part of your daily workflow (your choice of tools, within security guidelines), backed by a company-funded usage budget so you can iterate quickly without unnecessary friction. We encourage pragmatic AI-assisted development: faster prototyping, test generation, refactors, documentation, and incident follow-ups—always paired with strong code review and quality standards. You’ll also have access to frontier models (e.g., GPT-Codex 5/3, Claude Opus 4.7, Gemini 3 Pro). What Makes You a Great Fit: • You have strong coding skills and operational experience; you were responsible for operating the software you have built. • You have worked on a SaaS platform and dealt with common distributed systems problems (e.g. scalability, multi-tenancy, data isolation, HA, …) • You have excellent communication skills. You can express your ideas with great clarity in writing and during meetings. You know when to pick which mode of communication. • You are willing to work across teams. Your work has to be aligned with the needs of other squads and external stakeholders. You make your plans transparent, bring stakeholders on board, and are open to feedback and suggestions. • You are pragmatic; you prioritize progress over perfection; you can handle ambiguity. • We are using Golang on the backend; you are either already familiar with Golang or are excited to learn the language In the United States, the compensation range for this role is $174,986 - $209,983 USD. Actual compensation may vary based on level, experience, and skillset as assessed throughout the interview process. All of our roles include Restricted Stock Units (RSUs), giving every team member ownership in Grafana Labs' success. We believe in shared outcomes—RSUs help us stay aligned and invested as we scale globally. • Compensation ranges are country specific. If you are applying for this role from a different location than listed above, your recruiter will discuss your specific market’s defined pay range & benefits at the beginning of the process. Why You’ll Thrive at Grafana Labs: • 100% Remote, Global Culture - As a remote-only company, we bring together talent from around the world, united by a culture of collaboration and shared purpose. • Scaling Organization – Tackle meaningful work in a high-growth, ever-evolving environment. • Transparent Communication – Expect open decision-making and regular company-wide updates. • Innovation-Driven – Autonomy and support to ship great work and try new things. • Open Source Roots – Built on community-driven values that shape how we work. • Empowered Teams – High trust, low ego culture that values outcomes over optics. • Career Growth Pathways – Defined opportunities to grow and develop your career. • Approachable Leadership – Transparent execs who are involved, visible, and human. • Passionate People – Join a team of smart, supportive folks who care deeply about what they do. • In-Person onboarding - We want you to thrive from day 1 with your fellow new ‘Grafanistas’ to learn all about what we do and how we do it. • Balance is Key - We operate a global annual leave policy of 30 days per annum. 3 days of your annual leave entitlement are reserved for Grafana Shutdown Days to allow the team to really disconnect. *We will comply with local legislation where applicable. Equal Opportunity Employer: We will recruit, train, compensate and promote regardless of race, religion, color, national origin, gender, disability, age, veteran status, and all the other fascinating characteristics that make us different and unique. We believe that equality and diversity builds a strong organization and we’re working hard to make sure that’s the foundation of our organization as we grow. Grafana Labs may utilize AI tools in its recruitment process to assist in matching information provided in CVs to job postings. The recruitment team will continue to review inbound CVs manually to identify alignment with current openings. For information about how your personal data is used once you’ve applied to a job, check out our privacy policy.

API Design
Backend Development
Observability
Mentorship
Technical Writing
Verified Source
Posted 2 months ago
Grafana Labs

Senior Manager, Web Technology | US | Remote

Grafana LabsAnywhereFull-time
View Job
Compensation$184K - 221K a year

Lead the web development team to manage and evolve the marketing website with AI-enabled transformations and operational excellence. | Strong technical leadership in web engineering, AI-enabled web experiences, team mentorship, and operational discipline. | Grafana Labs is a remote-first, open-source powerhouse. There are more than 20M users of Grafana, the open source visualization tool, around the globe, monitoring everything from beehives to climate change in the Alps. The instantly recognizable dashboards have been spotted everywhere from a NASA launch and Minecraft HQ to Wimbledon and the Tour de France. Grafana Labs also helps more than 3,000 companies - including Bloomberg, JPMorgan Chase, and eBay - manage their observability strategies with the Grafana LGTM Stack, which can be run fully managed with Grafana Cloud or self-managed with the Grafana Enterprise Stack, both featuring scalable metrics (Grafana Mimir), logs (Grafana Loki), and traces (Grafana Tempo). We’re scaling fast and staying true to what makes us different: an open-source legacy, a global collaborative culture, and a passion for meaningful work. Our team thrives in an innovation-driven environment where transparency, autonomy, and trust fuel everything we do. You may not meet every requirement, and that’s okay. If this role excites you, we’d love you to raise your hand for what could be a truly career-defining opportunity. This is a remote position and we're considering candidates in the US. The Opportunity: Grafana Labs is building for an agentic world, and the marketing website is where that ambition becomes visible. As Senior Manager, Web Technology, you will own the technical execution, platform strategy, and operational excellence of grafana.com and its surrounding digital experiences. You will lead the web development team and serve as the marketing organization’s primary technical leader for the website. You will help define how AI changes the website itself, how the team builds and operates it, how content is structured for humans and machines, and how digital experiences move users from interest to action with less friction and more confidence. This role sits at the intersection of engineering craft, marketing velocity, AI-enabled transformation, and business impact. You will partner closely with creative, content, communications, demand generation, regional marketing and events, marketing operations, docs, infrastructure, security, and product teams to make the website more dependable, more intelligent, more self-serve, and more effective. You are technical enough to earn the trust of a staff engineer, while communicating clearly with executive and marketing stakeholders. What You’ll Be Doing: • Lead AI-enabled website transformation. Drive the evolution of grafana.com into an AI-first, agent-ready platform optimized for humans, search engines, AI assistants, agents, and LLMs. Build practical AI-enabled systems for content workflows, QA, personalization, localization, analytics, experimentation, and web operations. • Own the web platform. Ensure the marketing website is stable, secure, performant, dependable, and easy to evolve. Define practical standards for reliability, incident handling, operational visibility, platform health, performance, and security hygiene. • Improve delivery predictability. Own the operating system for web technology work including clear operating cadences for roadmap reviews, stakeholder updates, prioritization, and delivery accountability.Use automation and AI-assisted workflows to reduce coordination overhead and improve clarity. • Build better digital experiences. Partner with marketing, product, design, and engineering stakeholders to help create a more seamless path from website visit to product trial and activation across different user intents and journeys. • Scale self-serve web operations. Make the site easier for internal teams to edit and maintain through strong CMS models, reusable patterns, content and UX guardrails, documentation, and governance. • Lead and grow the web team. Lead with clarity, accountability, strong technical judgment, and open communication while helping the team balance operational rigor with thoughtful innovation. Manage, mentor, and develop a team of web developers by raising technical standards, building AI-first development practices, clarifying ownership, and helping each person increase their impact and strategic contribution. What Makes You a Great Fit: • You have a strong point of view on how AI will change web strategy, web operations, and digital user experiences, and you can turn that point of view into practical systems, workflows, experiments, and measurable improvements. • You have strong technical judgment and can earn the trust of experienced engineers while communicating clearly with marketing and executive stakeholders. • You have experience leading web engineering teams and building reliable, scalable, high-performing web platforms with strong operational discipline and delivery predictability. • You understand modern frontend engineering, CMS-driven websites, analytics, integrations, performance, security, launch operations, and AI-enabled digital experiences. • You can turn ambiguous business needs into clear technical plans, priorities, timelines, tradeoffs, and decisions. • You have a strong operational mindset and improve delivery through better systems, not last-minute heroics. • You care about team growth and know how to develop engineers through mentorship, accountability, and meaningful ownership. • You focus on measurable impact across reliability, content velocity, self-serve adoption, discoverability, funnel performance, AI-enabled efficiency, and user engagement. • Experience with Next.js, React, TypeScript, Tailwind, Storyblok, Vercel, Algolia, RudderStack, BigQuery, AirOps, and similar tools. • Experience using AI to improve engineering workflows, QA, documentation, planning, analytics, experimentation, or operational reporting. • Experience creating self-serve systems that help marketing, content, events, and regional teams move faster without compromising quality. • Experience building AI-enabled web experiences, agent-ready content systems, personalization, chat experiences, or LLM-optimized websites. • What success looks like: • Predictable, high-confidence delivery with clear priorities and fewer surprises • A stable, secure, high-performing web platform with strong operational visibility • Faster, more self-serve content operations across marketing teams • Improved discoverability, engagement, conversion, and activation through better digital experiences • Practical AI-enabled systems that improve team velocity, quality, and user experience Bonus points for: • Familiarity with AEO, GEO, SEO, structured content, and how AI assistants consume website information. • Experience supporting developer, infrastructure, observability, cloud, or open source audiences. Compensation & Rewards: In the US, the compensation range for this role is $183,858 - $220,630. Actual compensation may vary based on level, experience, and skillset as assessed throughout the interview process. All of our roles include Restricted Stock Units (RSUs), giving every team member ownership in Grafana Labs' success. We believe in shared outcomes—RSUs help us stay aligned and invested as we scale globally. • Compensation ranges are country specific. If you are applying for this role from a different location than listed above, your recruiter will discuss your specific market’s defined pay range & benefits at the beginning of the process. Why You’ll Thrive at Grafana Labs: • 100% Remote, Global Culture - As a remote-only company, we bring together talent from around the world, united by a culture of collaboration and shared purpose. • Scaling Organization – Tackle meaningful work in a high-growth, ever-evolving environment. • Transparent Communication – Expect open decision-making and regular company-wide updates. • Innovation-Driven – Autonomy and support to ship great work and try new things. • Open Source Roots – Built on community-driven values that shape how we work. • Empowered Teams – High trust, low ego culture that values outcomes over optics. • Career Growth Pathways – Defined opportunities to grow and develop your career. • Approachable Leadership – Transparent execs who are involved, visible, and human. • Passionate People – Join a team of smart, supportive folks who care deeply about what they do. • In-Person onboarding - We want you to thrive from day 1 with your fellow new ‘Grafanistas’ to learn all about what we do and how we do it. • Balance is Key - We operate a global annual leave policy of 30 days per annum. 3 days of your annual leave entitlement are reserved for Grafana Shutdown Days to allow the team to really disconnect. *We will comply with local legislation where applicable. Equal Opportunity Employer: We will recruit, train, compensate and promote regardless of race, religion, color, national origin, gender, disability, age, veteran status, and all the other fascinating characteristics that make us different and unique. We believe that equality and diversity builds a strong organization and we’re working hard to make sure that’s the foundation of our organization as we grow. Grafana Labs may utilize AI tools in its recruitment process to assist in matching information provided in CVs to job postings. The recruitment team will continue to review inbound CVs manually to identify alignment with current openings. #LI-Remote For information about how your personal data is used once you’ve applied to a job, check out our privacy policy.

Web Development
API Design
Mentorship
Verified Source
Posted 3 months ago
GL

Staff Backend Engineer - Grafana App Platform| US| Remote

Grafana LabsAnywhereFull-time
View Job
Compensation$85K - 150K a year

Refactor and scale Grafana's core architecture for multi-tenancy and observability platform support including coding, design documentation, and on-call rotations. | Strong Golang coding skills or willingness to learn, extensive SaaS platform operation experience, proficiency with distributed systems, and excellent communication skills. | Grafana Labs is a remote-first, open-source powerhouse. There are more than 20M users of Grafana, the open source visualization tool, around the globe, monitoring everything from beehives to climate change in the Alps. The instantly recognizable dashboards have been spotted everywhere from a NASA launch and Minecraft HQ to Wimbledon and the Tour de France. Grafana Labs also helps more than 3,000 companies -- including Bloomberg, JPMorgan Chase, and eBay -- manage their observability strategies with the Grafana LGTM Stack, which can be run fully managed with Grafana Cloud or self-managed with the Grafana Enterprise Stack, both featuring scalable metrics (Grafana Mimir), logs (Grafana Loki), and traces (Grafana Tempo). We’re scaling fast and staying true to what makes us different: an open-source legacy, a global collaborative culture, and a passion for meaningful work. Our team thrives in an innovation-driven environment where transparency, autonomy, and trust fuel everything we do. You may not meet every requirement, and that’s okay. If this role excites you, we’d love you to raise your hand for what could be a truly career-defining opportunity. This is a remote position. We are looking for candidates in the US and Canada. The Opportunity: As our Cloud / Grafana-as-a-service business continues to grow, we've started to change Grafana's core architecture to be fully multi-tenant and scalable, as well as a solid platform for our opinionated Cloud apps. We're looking for someone with a lot of experience with SaaS platforms at scale. We are turning Grafana into a proper observability app platform where OSS and proprietary apps can directly tap into dashboards, alerts, incidents, and telemetry and deliver even more integrated experiences. To get there, we need to refactor a big part of Grafana so that it’s simpler and standardized. Grafana is used by countless OSS and Cloud users across different platforms, so planning and rolling out changes safely to avoid service disruptions is crucial; we are looking for someone who is excited about this sort of work. What You’ll Be Doing: Coding new features, enhancing the operational experience of the systems you develop, and iteratively improving them based on production insights Authoring, contributing to and reviewing design documents Taking an active role in shaping the roadmap Mentoring and supporting other team members and collaborating with different teams across the organization Owning the experience of our customers by participating in weekday 12h x 5d and a separate weekend 24h x 2d on-call rotations. We invest heavily in developer productivity. You can use modern AI coding assistants as part of your daily workflow (your choice of tools, within security guidelines), backed by a company-funded usage budget so you can iterate quickly without unnecessary friction. We encourage pragmatic AI-assisted development: faster prototyping, test generation, refactors, documentation, and incident follow-ups—always paired with strong code review and quality standards. You’ll also have access to frontier models (e.g., GPT-Codex 5/3, Claude Opus 4.7, Gemini 3 Pro). What Makes You a Great Fit: You have strong coding skills and operational experience; you were responsible for operating the software you have built. You have worked on a SaaS platform and dealt with common distributed systems problems (e.g. scalability, multi-tenancy, data isolation, HA, …) You have excellent communication skills. You can express your ideas with great clarity in writing and during meetings. You know when to pick which mode of communication. You are willing to work across teams. Your work has to be aligned with the needs of other squads and external stakeholders. You make your plans transparent, bring stakeholders on board, and are open to feedback and suggestions. You are pragmatic; you prioritize progress over perfection; you can handle ambiguity. We are using Golang on the backend; you are either already familiar with Golang or are excited to learn the language In the United States, the compensation range for this role is $174,986 - $209,983 USD. Actual compensation may vary based on level, experience, and skillset as assessed throughout the interview process. All of our roles include Restricted Stock Units (RSUs), giving every team member ownership in Grafana Labs' success. We believe in shared outcomes—RSUs help us stay aligned and invested as we scale globally. *Compensation ranges are country specific. If you are applying for this role from a different location than listed above, your recruiter will discuss your specific market’s defined pay range & benefits at the beginning of the process. Why You’ll Thrive at Grafana Labs: 100% Remote, Global Culture - As a remote-only company, we bring together talent from around the world, united by a culture of collaboration and shared purpose. Scaling Organization – Tackle meaningful work in a high-growth, ever-evolving environment. Transparent Communication – Expect open decision-making and regular company-wide updates. Innovation-Driven – Autonomy and support to ship great work and try new things. Open Source Roots – Built on community-driven values that shape how we work. Empowered Teams – High trust, low ego culture that values outcomes over optics. Career Growth Pathways – Defined opportunities to grow and develop your career. Approachable Leadership – Transparent execs who are involved, visible, and human. Passionate People – Join a team of smart, supportive folks who care deeply about what they do. In-Person onboarding - We want you to thrive from day 1 with your fellow new ‘Grafanistas’ to learn all about what we do and how we do it. Balance is Key - We operate a global annual leave policy of 30 days per annum. 3 days of your annual leave entitlement are reserved for Grafana Shutdown Days to allow the team to really disconnect. *We will comply with local legislation where applicable. Equal Opportunity Employer: We will recruit, train, compensate and promote regardless of race, religion, color, national origin, gender, disability, age, veteran status, and all the other fascinating characteristics that make us different and unique. We believe that equality and diversity builds a strong organization and we’re working hard to make sure that’s the foundation of our organization as we grow. Grafana Labs may utilize AI tools in its recruitment process to assist in matching information provided in CVs to job postings. The recruitment team will continue to review inbound CVs manually to identify alignment with current openings. #LI-Remote For information about how your personal data is used once you’ve applied to a job, check out our privacy policy.

Backend Engineering
System Architecture
API Design
Distributed Systems
Technical Mentoring
Direct Apply
Posted 3 months ago
Grafana Labs

Staff Backend Engineer - Application Core Services, Stacks | USA | Remote

Grafana LabsAnywhereFull-time
View Job
Compensation$175K - 210K a year

Design, build, and operate backend systems for scalable cloud services with a focus on reliability and operational clarity. | Experience with distributed systems, Golang, Kubernetes, cloud infrastructure, and remote work, plus leadership and mentoring capabilities. | This is a remote opportunity and we would be interested in applicants located in USA time zones (EST + CST only at this time). Staff Backend Engineer - Application Core Services, Stacks The Opportunity: Application Core Services (AppCore) partners closely with our Cloud, Enterprise, and Grafana teams to deliver reliable internal and customer-facing systems that power critical parts of the Grafana business. We build on the grafana.com platform to create custom solutions and integrations across the many systems that support a modern software company. The team owns important domain areas that help keep both our customer workflows and internal business processes running smoothly. AppCore is made up of multiple squads, each focused on one or more of these domains. Our work includes maintaining the billing engine responsible for customer usage calculation, automating provisioning after a customer signs a contract, integrating with cloud marketplaces such as AWS, Azure, and GCP, and building and maintaining the user portal our customers rely on to manage their accounts. This is a team working at the intersection of product, platform, and business operations. The systems we build are critical to how Grafana scales. We are looking for engineers who enjoy solving complex workflow and systems problems, improving reliability and developer experience, and building software that directly supports both customers and internal stakeholders. As a company we are remote-first and global, we embrace people of different experiences and backgrounds to build diverse teams where every person brings a unique perspective to the software. Engineers at Grafana also have the opportunity to contribute to Open Source communities and collaborate across teams beyond their immediate scope. What You'll Be Doing: The AppCore Stacks squad owns the systems that create, configure, reconcile, migrate, and operate Grafana Cloud stacks at scale. A stack is the customer-facing Grafana Cloud environment that connects an organization to Grafana and the backend services it uses, including Mimir, Loki, Tempo, plugins, dashboards, data sources, and stack-level configuration. Our work sits at the intersection of product, platform, and operations. We build the control-plane services and workflows that keep stack state aligned across grafana.com, Stack State Service (SSS), Hosted Grafana, cloud regions, and the underlying Grafana Cloud infrastructure. When this domain works well, customers get reliable stack creation, safe configuration rollout, predictable migrations, and fewer manual operational interventions. • Design, build, and operate reconciliation systems, including the SSS backend, to track desired stack state, detect and repair drift across stack templates, grafana.com state, Hosted Grafana, and actual customer stack configuration • Collaborate across SSS, grafana.com, and deployment configurations to ensure stack lifecycle workflows remain reliable, observable, and resilient • Improve operational efficiency by reducing deployment complexity (e.g., aiming for single PR regional SSS deployment) and contributing to the Stack Config Reconciliation project • Manage rollout mechanisms for provisioned plugins, dashboards, data sources, Grafana versions, release channels, and stack-level configuration • Support new region and cluster rollouts, including the operational paths required to bring stacks online safely in new Grafana Cloud regions • Improve incident response and recovery paths for stack misalignment, reconciliation failures, plugin rollout issues, and Hosted Grafana integration failures • Partner with Product, Hosted Grafana, Infrastructure, Support, and adjacent AppCore squads on customer-impacting stack lifecycle work • Contribute to roadmap planning, technical design, OnCall improvements, and long-term simplification of stack operations • You will help own the production behavior of the systems you build. That includes improving runbooks, dashboards, alerts, reconciliation safety, rollout controls, and recovery procedures. You should be comfortable debugging across service boundaries and making careful changes in systems that affect customer stacks Of course, there is an on-call component to this role and one that we take seriously. As a company, we hire globally (remote-first) to ensure our on-call remains healthy and aligned to approximately 12 daylight hours per day. You will work closely with counterparts in other regions to provide balanced coverage and shared ownership. We invest heavily in developer productivity. You can use modern AI coding assistants as part of your daily workflow (your choice of tools, within security guidelines), backed by a company-funded usage budget so you can iterate quickly without unnecessary friction. We encourage pragmatic AI-assisted development: faster prototyping, test generation, refactors, documentation, and incident follow-ups-always paired with strong code review and quality standards. You'll also have access to frontier models (e.g., GPT-Codex 5/3, Claude Opus 4.6, Gemini 3 Pro). What Makes You a Great Fit: At Grafana, we actively embrace AI-assisted and agentic development practices, integrating these technologies into both our engineering workflows and the systems we deliver. We encourage our engineers to thoughtfully leverage AI tools to enhance every stage of the lifecycle, from design and implementation to testing, documentation, and operations. We also look for strategic opportunities to embed agentic capabilities within our services to eliminate toil, bolster reliability, and ensure that complex customer workflows remain resilient and safe. We are seeking a Staff Backend Engineer who thrives on building production systems where correctness, scalability, and operational clarity are paramount. As a remote-first organization, you should be comfortable collaborating asynchronously across time zones and taking full ownership of the critical systems powering Grafana Cloud. Our team is small and operates with a high degree of independence; you will be expected to lead major projects, coordinate across service boundaries, and help define the technical direction for our domain. You will be particularly successful in this role if you enjoy solving challenges related to stateful systems, eventual consistency, and reconciliation loops. We value engineers who can take ambiguous lifecycle requirements and transform them into explicit, modular solutions. You should be adept at breaking down complex systems work into safe, iterative increments while clearly communicating technical tradeoffs to both internal stakeholders and adjacent product teams. Some things you might be expected to do could include: • Writing efficient, readable, and easy to maintain code • Designing new microservices or systems • Collaborating with teammates and other departments to reach consensus on proposed solutions • Coordinating with product and UX when needed • Responding to customer requests and feedback • When ready, participating in our follow-the-sun OnCall rotation • Participating in team decisions, such as roadmap planning and prioritization Requirements: • You have at least 1 year of fully remote work experience • You have worked on a big SaaS platform and dealt with common distributed systems problems (e.g. scalability, multi-tenancy, data isolation, HA, ...) • Have professional experience with Golang and be willing to work across both backend service and application code • Care deeply about developer and user experience and the quality of the products that you work on • Have some experience with delivering projects from gathering requirements, and brainstorming ideas to shipping a product to the customer's hands in a self-driven way • You write clean, robust, well-tested software that other engineers can understand, operate, and maintain • Have experience with mentoring junior engineers in a collaborative but asynchronous environment • Can take on complex challenges and break them down to achieve tight learning loops: to analyze, design, and build modular solutions, deliver MVPs, gather data and feedback, and then progress iteratively • You are willing to work across teams. Your work has to be aligned with the needs of other squads and external stakeholders. You make your plans transparent, bring stakeholders on board, and are open to feedback and suggestions • Strong Kubernetes experience in AWS, GCP, or Azure, and familiarity with infrastructure-as-code tooling (Helm, Terraform, Jsonnet, etc.) • Experience participating in blameless incident response and writing high-quality post-incident reviews Bonus Points For: • Experience with TypeScript/Node.js • Experience with Kubernetes control-plane patterns, operators, reconcilers, or desired-state systems • Experience with Jsonnet/Tanka, Terraform, Flux, Argo, or similar deployment/configuration tooling • Experience working on SaaS provisioning, tenancy, regional expansion, plugin rollout, or customer lifecycle systems • Experience with incident response involving configuration drift, partial failure, or cross-service state mismatch Compensation & Rewards: In the United States, the Base compensation range for this role is USD 174,986 - USD 209,983. Actual compensation may vary based on level, experience, and skillset as assessed in the interview process. Benefits include equity, bonus (if applicable) and other benefits listed here. All of our roles include Restricted Stock Units (RSUs), giving every team member ownership in Grafana Labs' success. We believe in shared outcomes-RSUs help us stay aligned and invested as we scale globally. *Compensation ranges are country specific. If you are applying for this role from a different location than listed above, your recruiter will discuss your specific market's defined pay range & benefits at the beginning of the process.

API Design
Backend Development
System Architecture
Verified Source
Posted 3 months ago
Grafana Labs

Staff Backend Engineer - Application Core Services, Stacks | Canada | Remote

Grafana LabsAnywhereFull-time
View Job
Compensation$186K - 224K a year

Design, build, and operate backend systems ensuring reliability and scalability for Grafana Cloud stacks. | Experienced in backend engineering with distributed systems, Golang, Kubernetes, cloud infrastructure, and remote collaboration. | This is a remote opportunity and we would be interested in applicants located in Canadian time zones (EST + CST only at this time). Staff Backend Engineer - Application Core Services, Stacks The Opportunity: Application Core Services (AppCore) partners closely with our Cloud, Enterprise, and Grafana teams to deliver reliable internal and customer-facing systems that power critical parts of the Grafana business. We build on the grafana.com platform to create custom solutions and integrations across the many systems that support a modern software company. The team owns important domain areas that help keep both our customer workflows and internal business processes running smoothly. AppCore is made up of multiple squads, each focused on one or more of these domains. Our work includes maintaining the billing engine responsible for customer usage calculation, automating provisioning after a customer signs a contract, integrating with cloud marketplaces such as AWS, Azure, and GCP, and building and maintaining the user portal our customers rely on to manage their accounts. This is a team working at the intersection of product, platform, and business operations. The systems we build are critical to how Grafana scales. We are looking for engineers who enjoy solving complex workflow and systems problems, improving reliability and developer experience, and building software that directly supports both customers and internal stakeholders. As a company we are remote-first and global, we embrace people of different experiences and backgrounds to build diverse teams where every person brings a unique perspective to the software. Engineers at Grafana also have the opportunity to contribute to Open Source communities and collaborate across teams beyond their immediate scope. What You'll Be Doing: The AppCore Stacks squad owns the systems that create, configure, reconcile, migrate, and operate Grafana Cloud stacks at scale. A stack is the customer-facing Grafana Cloud environment that connects an organization to Grafana and the backend services it uses, including Mimir, Loki, Tempo, plugins, dashboards, data sources, and stack-level configuration. Our work sits at the intersection of product, platform, and operations. We build the control-plane services and workflows that keep stack state aligned across grafana.com, Stack State Service (SSS), Hosted Grafana, cloud regions, and the underlying Grafana Cloud infrastructure. When this domain works well, customers get reliable stack creation, safe configuration rollout, predictable migrations, and fewer manual operational interventions. • Design, build, and operate reconciliation systems, including the SSS backend, to track desired stack state, detect and repair drift across stack templates, grafana.com state, Hosted Grafana, and actual customer stack configuration • Collaborate across SSS, grafana.com, and deployment configurations to ensure stack lifecycle workflows remain reliable, observable, and resilient • Improve operational efficiency by reducing deployment complexity (e.g., aiming for single PR regional SSS deployment) and contributing to the Stack Config Reconciliation project • Manage rollout mechanisms for provisioned plugins, dashboards, data sources, Grafana versions, release channels, and stack-level configuration • Support new region and cluster rollouts, including the operational paths required to bring stacks online safely in new Grafana Cloud regions • Improve incident response and recovery paths for stack misalignment, reconciliation failures, plugin rollout issues, and Hosted Grafana integration failures • Partner with Product, Hosted Grafana, Infrastructure, Support, and adjacent AppCore squads on customer-impacting stack lifecycle work • Contribute to roadmap planning, technical design, OnCall improvements, and long-term simplification of stack operations • You will help own the production behavior of the systems you build. That includes improving runbooks, dashboards, alerts, reconciliation safety, rollout controls, and recovery procedures. You should be comfortable debugging across service boundaries and making careful changes in systems that affect customer stacks Of course, there is an on-call component to this role and one that we take seriously. As a company, we hire globally (remote-first) to ensure our on-call remains healthy and aligned to approximately 12 daylight hours per day. You will work closely with counterparts in other regions to provide balanced coverage and shared ownership. We invest heavily in developer productivity. You can use modern AI coding assistants as part of your daily workflow (your choice of tools, within security guidelines), backed by a company-funded usage budget so you can iterate quickly without unnecessary friction. We encourage pragmatic AI-assisted development: faster prototyping, test generation, refactors, documentation, and incident follow-ups-always paired with strong code review and quality standards. You'll also have access to frontier models (e.g., GPT-Codex 5/3, Claude Opus 4.6, Gemini 3 Pro). What Makes You a Great Fit: At Grafana, we actively embrace AI-assisted and agentic development practices, integrating these technologies into both our engineering workflows and the systems we deliver. We encourage our engineers to thoughtfully leverage AI tools to enhance every stage of the lifecycle, from design and implementation to testing, documentation, and operations. We also look for strategic opportunities to embed agentic capabilities within our services to eliminate toil, bolster reliability, and ensure that complex customer workflows remain resilient and safe. We are seeking a Staff Backend Engineer who thrives on building production systems where correctness, scalability, and operational clarity are paramount. As a remote-first organization, you should be comfortable collaborating asynchronously across time zones and taking full ownership of the critical systems powering Grafana Cloud. Our team is small and operates with a high degree of independence; you will be expected to lead major projects, coordinate across service boundaries, and help define the technical direction for our domain. You will be particularly successful in this role if you enjoy solving challenges related to stateful systems, eventual consistency, and reconciliation loops. We value engineers who can take ambiguous lifecycle requirements and transform them into explicit, modular solutions. You should be adept at breaking down complex systems work into safe, iterative increments while clearly communicating technical tradeoffs to both internal stakeholders and adjacent product teams. Some things you might be expected to do could include: • Writing efficient, readable, and easy to maintain code • Designing new microservices or systems • Collaborating with teammates and other departments to reach consensus on proposed solutions • Coordinating with product and UX when needed • Responding to customer requests and feedback • When ready, participating in our follow-the-sun OnCall rotation • Participating in team decisions, such as roadmap planning and prioritization Requirements: • You have at least 1 year of fully remote work experience • You have worked on a big SaaS platform and dealt with common distributed systems problems (e.g. scalability, multi-tenancy, data isolation, HA, ...) • Have professional experience with Golang and be willing to work across both backend service and application code • Care deeply about developer and user experience and the quality of the products that you work on • Have some experience with delivering projects from gathering requirements, and brainstorming ideas to shipping a product to the customer's hands in a self-driven way • You write clean, robust, well-tested software that other engineers can understand, operate, and maintain • Have experience with mentoring junior engineers in a collaborative but asynchronous environment • Can take on complex challenges and break them down to achieve tight learning loops: to analyze, design, and build modular solutions, deliver MVPs, gather data and feedback, and then progress iteratively • You are willing to work across teams. Your work has to be aligned with the needs of other squads and external stakeholders. You make your plans transparent, bring stakeholders on board, and are open to feedback and suggestions • Strong Kubernetes experience in AWS, GCP, or Azure, and familiarity with infrastructure-as-code tooling (Helm, Terraform, Jsonnet, etc.) • Experience participating in blameless incident response and writing high-quality post-incident reviews Bonus Points For: • Experience with TypeScript/Node.js • Experience with Kubernetes control-plane patterns, operators, reconcilers, or desired-state systems • Experience with Jsonnet/Tanka, Terraform, Flux, Argo, or similar deployment/configuration tooling • Experience working on SaaS provisioning, tenancy, regional expansion, plugin rollout, or customer lifecycle systems • Experience with incident response involving configuration drift, partial failure, or cross-service state mismatch Compensation & Rewards: In Canada, the Base compensation range for this role is CAD 186,368 - CAD 223,642. Actual compensation may vary based on level, experience, and skillset as assessed in the interview process. Benefits include equity, bonus (if applicable) and other benefits listed here. All of our roles include Restricted Stock Units (RSUs), giving every team member ownership in Grafana Labs' success. We believe in shared outcomes-RSUs help us stay aligned and invested as we scale globally. *Compensation ranges are country specific. If you are applying for this role from a different location than listed above, your recruiter will discuss your specific market's defined pay range & benefits at the beginning of the process.

Backend development
API design
System architecture
Verified Source
Posted 3 months ago
Grafana Labs

Backend Engineer | Mimir

Grafana LabsAnywhereFull-time
View Job
Compensation$128K - 153K a year

Design, build, operate, and maintain scalable backend systems and contribute to open-source projects in a remote team environment. | Experience with backend programming, cloud or systems engineering, collaboration in remote teams, on-call duties, and familiarity with observability systems. | Grafana Labs is a remote-first, open-source powerhouse. There are more than 20M users of Grafana, the open source visualization tool, around the globe, monitoring everything from beehives to climate change in the Alps. The instantly recognizable dashboards have been spotted everywhere from a NASA launch and Minecraft HQ to Wimbledon and the Tour de France. Grafana Labs also helps more than 3,000 companies -- including Bloomberg, JPMorgan Chase, and eBay -- manage their observability strategies with the Grafana LGTM Stack, which can be run fully managed with Grafana Cloud or self-managed with the Grafana Enterprise Stack, both featuring scalable metrics (Grafana Mimir), logs (Grafana Loki), and traces (Grafana Tempo). We’re scaling fast and staying true to what makes us different: an open-source legacy, a global collaborative culture, and a passion for meaningful work. Our team thrives in an innovation-driven environment where transparency, autonomy, and trust fuel everything we do. You may not meet every requirement, and that’s okay. If this role excites you, we’d love you to raise your hand for what could be a truly career-defining opportunity. This is a remote position. We are looking for candidates in the United States time zones. What is Grafana Cloud? Grafana Cloud is our composable observability platform that integrates metrics, logs, traces, and profiles with Grafana. It allows our customers to leverage the best open source observability software – including Prometheus, Mimir, Loki, Tempo, and Pyroscope – without the overhead of installing, maintaining and scaling their own observability stack. The Databases team owns and operates the telemetry databases that are Mimir for metrics, Loki for logs, Tempo for traces, and Pyroscope for profiles. Our databases are offered as a hosted service in Grafana Cloud, and additionally as on-premise solutions with Grafana Enterprise Metrics, Grafana Enterprise Logs, and Grafana Enterprise Traces. They are multi-tenant distributed systems implemented in Go and operating at scale on Kubernetes across all major Cloud service providers (AWS, GCP, Azure). As a company we are remote-first and global, we embrace people of different experiences and backgrounds to build diverse teams where every person brings a new perspective to the software. Mimir Squad The Mimir squad has 3 sub-squads, Ingest, Storage, and Query, which together maintain the Mimir OSS project, and additionally own and operate Grafana Cloud Metrics across 3 major cloud providers. Engineers on the team focus on optimizing the efficiency and resilience of processing, storing, and querying metrics at large volumes. These services operate at a large scale and performance is key to keeping the offering competitive and running smoothly. A Mimir engineer has various work streams. They are likely engaged in a larger project with another engineer, and they are also incorporating some performance and reliability improvements discovered through operating the system in production. They are also responsible for writing and reviewing PRs and design documents from other engineers in the squad, shepherding automated release rollouts, and participating in the on-call rotation for their systems. What will you be doing? • Take an active role in influencing our roadmap and your own career objectives • Work with your team to deliver new features, then use the results to iterate and improve. • Help your team drive projects from initial idea all the way to operations once it is in the hands of customers • Embrace our open-source culture and contribute to other projects that may not directly fall within your team’s scope • Build, operate, and maintain critical systems, owning the reliability, performance, and availability • Be a part of your team’s follow-the-sun on-call rotations and take ownership of the services you’re running • Support other team members, participate in design discussions and collaborate with the team • Learn new skills by gaining a deeper understanding of our cloud product and our customers and getting to know the codebase of a large distributed system As we are remote-first and our engineering organization is largely remote, we provide guidance and meet regularly using video calls, so an independent attitude and good communication skills are a must. We invest heavily in developer productivity. You can use modern AI coding assistants as part of your daily workflow (your choice of tools, within security guidelines), backed by a company-funded usage budget so you can iterate quickly without unnecessary friction. We encourage pragmatic AI-assisted development: faster prototyping, test generation, refactors, documentation, and incident follow-ups—always paired with strong code review and quality standards. You’ll also have access to frontier models (e.g., GPT-Codex 5/3, Claude Opus 4.6, Gemini 3 Pro). What Makes You a Great Fit: You are a motivated self starter with a bias towards action. You are customer focused. We build everything with our users in mind. You have a passion for building intuitive products that fit customers’ needs. You have good time management skills, which you leverage to work on the right things at the right time. • Pragmatism: You have a bias towards action, taking direction and building a plan of action to analyze, design, and build modular solutions, deliver MVPs, gather data and feedback and then progress iteratively • Collaboration and communication: The smallest unit we have is a squad. You’ll be working with your teammates in a fully remote setup. Good communication and time management skills are a must • AI: some experience using LLMs for day to day coding tasks and understanding a codebase. • Experience with at least one programming language. We use Go, but if you have familiarity with Python, C, C++, Rust or similar then that translates well • Some experience with delivering projects as a member of a larger team. Your experience includes gathering requirements and brainstorming ideas, all the way to shipping features to the customer’s hands. • Some experience with developing software that runs in the Cloud or some experience with systems engineering • Some experience with being on-call and following the DevOps model • Experience writing clean, robust, and performant software that is easily maintained by others • Some familiarity with observability systems, know when to use metrics, logs, traces, to debug a problem. Bonus Points For: • Experience working with Kubernetes • Experience working with queue systems, e.g. the Kafka protocol • Been a user of Grafana and Prometheus in operational roles (including on-call for your team at a previous employer or just using these tools on hobby/homelab projects) • Exposure to microservices architecture and distributed systems, or a desire to learn • Familiarity with the concept of infrastructure as code Compensation & Rewards: In the United States, the Base compensation range for this role is USD 127,651 - USD 153,180. Actual compensation may vary based on level, experience, and skillset as assessed in the interview process. Benefits include equity, bonus (if applicable) and other benefits listed here. • Compensation ranges are country-specific. If you are applying for this role from a different location than listed above, your recruiter will discuss your specific market’s defined pay range & benefits at the beginning of the process. Why You’ll Thrive at Grafana Labs: • 100% Remote, Global Culture - As a remote-only company, we bring together talent from around the world, united by a culture of collaboration and shared purpose. • Scaling Organization – Tackle meaningful work in a high-growth, ever-evolving environment. • Transparent Communication – Expect open decision-making and regular company-wide updates. • Innovation-Driven – Autonomy and support to ship great work and try new things. • Open Source Roots – Built on community-driven values that shape how we work. • Empowered Teams – High trust, low ego culture that values outcomes over optics. • Career Growth Pathways – Defined opportunities to grow and develop your career. • Approachable Leadership – Transparent execs who are involved, visible, and human. • Passionate People – Join a team of smart, supportive folks who care deeply about what they do. • In-Person onboarding - We want you to thrive from day 1 with your fellow new ‘Grafanistas’ to learn all about what we do and how we do it. • Balance is Key - We operate a global annual leave policy of 30 days per annum. 3 days of your annual leave entitlement are reserved for Grafana Shutdown Days to allow the team to really disconnect. *We will comply with local legislation where applicable. Remote Skills: Amazon Web Services (AWS), Analysis Skills, Apache Kafka, Apiary/Beekeeping, Artificial Intelligence (AI), Budgeting, C Programming Language, C++ Programming Language, Climate Change, Cloud Computing, Code Reviews, Communication Skills, Customer Relations, Data Collection, Debugging Skills, Design Document, Distributed Computing, Documentation, GCP (Good Clinical Practices), Leadership, Metrics, Microservices, Microsoft Windows Azure, On Call, Onboarding, Open Source, Operational Support Systems (OSS), People Management, Performance Management, Product Development, Production Systems, Programming Languages, Project Engineering, Pronto Software, Prototyping, Python Programming/Scripting Language, Quality Metrics, Refactoring, Reliability Engineering, Reporting Dashboards, Requirements Management, Software Development, Software Engineering, System Operations, Systems Engineering, Systems Maintenance, Team Lead/Manager, Team Player, Telemetry, Testing, Time Management About the Company: Grafana Labs

API Design
Backend Development
System Architecture
Observability
Microservices
Verified Source
Posted 3 months ago
Grafana Labs

Staff Backend Engineer – Application Core Services

Grafana LabsAnywhereFull-time
View Job
Compensation$90K - 150K a year

Design, build, and operate reconciliation systems and improve operational efficiency in a distributed SaaS environment. | Requires remote experience, Golang proficiency, Kubernetes and cloud infrastructure skills, mentoring experience, and incident response familiarity. | Job Description: • Design, build, and operate reconciliation systems • Collaborate across teams for stack lifecycle workflows • Improve operational efficiency • Manage rollout mechanisms for provisioned plugins • Support new region and cluster rollouts • Improve incident response for stack issues • Participate in OnCall rotation Requirements: • At least 1 year of fully remote work experience • Experience in large SaaS platforms and distributed systems problems • Proficient in Golang • Clean, robust, well-tested software development • Experience mentoring junior engineers • Strong Kubernetes experience in AWS, GCP, or Azure • Familiarity with infrastructure-as-code tooling (Helm, Terraform, Jsonnet, etc.) • Experience in blameless incident response Benefits: • Restricted Stock Units (RSUs) • Health insurance • 30 days annual leave • Flexible work hours • In-Person onboarding • Company-funded AI coding assistants usage budget • Professional development opportunities

API design
Backend development
Mentorship
Verified Source
Posted 3 months ago
GL

Staff Backend Engineer - Databases Tempo | US | Remote

Grafana LabsAnywhereFull-time
View Job
Compensation$120K - 160K a year

Lead technical initiatives for distributed tracing backend, focusing on architecture, performance, and mentoring. | Deep experience with distributed data systems, proficiency in Go, technical leadership, and operational mindset. | Grafana Labs is a remote-first, open-source powerhouse. There are more than 20M users of Grafana, the open source visualization tool, around the globe, monitoring everything from beehives to climate change in the Alps. The instantly recognizable dashboards have been spotted everywhere from a NASA launch and Minecraft HQ to Wimbledon and the Tour de France. Grafana Labs also helps more than 3,000 companies -- including Bloomberg, JPMorgan Chase, and eBay -- manage their observability strategies with the Grafana LGTM Stack, which can be run fully managed with Grafana Cloud or self-managed with the Grafana Enterprise Stack, both featuring scalable metrics (Grafana Mimir), logs (Grafana Loki), and traces (Grafana Tempo). We’re scaling fast and staying true to what makes us different: an open-source legacy, a global collaborative culture, and a passion for meaningful work. Our team thrives in an innovation-driven environment where transparency, autonomy, and trust fuel everything we do. You may not meet every requirement, and that’s okay. If this role excites you, we’d love you to raise your hand for what could be a truly career-defining opportunity. This is a remote position. We are seeking candidates in US and Canada. The Opportunity: We build Tempo, the open-source distributed tracing backend behind Grafana Cloud Traces and Grafana Enterprise Traces (GET). Tempo makes it easy to search traces, generate metrics from spans, and connect tracing data with logs, metrics, and profiles across the Grafana stack. 2026 is an inflection point for Tempo. After a major architectural upgrade and the launch of TraceQL metrics, we are shifting from foundational work to product and operational excellence, and evolving Tempo from a SaaS database into a platform that powers Grafana’s next generation of observability products (App Observability, Asserts, Traces Drilldown, and AI-driven assistants). Over the next year, you will help us: Make Grafana Cloud Traces “just work” for customers by eliminating rough edges, confusing limits, and hidden failure modes. Achieve operational excellence at scale as we grow from close to 50 cells today into triple digits this year, with autoscaling, parameterized rollouts, and aggressive toil reduction. Evolve Tempo into a platform enabler: higher-density APIs, trace aggregation, TraceQL metrics math, and machine/LLM-friendly interfaces that downstream products and agents can build on. Push performance further: faster query latency at hundreds of MB/s ingestion and performant 30-day query ranges to match competitors. Prepare Tempo for an agent-driven world: larger, burstier, higher-cardinality workloads, and new categories of AI-powered workflows, such as assistant-driven triage and “why is this slow?”- style investigations. What You’ll Be Doing: As a Staff Engineer on Tempo, you will set technical direction on the hardest problems in our roadmap and raise the bar across the team. Lead multi-quarter technical initiatives from problem framing through rollout, e.g., trace aggregation APIs, Limitless Tempo, autoscaling cells and customer limits, or query engine improvements. Own the architecture of core Tempo components: ingestion, storage, query, and metrics generation. Drive design reviews, make sharp trade-offs on performance, cost, and complexity, and document the “why” for the team. Design APIs for humans and agents. Shape the next generation of Tempo’s interfaces (structured, deterministic, discoverable) so that Act 3 products, LLM-driven assistants, and external integrators can build on Tempo reliably. Drive operational excellence. Own outcomes against concrete SLOs (P99 write latency, incident recurrence, TCO per ingested GB) and push the team toward Zero Ops through automation, parameterized rollouts, and actionable alerts. Partner with Product and sibling teams. Work closely with PMs and with App Observability, Asserts, Drilldown, and Grafana Assistant teams to understand how Tempo gets consumed and to ship what unblocks them. Mentor engineers. Raise the engineering bar through code review, design feedback, pairing on hard problems, and writing that leaves the team smarter than you found it. Participate in on-call for the services you help build, and be a force multiplier in incident response and post-incident learning. Contribute to open source. Tempo is OSS. You will engage the community, review external contributions, and help steer the project in the open. We invest heavily in developer productivity. You can use modern AI coding assistants as part of your daily workflow (your choice of tools, within security guidelines), backed by a company-funded usage budget so you can iterate quickly without unnecessary friction. We encourage pragmatic AI-assisted development: faster prototyping, test generation, refactors, documentation, and incident follow-ups—always paired with strong code review and quality standards. You’ll also have access to frontier models (e.g., GPT-Codex 5/3, Claude Opus 4.6, Gemini 3 Pro). Example problems you could work on These are the kinds of projects landing in 2026. Any one of them is a Staff-sized problem: Trace aggregation and higher-density APIs: extend TraceQL metrics, design LLM-friendly response types, and make Tempo a first-class data source for Grafana’s AI assistant. Autoscaling end to end: customer limits and Tempo cells, with hysteresis, predictive scaling for spikes, and safe scale-down. Agent-scale ingestion and query: guardrails for bursty, high-cardinality, agent-generated workloads. Query performance: new data formats, smarter query pipelines, targeted optimizations for common Drilldown and Traces workflows, and 30-day query ranges. Rollouts and multi-cell operations: parameterized rollouts, push-button deploys, and the tooling to grow safely into triple-digit cell counts without a proportional increase in alert noise. Limits and self-service: drive customer-facing configuration and observability so escalations trend toward zero. What Makes You a Great Fit: Technical leadership. A track record of leading complex, multi-quarter initiatives that spanned design, delivery, and operations, and made the teams around you better. Deep systems experience. Substantial hands-on experience building and operating distributed data systems in production: ingestion pipelines, storage engines, query execution, or similar. Strong software craftsmanship. You write clean, robust, performant software that others can maintain, and you know when to optimize vs. when to ship. Strong Go, or a path to it. We write Tempo in Go. Deep experience in other systems languages (Rust, C, C++) translates well. Operational mindset. You’ve owned production services, carried a pager, reduced toil, and treated SLOs as a product feature, not a chore. Customer focus and pragmatism. You break complex problems into short feedback loops: analyze, design, deliver an MVP, learn, iterate. Leadership through writing and collaboration. You lead through design docs, reviews, and shipped code, not hierarchy. You communicate clearly in a fully remote, asynchronous environment. Bonus Points For: Experience with tracing, OpenTelemetry, or large-scale observability systems. Experience designing query languages, SQL/TraceQL-like engines, or APIs intended to be consumed programmatically (by services or agents). Experience with columnar storage formats (e.g., Parquet) or purpose-built on-disk formats for analytical workloads. Experience operating multi-tenant, multi-cell SaaS infrastructure at scale on Kubernetes. Experience building for AI/LLM consumers: structured APIs, metadata/discovery endpoints, deterministic outputs, evaluation harnesses. Open-source contribution or maintainership, and comfort engaging a community in the open. Experience as an on-call user of Grafana, Prometheus, Loki, or Tempo in a previous role (or on a homelab). Experience in a fully remote, globally distributed team. How we work We are a remote-first team that meets regularly over video and does most of our work asynchronously, in writing. We value creativity, diverse perspectives, and clear communication. Tempo is relied upon by prominent global organizations to monitor critical applications and infrastructure, and we expect everyone on the team, including our Staff engineers, to contribute ideas that make it a more reliable, more useful, and more loved product. In the United States, the compensation range for this role is $174,986 - $209,983 USD. Actual compensation may vary based on level, experience, and skillset as assessed throughout the interview process. All of our roles include Restricted Stock Units (RSUs), giving every team member ownership in Grafana Labs' success. We believe in shared outcomes—RSUs help us stay aligned and invested as we scale globally. *Compensation ranges are country specific. If you are applying for this role from a different location than listed above, your recruiter will discuss your specific market’s defined pay range & benefits at the beginning of the process. Why You’ll Thrive at Grafana Labs: 100% Remote, Global Culture - As a remote-only company, we bring together talent from around the world, united by a culture of collaboration and shared purpose. Scaling Organization – Tackle meaningful work in a high-growth, ever-evolving environment. Transparent Communication – Expect open decision-making and regular company-wide updates. Innovation-Driven – Autonomy and support to ship great work and try new things. Open Source Roots – Built on community-driven values that shape how we work. Empowered Teams – High trust, low ego culture that values outcomes over optics. Career Growth Pathways – Defined opportunities to grow and develop your career. Approachable Leadership – Transparent execs who are involved, visible, and human. Passionate People – Join a team of smart, supportive folks who care deeply about what they do. In-Person onboarding - We want you to thrive from day 1 with your fellow new ‘Grafanistas’ to learn all about what we do and how we do it. Balance is Key - We operate a global annual leave policy of 30 days per annum. 3 days of your annual leave entitlement are reserved for Grafana Shutdown Days to allow the team to really disconnect. *We will comply with local legislation where applicable. Equal Opportunity Employer: We will recruit, train, compensate and promote regardless of race, religion, color, national origin, gender, disability, age, veteran status, and all the other fascinating characteristics that make us different and unique. We believe that equality and diversity builds a strong organization and we’re working hard to make sure that’s the foundation of our organization as we grow. Grafana Labs may utilize AI tools in its recruitment process to assist in matching information provided in CVs to job postings. The recruitment team will continue to review inbound CVs manually to identify alignment with current openings. #LI-Remote For information about how your personal data is used once you’ve applied to a job, check out our privacy policy.

API Design
Observability
Mentoring
Direct Apply
Posted 3 months ago
GL

Senior Fullstack Engineer - Observability Real User Monitoring (RUM) | US | Remote

Grafana LabsAnywhereFull-time
View Job
Compensation$90K - 150K a year

Build and evolve fullstack features including backend services, APIs, and frontend visualization for Real User Monitoring, and design systems for high-volume telemetry data ingestion and querying. | 5+ years fullstack engineering experience with strong backend fundamentals, proficiency in Go and TypeScript/React, and experience with distributed systems and observability concepts. | Grafana Labs is a remote-first, open-source powerhouse. There are more than 20M users of Grafana, the open source visualization tool, around the globe, monitoring everything from beehives to climate change in the Alps. The instantly recognizable dashboards have been spotted everywhere from a NASA launch and Minecraft HQ to Wimbledon and the Tour de France. Grafana Labs also helps more than 3,000 companies -- including Bloomberg, JPMorgan Chase, and eBay -- manage their observability strategies with the Grafana LGTM Stack, which can be run fully managed with Grafana Cloud or self-managed with the Grafana Enterprise Stack, both featuring scalable metrics (Grafana Mimir), logs (Grafana Loki), and traces (Grafana Tempo). We’re scaling fast and staying true to what makes us different: an open-source legacy, a global collaborative culture, and a passion for meaningful work. Our team thrives in an innovation-driven environment where transparency, autonomy, and trust fuel everything we do. You may not meet every requirement, and that’s okay. If this role excites you, we’d love you to raise your hand for what could be a truly career-defining opportunity. This is a full-time remote opportunity. We are considering candidates from US and Canada only. The opportunity Grafana Observability builds end-to-end observability that spans application, infrastructure, database, browser, and mobile. Our Real User Monitoring (RUM) initiative focuses on capturing, storing, and querying high-volume user interaction data from browsers and mobile devices, enabling teams to understand real-world user experiences at scale. We’re building systems that ingest and process massive amounts of telemetry—sessions, events, traces, and logs—and make them explorable in real time. This requires deep expertise in high-performance backend systems, columnar storage, and intuitive frontend experiences. Our solutions are tightly integrated with OpenTelemetry and Grafana Cloud. We care deeply about performance, cost-efficiency, and developer experience across the entire stack—from instrumentation to query layer to visualization. We value open standards, great developer experience, and doing the hard engineering needed to ship reliable software at scale. You may not meet every requirement below. If this role excites you, please raise your hand. What you’ll be doing Build and evolve fullstack features for RUM, spanning backend services, APIs, storage systems, and frontend user experiences. Design and implement systems that ingest, store, and query high-cardinality, high-volume telemetry data using columnar/analytical databases. Develop performant query layers and APIs that power real-time exploration of user sessions, traces, and events. Contribute to frontend applications that visualize RUM data, enabling users to debug performance issues and understand user behavior. Work on data modeling, indexing strategies, and query optimization to ensure low-latency, cost-efficient analytics at scale. Collaborate closely with SDK engineers (browser and mobile) to ensure high-quality data ingestion and schema evolution. Own projects end-to-end: from design and implementation to deployment, monitoring, and iteration. Break down complex, ambiguous problems into incremental deliverables and iterate quickly based on feedback. Ensure quality through testing, observability of your own systems, documentation, and smooth upgrade paths. Collaborate cross-functionally with backend, frontend, product, and solutions engineering to deliver cohesive observability workflows. Support teammates, participate in technical design discussions and help shape the RUM roadmap. We invest heavily in developer productivity. You can use modern AI coding assistants as part of your daily workflow (your choice of tools, within security guidelines), backed by a company-funded usage budget so you can iterate quickly without unnecessary friction. We encourage pragmatic AI-assisted development: faster prototyping, test generation, refactors, documentation, and incident follow-ups—always paired with strong code review and quality standards. You’ll also have access to frontier models (e.g., GPT-Codex 5/3, Claude Opus 4.6, Gemini 3 Pro). What makes you a great fit 5+ years of fullstack engineering experience with strong backend fundamentals Backend experience (Go is preferred) and frontend experience, we use TypeScript and React Experience building or operating distributed systems in production (e.g., Kafka, WarpStream, ClickHouse, Cassandra, Postgres) Familiarity with cloud-native systems (Docker, Kubernetes, AWS, GCP, Azure) Experience working with high-throughput, high-cardinality data (logs, metrics, traces, events) Strong understanding of data modeling, query optimization, and performance tradeoffs Experience designing and building APIs and distributed services Experience building data-heavy UIs (dashboards, query tools, debugging interfaces) Familiarity with observability concepts (traces, logs, metrics) and/or OpenTelemetry Strong communication skills and ability to work in a remote, distributed team Pragmatic, self-driven, and c omfortable navigating ambiguity Customer-focused mindset with a passion for developer experience Bonus / nice-to-have Experience with browser or mobile instrumentation (RUM SDKs, telemetry collection). Mobile development experience (iOS or Android) or familiarity with mobile performance and telemetry. Contributions to OpenTelemetry or other observability OSS. Experience building developer-facing platforms or observability products. Familiarity with session replay, sampling strategies, or user behavior analytics systems. In the United States, the compensation range for this role is $154,445 - $185,334 USD. Actual compensation may vary based on level, experience, and skillset as assessed throughout the interview process. All of our roles include Restricted Stock Units (RSUs), giving every team member ownership in Grafana Labs' success. We believe in shared outcomes—RSUs help us stay aligned and invested as we scale globally. *Compensation ranges are country specific. If you are applying for this role from a different location than listed above, your recruiter will discuss your specific market’s defined pay range & benefits at the beginning of the process. Why You’ll Thrive at Grafana Labs: 100% Remote, Global Culture - As a remote-only company, we bring together talent from around the world, united by a culture of collaboration and shared purpose. Scaling Organization – Tackle meaningful work in a high-growth, ever-evolving environment. Transparent Communication – Expect open decision-making and regular company-wide updates. Innovation-Driven – Autonomy and support to ship great work and try new things. Open Source Roots – Built on community-driven values that shape how we work. Empowered Teams – High trust, low ego culture that values outcomes over optics. Career Growth Pathways – Defined opportunities to grow and develop your career. Approachable Leadership – Transparent execs who are involved, visible, and human. Passionate People – Join a team of smart, supportive folks who care deeply about what they do. In-Person onboarding - We want you to thrive from day 1 with your fellow new ‘Grafanistas’ to learn all about what we do and how we do it. Balance is Key - We operate a global annual leave policy of 30 days per annum. 3 days of your annual leave entitlement are reserved for Grafana Shutdown Days to allow the team to really disconnect. *We will comply with local legislation where applicable. Equal Opportunity Employer: We will recruit, train, compensate and promote regardless of race, religion, color, national origin, gender, disability, age, veteran status, and all the other fascinating characteristics that make us different and unique. We believe that equality and diversity builds a strong organization and we’re working hard to make sure that’s the foundation of our organization as we grow. Grafana Labs may utilize AI tools in its recruitment process to assist in matching information provided in CVs to job postings. The recruitment team will continue to review inbound CVs manually to identify alignment with current openings. #LI-Remote For information about how your personal data is used once you’ve applied to a job, check out our privacy policy.

TypeScript
API Design
System Architecture
Observability
Frontend Development
Direct Apply
Posted 3 months ago
GL

Staff Backend Engineer - Session Replay | USA | Remote

Grafana LabsAnywhereFull-time
View Job
Compensation$120K - 180K a year

Own end-to-end technical direction for backend architecture and data system design, partnering with cross-functional teams and mentoring staff. | Strong experience designing data-intensive systems, proficiency in Go, TypeScript, React, and ability to communicate complex technical decisions in a remote-first environment. | Grafana Labs is a remote-first, open-source powerhouse. There are more than 20M users of Grafana, the open source visualization tool, around the globe, monitoring everything from beehives to climate change in the Alps. The instantly recognizable dashboards have been spotted everywhere from a NASA launch and Minecraft HQ to Wimbledon and the Tour de France. Grafana Labs also helps more than 3,000 companies -- including Bloomberg, JPMorgan Chase, and eBay -- manage their observability strategies with the Grafana LGTM Stack, which can be run fully managed with Grafana Cloud or self-managed with the Grafana Enterprise Stack, both featuring scalable metrics (Grafana Mimir), logs (Grafana Loki), and traces (Grafana Tempo). We’re scaling fast and staying true to what makes us different: an open-source legacy, a global collaborative culture, and a passion for meaningful work. Our team thrives in an innovation-driven environment where transparency, autonomy, and trust fuel everything we do. You may not meet every requirement, and that’s okay. If this role excites you, we’d love you to raise your hand for what could be a truly career-defining opportunity. This is a remote position. We are looking for candidates in the EST or CT timezone in the United States or Canada. The Opportunity: The Session Replay squad is building a new Grafana Cloud product that helps our customers to understand what users actually experienced when something goes wrong. Session Replay connects frontend signals (errors, performance, synthetic checks) to concrete session-level evidence, enabling faster and more confident investigation of production issues. This team works at the intersection of: frontend observability, backend data processing and storage at scale, debugging workflows across products, privacy and access control, performance and cost constraints at scale. Session Replay is still early, which means you’ll help shape both what we build and how it fits into Grafana Cloud. A key part of our next phase is evolving the backend architecture of capturing sessions, including a migration toward columnar/analytical solution as a primary storage and query engine for high-volume session data. As a company we are remote-first and global, we embrace people of different experiences and backgrounds to build diverse teams where every person brings a new perspective to the software. We are looking for Staff Software Engineers who are passionate about working with data and providing seamless experiences for our customers to join our growing team! Our stack is Golang, Typescript and React, but we build tools for people using many other stacks. What you'll be doing: As a Staff Software Engineer, you will operate as a technical leader and systems thinker, driving both product direction and architectural evolution, specifically, you will: Own end-to-end technical direction for Session Replay, spanning frontend, backend, and data systems Drive the evolution of our backend architecture, including: Designing systems around columnar/analytical data storage for large-scale session data Defining data models, ingestion pipelines, and query patterns Lead the design of investigation workflows, connecting replay with logs, metrics, traces and other telemetry across Grafana Cloud Make high-leverage architectural decisions that impact multiple teams and products Partner with teams across Grafana (Frontend Observability, Synthetic Monitoring, Core Grafana) to build cohesive cross-product experiences Improve engineering standards, patterns, and operational practices within the team Mentor engineers and help grow technical leadership within the team Technologies you'll work with: Go (backend services and APIs) Columnar/Analytical data storage (core data storage and querying) Object storage (S3, GCS, Azure Blob Storage) and MySQL TypeScript / React (user-facing workflows) Grafana ecosystem (Mimir, Loki, Tempo, etc.) What Makes You a Great Fit: You are comfortable working in a remote-first company; communication is key. For us, working together means being collaborative, friendly, kind, and respectful. We operate by consensus. You can contribute to a discussion, disagree constructively, and commit to the team’s decision. You are able to communicate design decisions clearly in written and spoken English. Ability to reason about data-intensive systems (ingestion, storage, querying, cost trade-offs) You are comfortable owning features in ambiguous problem spaces. We are a small team, working remotely, on a product that will be used by engineers all over the world – the ability to work on your own is crucial. You have a good understanding of a software development process that takes a user-centered approach. You easily build an understanding of the users’ context and goals which will help you build the right solution with the maximum value. You enjoy working on complex solutions – Grafana is a highly technological solution and has avid followers who rely on it every day and who care deeply about their workflows. You value code maintainability, readability & automation. Bonus Points For: Experience with columnar/analytical databases Experience with observability tools (Grafana, Datadog, New Relic, Sentry, etc.). Experience building debugging or developer-focused tools. Familiarity with privacy, security, and access control in data-heavy systems Experience working on performance-sensitive systems (large datasets, real-time queries, session data) Compensation & Rewards: In the United States, the Base compensation range for this role is USD 174,986 - USD 218,733. Actual compensation may vary based on level, experience, and skillset as assessed in the interview process. Benefits include equity, bonus (if applicable) and other benefits listed here. *Compensation ranges are country-specific. If you are applying for this role from a different location than listed above, your recruiter will discuss your specific market’s defined pay range & benefits at the beginning of the process. *Compensation ranges are country specific. If you are applying for this role from a different location than listed above, your recruiter will discuss your specific market’s defined pay range & benefits at the beginning of the process. Why You’ll Thrive at Grafana Labs: 100% Remote, Global Culture - As a remote-only company, we bring together talent from around the world, united by a culture of collaboration and shared purpose. Scaling Organization – Tackle meaningful work in a high-growth, ever-evolving environment. Transparent Communication – Expect open decision-making and regular company-wide updates. Innovation-Driven – Autonomy and support to ship great work and try new things. Open Source Roots – Built on community-driven values that shape how we work. Empowered Teams – High trust, low ego culture that values outcomes over optics. Career Growth Pathways – Defined opportunities to grow and develop your career. Approachable Leadership – Transparent execs who are involved, visible, and human. Passionate People – Join a team of smart, supportive folks who care deeply about what they do. In-Person onboarding - We want you to thrive from day 1 with your fellow new ‘Grafanistas’ to learn all about what we do and how we do it. Balance is Key - We operate a global annual leave policy of 30 days per annum. 3 days of your annual leave entitlement are reserved for Grafana Shutdown Days to allow the team to really disconnect. *We will comply with local legislation where applicable. Equal Opportunity Employer: We will recruit, train, compensate and promote regardless of race, religion, color, national origin, gender, disability, age, veteran status, and all the other fascinating characteristics that make us different and unique. We believe that equality and diversity builds a strong organization and we’re working hard to make sure that’s the foundation of our organization as we grow. Grafana Labs may utilize AI tools in its recruitment process to assist in matching information provided in CVs to job postings. The recruitment team will continue to review inbound CVs manually to identify alignment with current openings. #LI-Remote For information about how your personal data is used once you’ve applied to a job, check out our privacy policy.

System architecture
TypeScript
API design
Observability
Mentorship
Direct Apply
Posted 3 months ago

Ready to join Grafana Labs?

Create tailored applications specifically for Grafana Labs with our AI-powered resume builder

Get Started for Free

Ready to have AI work for you in your job search?

Sign-up for free and start using JobLogr today!

Get Started »
JobLogr badgeTinyLaunch BadgeJobLogr - AI Job Search Tools to Land Your Next Job Faster than Ever | Product Hunt