Find Your Next Opportunity
Browse through hundreds of job listings
Staff AI Engineer - Grafana Labs
51 viewsGrafana Labs, the company behind the open observability cloud, is founded on the principles of open source, open standards, open ecosystems, and open culture. Grafana Cloud, our fully managed observability platform, is flexible and built for scale. With Grafana Cloud's actually useful AI, organizations can see, understand, and act on all their disparate data to move at the speed of their ambitions. Today, more than 35 million users and 7,000+ customers – including Anthropic, Bloomberg, NVIDIA, Microsoft, and Salesforce – trust Grafana Labs to ensure reliability of their applications and systems, resolve incidents quickly, and optimize their telemetry to reduce noise and cost. We are a 100% remote company with 1,600+ team members across 40+ countries, and we’re backed by leading investors including Lightspeed Venture Partners, Sequoia Capital, GIC, Coatue, J.P. Morgan, CapitalG, and Lead Edge Capital. Learn more at grafana.com and follow us on LinkedIn and X. We’re scaling fast and staying true to what makes us different: an open-source legacy, a global collaborative culture, and a passion for meaningful work. Our team thrives in an innovation-driven environment where transparency, autonomy, and trust fuel everything we do. You may not meet every requirement, and that’s okay. If this role excites you, we’d love you to raise your hand for what could be a truly career-defining opportunity.This is a remote opportunity and we would be interested in applicants from USA time zones only at this time. Staff AI Engineer The Opportunity: Grafana's Revenue Operations organization is looking for a Staff AI Engineer to own the AI agent infrastructure and automation platform that powers our GTM teams. You'll build multi-agent architectures, LLM integrations, and backend services that connect AI models to internal and third-party data platforms. You'll ship production systems that teams depend on daily. This is a high-autonomy role where you own the technical direction. You'll identify the highest-leverage problems across Sales, Customer Success, and Marketing, design the solutions, and ship them. You'll define the technical direction for the automation platform—data models, API contracts, shared libraries, reference architectures—and partner with Data Engineering, GTM Systems, Field Operations, and GTM leadership to build scalable, self-service automation that eliminates manual work and drives operational efficiency. What You’ll Be Doing: Agentic Tool Development Own end-to-end development of multi-agent AI systems, from architecture and implementation through testing, deployment, and ongoing operation Build modular, composable agentic systems using orchestration frameworks (LangChain, CrewAI, Anthropic MCP, or similar) that operate 24/7 across teams Develop reusable agentic skills that agents invoke across interfaces (Slack, dashboards, internal apps, CLIs) Implement observability and feedback loops including logging, performance metrics, prompt iteration, model evaluation, and cost management Establish governance and compliance standards for AI workflows including access controls, audit trails, PII handling, and human-in-the-loop escalation paths Systems Integration & Backend Services Build MCP servers, APIs, CLIs, and microservices connecting AI models to business systems (BigQuery, Slack, CRMs, email, calendars, analytics tools) Architect data flows for retrieval-augmented generation (RAG), connecting LLMs to internal knowledge bases, customer data, and real-time business context Build serverless or containerized services (GCP Cloud Functions, Cloud Run) that scale with usage and integrate with Grafana's cloud infrastructure Automation & Workflow Manufacturing Partner with RevOps, Demand Generation, Regional Marketing, and SDR teams to scope high-impact automation problems, identify bottlenecks, and build solutions with measurable business outcomes Design and deploy workflows using orchestration tools (n8n, Workato, or custom platforms) with CI/CD, testing, and production reliability standards Build systems designed for self-service with documentation, playbooks, and enablement materials that let partner teams operate independently We invest heavily in developer productivity. You'll have access to AI coding assistants (Claude Code, Gemini CLI, OpenAI Codex, and others of your choice within security guidelines). We encourage pragmatic AI-assisted development paired with strong code review and quality standards. What Makes You a Great Fit / Requirements: 8+ years of software engineering experience with depth in backend development, systems integration, or data/analytics engineering 2+ years hands-on experience applying LLMs/AI to production workflows, not just prototypes Strong proficiency in Python and JavaScript/Node.js with Git-based workflows, code review practices, and testing discipline Hands-on experience with LLM frameworks and patterns including prompt engineering, RAG, function calling/tool use, structured output parsing, and evaluation Experience building and operating multi-agent systems at scale including agent decomposition, orchestration patterns (sequential chains, router/dispatcher, parallel fan-out), state management, and production monitoring You diagnose business problems before writing code. You think in workflows and outcomes, not just functions. Deep familiarity with Google Cloud Platform, BigQuery, and serverless/containerized services (Cloud Functions, Cloud Run) Understanding of LLM failure modes and production mitigations including confidence thresholds, fallback logic, human escalation, and cost/latency management Proven ability to identify high-leverage problems, push back on low-impact requests, and deliver end-to-end with minimal direction Fluent with AI-assisted development tools (GitHub Copilot, Cursor, Claude Code). You use AI to build AI systems Clear technical communicator—you can explain complex systems in simple terms to both engineers and business stakeholders Bonus Points For: Experience with frontend frameworks & tooling (React, Slack Block Kit, dashboard components) to build user-facing interfaces for AI tools Familiarity with GTM platforms like Salesforce, HubSpot, Outreach, Gainsight, or similar CRM/sales engagement tools Experience with vector databases or retrieval pipelines (Pinecone, Weaviate, ChromaDB, pgvector, or similar) Prior work automating sales, customer success, or marketing workflows in a B2B SaaS environment Experience with workflow automation platforms like n8n, Prefect, Clay, PhantomBuster, Apify, Dust, or similar tools Familiarity with Model Context Protocol (MCP) or similar standards for connecting AI systems to data sources and tools Exposure to observability tools for AI systems (LangSmith, Weights & Biases, custom logging/evaluation frameworks) Experience working in Revenue Operations, GTM Analytics, or Sales Operations environments Previous experience in open source or developer-focused SaaS companies—Grafana is built on OSS and we value engineers who share that DNA Compensation & Rewards: In the United States, the Base compensation range for this role is USD 174,986 - USD 220,000. Actual compensation may vary based on level, experience, and skillset as assessed in the interview process. Benefits include equity, bonus (if applicable) and other benefits listed here. All of our roles include Restricted Stock Units (RSUs), giving every team member ownership in Grafana Labs' success. We believe in shared outcomes—RSUs help us stay aligned and invested as we scale globally. *Compensation ranges are country specific. If you are applying for this role from a different location than listed above, your recruiter will discuss your specific market’s defined pay range & benefits at the beginning of the process. Why You’ll Thrive at Grafana Labs: 100% Remote, Global Culture - As a remote-only company, we bring together talent from around the world, united by a culture of collaboration and shared purpose. Scaling Organization – Tackle meaningful work in a high-growth, ever-evolving environment. Transparent Communication – Expect open decision-making and regular company-wide updates. Innovation-Driven – Autonomy and support to ship great work and try new things. Open Source Roots – Built on community-driven values that shape how we work. Empowered Teams – High trust, low ego culture that values outcomes over optics. Career Growth Pathways – Defined opportunities to grow and develop your career. Approachable Leadership – Transparent execs who are involved, visible, and human. Passionate People – Join a team of smart, supportive folks who care deeply about what they do. In-Person onboarding - We want you to thrive from day 1 with your fellow new ‘Grafanistas’ to learn all about what we do and how we do it. Balance is Key - We operate a global annual leave policy of 30 days per annum. 3 days of your annual leave entitlement are reserved for Grafana Shutdown Days to allow the team to really disconnect. *We will comply with local legislation where applicable. Equal Opportunity Employer: Grafana Labs is an equal opportunities employer. We welcome applications from everyone regardless of race, colour, nationality, origin, caste, sex, gender reassignment identity or expression, sexual orientation, age, religion or belief, disability, veteran status, genetic information, pregnancy, maternity, marital, family or carer status, or any other characteristic which is protected by local law. We believe that equality and diversity build a strong organisation, and we work hard to ensure that is the foundation of our organisation as we grow. Grafana Labs may utilize AI tools in its recruitment process to assist in matching information provided in CVs to job postings. The recruitment team will continue to review inbound CVs manually to identify alignment with current openings. #LI-Remote For information about how your personal data is used once you’ve applied to a job, check out our privacy policy.
United States (Remote), No Exact Region, No Exact Location
Posted: 13/07/2026 | Deadline: 12/08/2026
Grafana Labs, the company behind the open observability cloud, is founded on the principles of open source, open standards, open ecosystems, and open culture. Grafana Cloud, our fully managed observability platform, is flexible and built for scale. With Grafana Cloud's actually useful AI, organizations can see, understand, and act on all their disparate data to move at the speed of their ambitions. Today, more than 35 million users and 7,000+ customers – including Anthropic, Bloomberg, NVIDIA, Microsoft, and Salesforce – trust Grafana Labs to ensure reliability of their applications and systems, resolve incidents quickly, and optimize their telemetry to reduce noise and cost. We are a 100% remote company with 1,600+ team members across 40+ countries, and we’re backed by leading investors including Lightspeed Venture Partners, Sequoia Capital, GIC, Coatue, J.P. Morgan, CapitalG, and Lead Edge Capital. Learn more at grafana.com and follow us on LinkedIn and X. We’re scaling fast and staying true to what makes us different: an open-source legacy, a global collaborative culture, and a passion for meaningful work. Our team thrives in an innovation-driven environment where transparency, autonomy, and trust fuel everything we do. You may not meet every requirement, and that’s okay. If this role excites you, we’d love you to raise your hand for what could be a truly career-defining opportunity.Staff Backend Engineer - Adaptive Telemetry, Databases This is a remote position. We are looking for candidates in the USA time zones. What is Grafana Cloud? Grafana Cloud is our composable observability platform that integrates metrics, logs, traces, and profiles with Grafana. It allows our customers to leverage the best open source observability software – including Prometheus, Mimir, Loki, Tempo, and Pyroscope – without the overhead of installing, maintaining and scaling their own observability stack. The Databases department owns and operates the telemetry databases that are Mimir for metrics, Loki for logs, Tempo for traces, and Pyroscope for profiles. We offer our databases as a Cloud service supporting Grafana Cloud. Adaptive Telemetry Group The Adaptive Telemetry group, part of the Databases department, has the mission of ensuring that all telemetry stored in our databases is worthy of attention. Under that mission, the group is responsible for the development of Adaptive Metrics, Adaptive Logs, Adaptive Traces and Adaptive Profiles. Our Adaptive Telemetry solutions give users the ability to control and optimize their telemetry data. These solutions ensure that data storage is optimized based on individual usage patterns, so only the most valuable data is retained. As a company we are remote-first and global, we embrace people of different experiences and backgrounds to build diverse teams where every person brings a new perspective to the software. What will you be doing: Drive technical strategy and roadmap. Proactively define the architectural vision, prioritize work that unlocks major product or platform improvements, and influence product and engineering decisions. Lead end-to-end delivery of large, cross-functional projects. Own planning, design, execution, rollout and long-term operation of large initiatives. Own architecture, reliability, performance and cost for critical systems. Make pragmatic architecture choices that balance scalability, availability, latency and cost while ensuring systems remain maintainable and evolvable. Define SLOs/SLIs and lead incident response. Establish measurable reliability targets, run high-severity incident response, lead blameless post-mortems, and drive systemic fixes and automation to prevent recurrence. Improve observability, automation and operational readiness. Champion telemetry, alerting, runbooks, capacity planning and automation efforts that reduce toil, speed debugging and lower MTTR. Align stakeholders and remove blockers. Coordinate across Product, Design and other teams to align priorities, negotiate tradeoffs, and unblock delivery for large initiatives. Mentor and grow engineering talent. Coach senior and mid-level engineers, lead design reviews, raise engineering standards, and help teammates make sound technical tradeoffs. Represent engineering internally and externally. Communicate technical strategy clearly to non-engineering stakeholders and represent the team in cross-team planning. We invest heavily in developer productivity. You can use modern AI coding assistants as part of your daily workflow (your choice of tools, within security guidelines), backed by a company-funded usage budget so you can iterate quickly without unnecessary friction. We encourage pragmatic AI-assisted development: faster prototyping, test generation, refactors, documentation, and incident follow-ups—always paired with strong code review and quality standards. You’ll also have access to frontier models (e.g., GPT-Codex 5/3, Claude Opus 4.6, Gemini 3 Pro). What makes you a great fit: You are a motivated self starter with a bias towards action. You are customer focused. We build everything with our users in mind. You have a passion for creating intuitive products that fit customers’ needs Proven delivery of large distributed systems. Experience shipping and operating complex systems that span multiple teams, with clear evidence of technical leadership and impact. Strong systems-design instincts. Deep understanding of tradeoffs around latency, consistency, availability, scaling and cost. Hands-on cloud and platform experience. Solid experience with cloud-native architectures (microservices, containers/Kubernetes, IaC) and the operational practices that keep them healthy. Reliability and performance ownership. Comfortable defining SLOs/SLIs, doing capacity planning, tuning performance, and driving reliability work end-to-end. Excellent coding and design skills. You write clear, maintainable, well-tested code and can lead technical designs — we use Go, but Python/C/C++/Rust or similar translate well. Comfort with AI-assisted development. We embrace AI and agentic development so we expect you to be curious and comfortable using AI-powered developer tools and ideally have practical experience folding them into a team’s workflow. Experience with messaging and telemetry. Familiarity with streaming/messaging systems (e.g., Kafka) and observability tooling (Prometheus/Grafana or equivalents). Influence without authority. Ability to align cross-functional stakeholders, set priorities and drive outcomes in a remote-first environment. Strong communicator. Clear written and verbal communication that works across engineers and non-technical stakeholders. Compensation & Rewards: In the United States, the Base compensation range for this role is USD 174,986 - USD 209,983. Actual compensation may vary based on level, experience, and skillset as assessed in the interview process. Benefits include equity, bonus (if applicable) and other benefits listed here. *Compensation ranges are country-specific. If you are applying for this role from a different location than listed above, your recruiter will discuss your specific market’s defined pay range & benefits at the beginning of the process. *Compensation ranges are country specific. If you are applying for this role from a different location than listed above, your recruiter will discuss your specific market’s defined pay range & benefits at the beginning of the process. Why You’ll Thrive at Grafana Labs: 100% Remote, Global Culture - As a remote-only company, we bring together talent from around the world, united by a culture of collaboration and shared purpose. Scaling Organization – Tackle meaningful work in a high-growth, ever-evolving environment. Transparent Communication – Expect open decision-making and regular company-wide updates. Innovation-Driven – Autonomy and support to ship great work and try new things. Open Source Roots – Built on community-driven values that shape how we work. Empowered Teams – High trust, low ego culture that values outcomes over optics. Career Growth Pathways – Defined opportunities to grow and develop your career. Approachable Leadership – Transparent execs who are involved, visible, and human. Passionate People – Join a team of smart, supportive folks who care deeply about what they do. In-Person onboarding - We want you to thrive from day 1 with your fellow new ‘Grafanistas’ to learn all about what we do and how we do it. Balance is Key - We operate a global annual leave policy of 30 days per annum. 3 days of your annual leave entitlement are reserved for Grafana Shutdown Days to allow the team to really disconnect. *We will comply with local legislation where applicable. Equal Opportunity Employer: Grafana Labs is an equal opportunities employer. We welcome applications from everyone regardless of race, colour, nationality, origin, caste, sex, gender reassignment identity or expression, sexual orientation, age, religion or belief, disability, veteran status, genetic information, pregnancy, maternity, marital, family or carer status, or any other characteristic which is protected by local law. We believe that equality and diversity build a strong organisation, and we work hard to ensure that is the foundation of our organisation as we grow. Grafana Labs may utilize AI tools in its recruitment process to assist in matching information provided in CVs to job postings. The recruitment team will continue to review inbound CVs manually to identify alignment with current openings. #LI-Remote For information about how your personal data is used once you’ve applied to a job, check out our privacy policy.
United States (Remote), No Exact Region, No Exact Location
Posted: 13/07/2026 | Deadline: 12/08/2026
About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems.About the Role We're looking for experienced engineers to build and scale the database infrastructure that powers Claude's products and Anthropic's research. As a Software Engineer on the Databases team, you'll architect and operate the systems that let millions of users interact with Claude while also supporting frontier AI research workloads. You'll help set the database strategy for Anthropic: designing systems that handle billions of API requests, building storage that runs cleanly across GCP, AWS, and a range of deployment models, and creating the reliable data layer that lets research move fast. The Databases team spans three areas: the core database platform (data plane and control plane), data movement (migrations, backfill, and change data capture), and caching. We're hiring across all three. Key Responsibilities Drive the technical direction for database solutions used across Product and Research Design and implement database solutions that scale to support millions of users across Claude's product ecosystem Build and scale database systems through 100x+ growth while maintaining reliability and performance Build the database platform that lets Anthropic engineers ship without thinking about databases or scaling. Architect data storage solutions that operate across GCP, AWS, first-party deployments, third-party deployments, and other environments Develop database infrastructure that serves both product and research workloads with different performance characteristics Build data movement infrastructure (migration tooling, backfill, and change data capture pipelines) that safely consolidates and moves data across the organization Design and operate caching infrastructure, including CDC-driven cache invalidation, that keeps Anthropic's hottest paths fast and correct Partner with product and research teams to understand data requirements and build infrastructure that accelerates their work Optimize database performance, reliability, and cost efficiency at scale Make build-vs-buy decisions for database technologies Minimum Qualifications Significant experience as a software engineer building and operating production database or storage systems Deep knowledge of distributed database architectures and OLTP systems at scale Proficiency with SQL and at least one major relational or distributed database engine (e.g., PostgreSQL, MySQL, Spanner, CockroachDB, DynamoDB) Track record of leading large, complex infrastructure projects as an engineer or tech lead Ability to balance moving quickly with the reliability needs of production systems Strong technical leadership and cross-functional collaboration skills Preferred Qualifications 10+ years building and scaling database systems, with 3+ years leading large-scale projects or teams Experience scaling databases through periods of rapid growth at high-growth companies Experience operating Spanner, CockroachDB, TiDB, AlloyDB, or other globally distributed SQL databases in production Experience with Redis, Temporal, vector databases, or async job processing frameworks Experience with change data capture (Debezium or similar), large-scale data migration, or streaming data infrastructure Experience building multi-cloud or hybrid cloud database solutions Knowledge of database orchestration and automation at scale Contributions to database internals, storage engines, or related open source projects Note: Prior AI/ML infrastructure experience is not required. We value deep infrastructure and database expertise from any domain. Deadline to apply: None. Applications will be reviewed on a rolling basis.The annual compensation range for this role is listed below. For sales roles, the range provided is the role’s On Target Earnings ("OTE") range, meaning that the range includes both the sales commissions/sales bonuses target and annual base salary for the role.Annual Salary:$320,000—$485,000 USDLogistics Minimum education: Bachelor’s degree or an equivalent combination of education, training, and/or experience Required field of study: A field relevant to the role as demonstrated through coursework, training, or professional experience Minimum years of experience: Years of experience required will correlate with the internal job level requirements for the position Location-based hybrid policy: Currently, we expect all staff to be in one of our offices at least 25% of the time. However, some roles may require more time in our offices. Visa sponsorship: We do sponsor visas! However, we aren't able to successfully sponsor visas for every role and every candidate. But if we make you an offer, we will make every reasonable effort to get you a visa, and we retain an immigration lawyer to help with this. We encourage you to apply even if you do not believe you meet every single qualification. Not all strong candidates will meet every single qualification as listed. Research shows that people who identify as being from underrepresented groups are more prone to experiencing imposter syndrome and doubting the strength of their candidacy, so we urge you not to exclude yourself prematurely and to submit an application if you're interested in this work. We think AI systems like the ones we're building have enormous social and ethical implications. We think this makes representation even more important, and we strive to include a range of diverse perspectives on our team.Your safety matters to us. To protect yourself from potential scams, remember that Anthropic recruiters only contact you from @anthropic.com email addresses. In some cases, we may partner with vetted recruiting agencies who will identify themselves as working on behalf of Anthropic. Be cautious of emails from other domains. Legitimate Anthropic recruiters will never ask for money, fees, or banking information before your first day. If you're ever unsure about a communication, don't click any links—visit anthropic.com/careers directly for confirmed position openings. How we're different We believe that the highest-impact AI research will be big science. At Anthropic we work as a single cohesive team on just a few large-scale research efforts. And we value impact — advancing our long-term goals of steerable, trustworthy AI — rather than work on smaller and more specific puzzles. We view AI research as an empirical science, which has as much in common with physics and biology as with traditional efforts in computer science. We're an extremely collaborative group, and we host frequent research discussions to ensure that we are pursuing the highest-impact work at any given time. As such, we greatly value communication skills. The easiest way to understand our research directions is to read our recent research. This research continues many of the directions our team worked on prior to Anthropic, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute, Concrete Problems in AI Safety, and Learning from Human Preferences. Come work with us! Anthropic is a public benefit corporation headquartered in San Francisco. We offer competitive compensation and benefits, optional equity donation matching, generous vacation and parental leave, flexible working hours, and a lovely office space in which to collaborate with colleagues. Guidance on Candidates' AI Usage: Learn about our policy for using AI in our application process.
San Francisco, NY | Seattle, WA
Posted: 26/07/2026 | Deadline: 25/08/2026
About Us At Cloudflare, we are on a mission to help build a better Internet. Today the company runs one of the world’s largest networks that powers millions of websites and other Internet properties for customers ranging from individual bloggers to SMBs to Fortune 500 companies. Cloudflare protects and accelerates any Internet application online without adding hardware, installing software, or changing a line of code. Internet properties powered by Cloudflare all have web traffic routed through its intelligent global network, which gets smarter with every request. As a result, they see significant improvement in performance and a decrease in spam and other attacks. Cloudflare was named to Entrepreneur Magazine’s Top Company Cultures list and ranked among the World’s Most Innovative Companies by Fast Company. At Cloudflare, we’re not looking for people who wait for a polished roadmap; we’re looking for the builders who see the cracks in the Internet that everyone else has simply learned to live with. We value candidates who have the instinct to spot a "normalized" problem and the AI-native curiosity to create a solution using the latest tools. Our culture is built on iteration, leveraging AI to ship faster today to make it better tomorrow, while ensuring that every improvement, no matter how small, is shared across the team to lift everyone up. If you’re the type of person who values curiosity over bureaucracy, and that AI is a partner in solving tough problems to keep the Internet moving forward, you’ll fit right in. Available Locations | Austin, TX or Seattle, WA About the Role Emerging Technologies & Incubation (ETI) is where new and bold products are built and released within Cloudflare. Rather than being constrained by the structures which make Cloudflare a massively successful business, we are able to leverage them to deliver entirely new tools and products to our customers. Cloudflare’s edge and network make it possible to solve problems at massive scale and efficiency which would be impossible for almost any other organization. Is this role a fit for you? We are looking for engineers who build infrastructure software from scratch, not those who operate it. If your core expertise is configuring, managing, scaling, or writing applications on top of existing data platforms like Kafka, Flink, Spark, Databricks, or Snowflake, this role is not the right fit. We are seeking the engineers who can write the code that replaces these tools. Responsibilities As a Senior Systems Engineer focused on Cloudflare's core storage and streaming products, you will design and implement the foundational layers of our next-generation data services. We work backward from developer needs, reimagining solutions by leveraging our unique global network. You will own your code from inception to release, delivering solutions at all layers of the software stack to empower Cloudflare customers. On any given day, you might be architecting a new globally distributed data consistency model, optimizing storage engine performance, or designing a novel API for a new data-centric product. You can expect to interact with a variety of languages and technologies including, but not limited to JavaScript, Typescript, Rust, and C++. Desirable Skills, Knowledge, and Experience Minimum 6+ years of experience designing, building, and deploying the internals of large-scale distributed systems or storage engines from scratch. Direct experience writing core components of a storage or streaming system. Proven track record of debugging, profiling, and fixing complex race conditions, split-brain scenarios, and network partition failures at the infrastructure layer. Deep, expert-level proficiency in Rust or C++ applied to systems programming. Bonus Points Experience building time-series databases, log-structured merge-trees, or real-time stream processing engines from the ground up. Academic or professional contributions to distributed systems research (e.g., consensus, conflict-free replicated data types (CRDTs), or database theory). Compensation Compensation may be adjusted depending on work location. For New York City, New Jersey, Washington, Washington DC, and California (excluding Bay Area) based hires: Estimated annual salary of $185,000 - $254,000 Equity This role is eligible to participate in Cloudflare’s equity plan. Benefits Cloudflare offers a complete package of benefits and programs to support you and your family. Our benefits programs can help you pay health care expenses, support caregiving, build capital for the future and make life a little easier and fun! The below is a description of our benefits for employees in the United States, and benefits may vary for employees based outside the U.S. Health & Welfare Benefits Medical/Rx Insurance Dental Insurance Vision Insurance Flexible Spending Accounts Commuter Spending Accounts Fertility & Family Forming Benefits On-demand mental health support and Employee Assistance Program Global Travel Medical Insurance Financial Benefits Short and Long Term Disability Insurance Life & Accident Insurance 401(k) Retirement Savings Plan Employee Stock Participation Plan Time Off Flexible paid time off covering vacation and sick leave Leave programs, including parental, pregnancy health, medical, and bereavement leave What Makes Cloudflare Special? We’re not just a highly ambitious, large-scale technology company. We’re a highly ambitious, large-scale technology company with a soul. Fundamental to our mission to help build a better Internet is protecting the free and open Internet. Project Galileo: Since 2014, we've equipped more than 2,400 journalism and civil society organizations in 111 countries with powerful tools to defend themselves against attacks that would otherwise censor their work, technology already used by Cloudflare’s enterprise customers--at no cost. Athenian Project: In 2017, we created the Athenian Project to ensure that state and local governments have the highest level of protection and reliability for free, so that their constituents have access to election information and voter registration. Since the project, we've provided services to more than 425 local government election websites in 33 states. 1.1.1.1: We released 1.1.1.1 to help fix the foundation of the Internet by building a faster, more secure and privacy-centric public DNS resolver. This is available publicly for everyone to use - it is the first consumer-focused service Cloudflare has ever released. Here’s the deal - we don’t store client IP addresses never, ever. We will continue to abide by our privacy commitment and ensure that no user data is sold to advertisers or used to target consumers. Sound like something you’d like to be a part of? We’d love to hear from you! Please note that applicants who progress to the offer stage of the interview process may be asked to attend an in-person interview within one of the Cloudflare Offices or Cloudflare Hubs. More details about this will be available at that stage of the interview process. This position may require access to information protected under U.S. export control laws, including the U.S. Export Administration Regulations. Please note that any offer of employment may be conditioned on your authorization to receive software or technology controlled under these U.S. export laws without sponsorship for an export license. Cloudflare is proud to be an equal opportunity employer. We are committed to providing equal employment opportunity for all people and place great value in both diversity and inclusiveness. All qualified applicants will be considered for employment without regard to their, or any other person's, perceived or actual race, color, religion, sex, gender, gender identity, gender expression, sexual orientation, national origin, ancestry, citizenship, age, physical or mental disability, medical condition, family care status, or any other basis protected by law. We are an AA/Veterans/Disabled Employer. Cloudflare provides reasonable accommodations to qualified individuals with disabilities. Please tell us if you require a reasonable accommodation to apply for a job. Examples of reasonable accommodations include, but are not limited to, changing the application process, providing documents in an alternate format, using a sign language interpreter, or using specialized equipment. If you require a reasonable accommodation to apply for a job, please contact us via e-mail at hr@cloudflare.com or via mail at 101 Townsend St. San Francisco, CA 94107.
Seattle, Washington, United States
Posted: 26/07/2026 | Deadline: 25/08/2026
About Us At Cloudflare, we are on a mission to help build a better Internet. Today the company runs one of the world’s largest networks that powers millions of websites and other Internet properties for customers ranging from individual bloggers to SMBs to Fortune 500 companies. Cloudflare protects and accelerates any Internet application online without adding hardware, installing software, or changing a line of code. Internet properties powered by Cloudflare all have web traffic routed through its intelligent global network, which gets smarter with every request. As a result, they see significant improvement in performance and a decrease in spam and other attacks. Cloudflare was named to Entrepreneur Magazine’s Top Company Cultures list and ranked among the World’s Most Innovative Companies by Fast Company. At Cloudflare, we’re not looking for people who wait for a polished roadmap; we’re looking for the builders who see the cracks in the Internet that everyone else has simply learned to live with. We value candidates who have the instinct to spot a "normalized" problem and the AI-native curiosity to create a solution using the latest tools. Our culture is built on iteration, leveraging AI to ship faster today to make it better tomorrow, while ensuring that every improvement, no matter how small, is shared across the team to lift everyone up. If you’re the type of person who values curiosity over bureaucracy, and that AI is a partner in solving tough problems to keep the Internet moving forward, you’ll fit right in. Software Engineer, AI Agents Location: India, Bangalore Exp: 5 - 10yrs Why This Role Matters At Cloudflare, we're building industrial-scale AI agents that support customers directly. This isn't research theater. Your code will power real customer interactions from day one, at global scale. Cloudflare already has the parts. You will assemble Workers, Durable Objects, KV, R2, D1, Vectorize, Workers AI, AI Gateway, and the Agent SDK into real agents customers use every day. Role Intent Ship production agents on the Cloudflare stack. Build, deploy, learn, repeat. Your code is the front door for Cloudflare customers. What You Will Do Build agents on Workers with Durable Objects for state and short term memory Wire tools with the Agent SDK, MCP, and function calling Use Vectorize, KV, R2, and D1 for semantic memory, cache, files, and config Run models through Workers AI and AI Gateway; integrate third parties when needed Create evals, guardrails, and audits. Measure, tune, re-ship fast Build agents that summarize, propose fixes, and escalate cleanly to humans Expose agent health and metrics in transparent dashboards. No mystery boxes Integrate with queues and webhooks; publish events on Queues or Pub/Sub Cut cost per case and time to first response. Prove it with data. Take end to end ownership including on call for what you ship (with team support) Design and maintain robust observability for distributed AI workflows, implementing structured logging and end-to-end tracing across async service boundaries to ensure visibility into agent reasoning and execution. Architect security boundaries for agent-led operations; implementing secure credential handling, multi-layer approval gates, and fine-grained trust scoping for mutative actions. Must Have Demonstrated success shipping production systems. Repos and releases that show real work. Strong in TypeScript or Rust on Workers. HTTP, queues, async, performance Fluency with Durable Objects, KV or R2, and either D1 or Postgres Hands on with model tooling. Prompt I/O, tool calling, evals, safety checks Observability mindset. Logs, traces, metrics, redlines Experience with a2a/multi-agent frameworks Experience developing LLM evaluation frameworks; automated scoring systems, CI-integrated quality gates. Bias for simple, scalable designs Nice to Have Workers AI, AI Gateway, and Vectorize in production Salesforce or Service Cloud experience. Webhooks and case APIs Security depth. Prompt injection protection, secrets detection, PII handling OSS agent frameworks. Know what to borrow and what to throw away. How We Build Align fast on what matters. Divide and conquer. Own your piece. Ship. Watch customers use it. Learn and repeat. Why Join Cloudflare in India? Impact at global scale: Your code will serve Cloudflare's customers across every region. Tens of millions of Internet properties depend on us. Work on the edge: Few companies give engineers the chance to build AI directly into an edge platform that runs in 300+ cities worldwide. Career growth: As one of the early engineers in our India based AI team, you'll have visibility, leadership opportunities, and a direct hand in shaping Cloudflare's AI roadmap. Culture of ownership: We believe in autonomy, accountability, and trust. Engineers here own outcomes, not just tickets. Learn and grow fast: Collaborate with peers across Support, Product, Security, and AI Platform teams. We encourage knowledge sharing, mentorship, and continuous learning. Interview Signal Expect to demonstrate your ability to: Build a mini agent on Workers using the Agent SDK Store session memory in Durable Objects Add semantic recall with Vectorize Ship behind a KV flag with traces and observability Push to production fast and take ownership Team Mission The Agent Tech team owns the end to end stack for customer facing agents on Cloudflare. Everything runs at the edge. Core Stack: Workers, Durable Objects, KV, R2, D1, Queues, Pub/Sub, Vectorize, Workers AI, AI Gateway, Pages, Zero Trust. Principles: Ship fast. Measure truth. Simplify relentlessly. Own outcomes. Fraud Alert: Do not fall victim to recruitment fraud. Cloudflare never charges application fees or requires candidates to purchase third-party certifications or training as a condition of employment. All official communication comes strictly from @cloudflare.com email addresses.What Makes Cloudflare Special? We’re not just a highly ambitious, large-scale technology company. We’re a highly ambitious, large-scale technology company with a soul. Fundamental to our mission to help build a better Internet is protecting the free and open Internet. Project Galileo: Since 2014, we've equipped more than 2,400 journalism and civil society organizations in 111 countries with powerful tools to defend themselves against attacks that would otherwise censor their work, technology already used by Cloudflare’s enterprise customers--at no cost. Athenian Project: In 2017, we created the Athenian Project to ensure that state and local governments have the highest level of protection and reliability for free, so that their constituents have access to election information and voter registration. Since the project, we've provided services to more than 425 local government election websites in 33 states. 1.1.1.1: We released 1.1.1.1 to help fix the foundation of the Internet by building a faster, more secure and privacy-centric public DNS resolver. This is available publicly for everyone to use - it is the first consumer-focused service Cloudflare has ever released. Here’s the deal - we don’t store client IP addresses never, ever. We will continue to abide by our privacy commitment and ensure that no user data is sold to advertisers or used to target consumers. Sound like something you’d like to be a part of? We’d love to hear from you! Please note that applicants who progress to the offer stage of the interview process may be asked to attend an in-person interview within one of the Cloudflare Offices or Cloudflare Hubs. More details about this will be available at that stage of the interview process. This position may require access to information protected under U.S. export control laws, including the U.S. Export Administration Regulations. Please note that any offer of employment may be conditioned on your authorization to receive software or technology controlled under these U.S. export laws without sponsorship for an export license. Cloudflare is proud to be an equal opportunity employer. We are committed to providing equal employment opportunity for all people and place great value in both diversity and inclusiveness. All qualified applicants will be considered for employment without regard to their, or any other person's, perceived or actual race, color, religion, sex, gender, gender identity, gender expression, sexual orientation, national origin, ancestry, citizenship, age, physical or mental disability, medical condition, family care status, or any other basis protected by law. We are an AA/Veterans/Disabled Employer. Cloudflare provides reasonable accommodations to qualified individuals with disabilities. Please tell us if you require a reasonable accommodation to apply for a job. Examples of reasonable accommodations include, but are not limited to, changing the application process, providing documents in an alternate format, using a sign language interpreter, or using specialized equipment. If you require a reasonable accommodation to apply for a job, please contact us via e-mail at hr@cloudflare.com or via mail at 101 Townsend St. San Francisco, CA 94107.
Bengaluru, Karnataka, India
Posted: 26/07/2026 | Deadline: 25/08/2026
Who we are About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies—from the world’s largest enterprises to the most ambitious startups—use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone’s reach while doing the most important work of your career. About the team In this role, you’ll join Stripe’s Vulnerability Management team, whose mission is to “Surface vulnerabilities at scale across Stripe.” Our vision is to create a culture of continuous excellence in managing vulnerabilities. The bug bounty program is an important pillar of this mission, acting as a critical line of defense in Stripe’s security “immune system.” What you’ll do We seek a highly technical and detail-oriented Security Analyst to join our team, focusing on the front lines of bug bounty triage and researcher engagement. In this role, you’ll be responsible for the end-to-end lifecycle of security vulnerability reports from our bug bounty program. You’ll own the overall effectiveness of Stripe’s bug bounty program with autonomy to implement continuous improvements (e.g., researcher campaigns, scoring transparency). You’ll play a key role in understanding the root cause of vulnerabilities, coordinating timely resolutions, and directly impacting the security posture of Stripe’s products. A core aspect of this role is developing a deep understanding of Stripe and acquired company products, assets, and their configuration to effectively assess and prioritize vulnerabilities. Responsibilities Analyze, assess, reproduce, and triage incoming security vulnerability reports from the bug bounty program Communicate clearly and effectively with security researchers to follow up on unclear reports, drive report clarity, and increase engagement with top hackers Understand the root cause of security vulnerabilities to help product and engineering teams fix them, and advise on the right mitigation strategies Drive the lifecycle of submissions through to resolution, coordinating with product and engineering stakeholders Act as the security bridge between external researchers and internal teams to facilitate rapid and effective remediation Conduct in-depth data analysis on bug reports and vulnerability patterns to identify systemic risks and inform new security initiatives Provide tactical support for vulnerability management triage processes to augment the team as needed Prepare and implement improvements to the overall bug bounty program Provide feedback and requirements for tool development to enhance triage and security workflows, leveraging opportunities for automation Who you are We’re looking for someone who meets the minimum requirements to be considered for the role. If you meet these requirements, you are encouraged to apply. The preferred qualifications are a bonus, not a requirement. Minimum requirements Proven ability to follow bug reports and accurately triage security vulnerabilities Familiarity with web security issues and exploit methodologies (e.g., OWASP Top 10, CWEs) Competent in offensive security tools (e.g., Burp Suite, custom scripting) Ability to think like an attacker to understand the impact of vulnerabilities Proficient in clear communication, conveying technical concepts to various stakeholders Experience in one of the following areas Bug bounty program or triaging security vulnerability reports Knowledge of Stripe products and general security expertise Preferred qualifications Experience in technical support, operations, or similar roles with technical systems exposure Prior participation in or experience with bug bounty programs Experience analyzing source code for security vulnerabilities Proficiency in scripting languages (e.g., Python, Ruby) for automation Familiarity with cloud-based services (e.g., AWS, GCP) Certifications such as OSWA or BSCP
Remote, North America, No Exact Location
Posted: 26/07/2026 | Deadline: 25/08/2026
Who we are At Twilio, we’re shaping the future of communications, all from the comfort of our homes. We deliver innovative solutions to hundreds of thousands of businesses and empower millions of developers worldwide to craft personalized customer experiences. Our dedication to remote-first work, and strong culture of connection and global inclusion means that no matter your location, you’re part of a vibrant team with diverse experiences making a global impact each day. As we continue to revolutionize how the world interacts, we’re acquiring new skills and experiences that make work feel truly rewarding. Your career at Twilio is in your hands.We use Artificial Intelligence (AI) to help make our hiring process efficient. That said, every hiring decision is made by real Twilions! .See yourself at Twilio Join the team as Twilio’s next Senior Principal Field Architect - AI Agents About the job The Senior Principal Field Engineer is a critical, highly visible leader on the Field CTO team who bridges the gap between Twilio's product vision and our customers' needs. In the rapidly evolving era of AI, Twilio takes an open and flexible approach: we want developers to use their preferred AI coding tools and agent builder platforms to get up and running extremely quickly with our platform of APIs across a variety of services that address many use cases. In this P7 Distinguished Engineer role, you will be in R&D and partner with Product, Engineering, Sales and GTM teams to spearhead the launch and adoption of our AI-native platform services. You will be the field engineering linchpin ensuring that the newly launched Twilio Conversations suite and Twilio Agent Connect (TAC) integrate flawlessly with the industry's top AI tools. Your mission is to ensure that insights from the field directly influence product strategy while actively building deep, functional partnerships with the companies driving the AI revolution. You will derisk adoption, increase partner integrations to accelerate adoption and create repeatable blueprints for mass adoption. Responsibilities In this role, you’ll: Ecosystem Engineering & Integration: Engineer seamless integrations between Twilio Agent Connect (TAC), the Twilio Conversations suite, and hyperscaler AI platforms (AWS Bedrock, Azure Foundry, GCP Vertex AI Agent Builder, and Meta's Business Agent platform for example). Champion the Builder’s Mindset: Demonstrate a deep, practical understanding of modern software development. You must know how developers are using AI coding tools like Claude, Gemini, and Cursor to build conversational AI (omnichannel customer support agents, for example) on top of Twilio today and in the future. Strategic Partnerships: Act as a technical liaison and partnership builder with our Partner Teams with leading AI tool and cloud companies, ensuring Twilio's APIs work perfectly within their ecosystems to accelerate our customers' time-to-market. Roadmap & Vision Leadership: Present Twilio's technical roadmap and vision to customers. Consolidate insights into structured inputs for planning to help shape R&D priorities. Sales & Go-to-Market Enablement: Support and provide direction to Sales and Specialist teams on the latest positioning, use cases, and technical roadmaps. Work with sales enablement to ensure teams are covering the right topics. Customer Discovery: Partner with the Specialist team to capture technical and business requirements with key enterprise accounts identified as design partners. Qualifications Twilio values diverse experiences from all kinds of industries, and we encourage everyone who meets the required qualifications to apply. If your career is just starting or hasn't followed a traditional path, don't let that stop you from considering Twilio. We are always looking for people who will bring something new to the table! *Required: 10+ years of engineering experience coupled with 8+ years in enterprise technology strategy technical product management, cloud engineering, or executive consulting roles. Cloud & AI Expertise: Deep architectural understanding of major cloud hyperscalers (AWS, Azure, GCP) and hands-on familiarity with conversational AI, chatbots, cognitive services, and machine learning architectures and engineering practices. Engineering Nuance: A strong technical background to discuss complex system architectures and integrate AI technologies. . Strategic Thinking & Leadership: Ability to develop and execute long-term solution architectures, roadmaps, and cross-functional initiatives at an enterprise scale with significant business impact. Technical & Business Acumen: Exceptional communication skills to articulate complex concepts to both highly technical developers and business-focused executives. Customer Focus & Agility: Proven ability to advocate for customer-centric engineering practices and thrive in a fast-paced, evolving environment with comfort operating in areas of ambiguity. Desired: Proven track record of supporting and enabling sales and partner ecosystems with product knowledge and GTM materials. Experience building modern digital customer engagement experiences using cloud and AI services to help customers identify and resolve issues. Advanced degree (Master's in Computer Science, Physics, or related field), or Bachelor's degree from an accredited university or equivalent experience. Location This role will be remote, but is not eligible to be hired in San Francisco, CA, Oakland, CA, San Jose, CA, or the surrounding areas. Travel We prioritize connection and opportunities to build relationships with our customers and each other. For this role, approximately 20% travel is anticipated to help you connect in-person in a meaningful way. What We Offer Working at Twilio offers many benefits, including competitive pay, generous time off, ample parental and wellness leave, healthcare, a retirement savings program, and much more. Offerings vary by location. Compensation Please note the salary range information provided applies only to candidates residing in California, Colorado, Hawaii, Illinois, Maryland, Massachusetts, Minnesota, New Jersey, New York, Vermont, Washington D.C., and Washington State due to local requirements. Compensation for candidates in other locations will be discussed during the hiring process. Please note that hiring for this role is not restricted to the locations listed above. The estimated pay ranges for this role are as follows: Based in Colorado, Hawaii, Illinois, Maryland, Massachusetts, Minnesota, Vermont or Washington D.C. : $284,320 - $355,400. Based in New York, New Jersey, Washington State, or California (outside of the San Francisco Bay area): $292,080 -- $365,100. Based in the San Francisco Bay area, California: $324,480 - $405,600. This role may be eligible to participate in Twilio’s equity plan and corporate bonus plan. All roles are generally eligible for the following benefits: health care insurance, 401(k) retirement account, paid sick time, paid personal time off, paid parental leave. The successful candidate’s starting salary will be determined based on permissible, non-discriminatory factors such as skills, experience, and geographic location. Application deadline information Applications for this role are intended to be accepted until August 3, 2026 but may change based on business needs. Twilio thinks big. Do you? We like to solve problems, take initiative, pitch in when needed, and are always up for trying new things. That's why we seek out colleagues who embody our values — something we call Twilio Magic. Additionally, we empower employees to build positive change in their communities by supporting their volunteering and donation efforts. So, if you're ready to unleash your full potential, do your best work, and be the best version of yourself, apply now! If this role isn't what you're looking for, please consider other open positions. Twilio is proud to be an equal opportunity employer. We do not discriminate based upon race, religion, color, national origin, sex (including pregnancy, childbirth, reproductive health decisions, or related medical conditions), sexual orientation, gender identity, gender expression, age, status as a protected veteran, status as an individual with a disability, genetic information, political views or activity, or other applicable legally protected characteristics. We also consider qualified applicants with criminal histories, consistent with applicable federal, state and local law. Qualified applicants with arrest or conviction records will be considered for employment in accordance with the Los Angeles County Fair Chance Ordinance for Employers and the California Fair Chance Act. Additionally, Twilio participates in the E-Verify program in certain locations, as required by law.
Remote, Remote, Remote
Posted: 26/07/2026 | Deadline: 25/08/2026
Grafana Labs, the company behind the open observability cloud, is founded on the principles of open source, open standards, open ecosystems, and open culture. Grafana Cloud, our fully managed observability platform, is flexible and built for scale. With Grafana Cloud's actually useful AI, organizations can see, understand, and act on all their disparate data to move at the speed of their ambitions. Today, more than 35 million users and 7,000+ customers – including Anthropic, Bloomberg, NVIDIA, Microsoft, and Salesforce – trust Grafana Labs to ensure reliability of their applications and systems, resolve incidents quickly, and optimize their telemetry to reduce noise and cost. We are a 100% remote company with 1,600+ team members across 40+ countries, and we’re backed by leading investors including Lightspeed Venture Partners, Sequoia Capital, GIC, Coatue, J.P. Morgan, CapitalG, and Lead Edge Capital. Learn more at grafana.com and follow us on LinkedIn and X. We’re scaling fast and staying true to what makes us different: an open-source legacy, a global collaborative culture, and a passion for meaningful work. Our team thrives in an innovation-driven environment where transparency, autonomy, and trust fuel everything we do. You may not meet every requirement, and that’s okay. If this role excites you, we’d love you to raise your hand for what could be a truly career-defining opportunity.This is a remote position. We are looking for candidates in the UK, Germany, Spain, Sweden and Ireland only. The Opportunity We build Pyroscope, the open-source continuous profiling database behind Grafana Cloud Profiles. Pyroscope gives engineers code-level visibility into how their applications use CPU and memory, down to the specific line of code, and connects that signal with metrics, logs, and traces across the Grafana stack. 2026 is an inflection point for Pyroscope. We completed the rollout of our V2 architecture, cutting production costs by 92% while significantly improving scalability, and grew the number of customers using the product by 64% last year. The product is moving beyond profiling-savvy users, and our focus is shifting from standalone feature development to reducing friction across the full experience: better onboarding, deeper integration with the rest of Grafana, and profiles that work for operational and SRE workflows, not just performance specialists. Over the next year, you will help us: Ship Adaptive Profiles as the default ingestion strategy, so customers automatically collect the profiles they need at a cost that makes sense. Turn Pyroscope into a platform capability inside Grafana: bi-directional trace-to-profile correlation, integration with Kubernetes Monitoring and App Observability, and profiles surfaced where engineers already start their investigations. Prepare Pyroscope for an agent-driven world: APIs, CLI, and docs designed so AI agents can drive profiling end to end, covering the full performance optimization lifecycle from finding an issue to verifying the fix. Push operational excellence and TCO further: generalized autoscaling, better UX for long queries, and BYOC automation that brings up a new cell with zero manual intervention. Double adoption by making profiling easy to start with: guided onboarding, use-case-driven docs, and assisted instrumentation so users get value in minutes instead of days. What You'll Be Doing As a Senior Engineer on Pyroscope, you will own meaningful projects end to end and help shape what the team builds next. Lead projects from concept to rollout, e.g. Read path efficiency improvement, or trace-to-profile correlation, including design, delivery, operations, and customer follow-up. Design, build, and operate core components of a distributed database: ingestion, storage, and query, making sharp trade-offs on performance, cost, and complexity. Wear the product hat. We have no dedicated PM: engineers on this team join customer calls, translate pain points into manageable deliverables, and directly shape the roadmap. Your questions and ideas will steer what we build. Drive operational excellence. Own outcomes against concrete SLOs and unit cost targets, reduce toil through automation, and make on-call quieter every quarter. Partner across Grafana. Work closely with App Observability, Alloy, Tempo, and the other Databases squads to make profiles useful wherever engineers work. Support your teammates through design conversations, code review, and pairing in a fully remote setup. Participate in on-call (EMEA rotation) for the services you build, and treat incident response and post-incident learning as part of the craft. Contribute to open source. Pyroscope is OSS. You will engage the community, review external contributions, and help steer the project in the open. We invest heavily in developer productivity. You can use modern AI coding assistants as part of your daily workflow (your choice of tools, within security guidelines), backed by a company-funded usage budget so you can iterate quickly without unnecessary friction. We encourage pragmatic AI-assisted development: faster prototyping, test generation, refactors, documentation, and incident follow-ups, always paired with strong code review and quality standards. You'll also have access to frontier models (e.g., GPT-5.6, Claude Fable, Gemini Pro 3.1). Example problems you could work on These are the kinds of projects landing this year: Adaptive Profiles: sampling and aggregation strategies that keep the signal while cutting cost and noise, plus the ingestion metrics that make its behavior transparent to customers. Large queries: query fairness, async execution, and autoscaling of the read path so long queries from our largest tenants run efficiently at any scale. Traces and profiles together: bi-directional correlation between traces and profiles, so a slow request links straight to the code that burned the CPU. Agent-ready Pyroscope: structured, deterministic APIs and a CLI that AI assistants can drive reliably, plus the benchmarks to prove tool quality and catch regressions. BYOC and regions: cell lifecycle automation, from provisioning to migration, targeting a new cell end to end with zero manual steps. OTel-native profiling: evolve ingestion, storage, and query as OpenTelemetry profiling matures, keeping Pyroscope the natural backend for OTel profiles. What Makes You a Great Fit Solid experience with a systems language. We write Pyroscope in Go; experience with one or more programming languages (e.g. Rust, C, C++, Python, Java, etc). Distributed systems in production. You have built and operated cloud services and understand what it takes to keep a multi-tenant data system fast, reliable, and affordable. Product sense. You are comfortable talking to customers, sitting with ambiguity, and breaking fuzzy problems into manageable deliverables. Since we have no dedicated PM, this matters as much as your code. Curiosity and courage. Pyroscope is nearing feature completeness while still refining product-market fit. We need someone who asks hard questions and challenges the status quo when appropriate. Strong software craftsmanship. You write clean, robust, performant software that others can maintain, and you know when to optimize versus when to ship. Operational mindset. You have carried a pager, done SRE-style work or infrastructure as code, and treat reliability as a feature. Pragmatism. You break complex problems into short feedback loops: analyze, design, deliver an MVP, learn, iterate. Clear communication. You work well in a fully remote, asynchronous environment and lead through writing, reviews, and shipped code. Bonus Points For Experience with profiling and performance engineering: flamegraphs, pprof, perf, or similar tooling. Experience with OpenTelemetry or large-scale observability systems. Experience operating multi-tenant SaaS infrastructure at scale on Kubernetes. Experience building for AI/LLM consumers: structured APIs, metadata and discovery endpoints, deterministic outputs. Open-source contribution or maintainership, and comfort engaging a community in the open. Experience as an on-call user of Grafana, Prometheus, Pyroscope, or similar in a previous role (or on a homelab). Experience in a fully remote, globally distributed team. How we work We are a remote-first team that meets regularly over video and does most of our work asynchronously, in writing. We value creativity, diverse perspectives, and clear communication. Pyroscope is relied upon by prominent global organizations to optimize critical applications and infrastructure, and we expect everyone on the team to contribute ideas that make it a more reliable, more useful, and more loved product. In UK, the compensation range for this role is £91,755 - £110,106. Actual compensation may vary based on level, experience, and skillset as assessed throughout the interview process. All of our roles include Restricted Stock Units (RSUs), giving every team member ownership in Grafana Labs' success. We believe in shared outcomes—RSUs help us stay aligned and invested as we scale globally. *Compensation ranges are country specific. If you are applying for this role from a different location than listed above, your recruiter will discuss your specific market’s defined pay range & benefits at the beginning of the process. Why You’ll Thrive at Grafana Labs: 100% Remote, Global Culture - As a remote-only company, we bring together talent from around the world, united by a culture of collaboration and shared purpose. Scaling Organization – Tackle meaningful work in a high-growth, ever-evolving environment. Transparent Communication – Expect open decision-making and regular company-wide updates. Innovation-Driven – Autonomy and support to ship great work and try new things. Open Source Roots – Built on community-driven values that shape how we work. Empowered Teams – High trust, low ego culture that values outcomes over optics. Career Growth Pathways – Defined opportunities to grow and develop your career. Approachable Leadership – Transparent execs who are involved, visible, and human. Passionate People – Join a team of smart, supportive folks who care deeply about what they do. In-Person onboarding - We want you to thrive from day 1 with your fellow new ‘Grafanistas’ to learn all about what we do and how we do it. Balance is Key - We operate a global annual leave policy of 30 days per annum. 3 days of your annual leave entitlement are reserved for Grafana Shutdown Days to allow the team to really disconnect. *We will comply with local legislation where applicable. Equal Opportunity Employer: Grafana Labs is an equal opportunities employer. We welcome applications from everyone regardless of race, colour, nationality, origin, caste, sex, gender reassignment identity or expression, sexual orientation, age, religion or belief, disability, veteran status, genetic information, pregnancy, maternity, marital, family or carer status, or any other characteristic which is protected by local law. We believe that equality and diversity build a strong organisation, and we work hard to ensure that is the foundation of our organisation as we grow. Grafana Labs may utilize AI tools in its recruitment process to assist in matching information provided in CVs to job postings. The recruitment team will continue to review inbound CVs manually to identify alignment with current openings. #LI-Remote For information about how your personal data is used once you’ve applied to a job, check out our privacy policy.
United Kingdom (Remote), No Exact Region, No Exact Location
Posted: 26/07/2026 | Deadline: 25/08/2026
About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems.About the role Data Infrastructure designs, operates, and scales secure, privacy-respecting systems that power data-driven decisions across Anthropic. Our mission is to provide data processing, storage, and access that are trusted, fast, and easy to use. We're looking for infrastructure engineers who thrive working at the intersection of data systems, security, and scalability. You'll tackle diverse challenges ranging from building financial reporting pipelines to architecting access control systems to ensuring cloud storage reliability. This role offers the opportunity to work directly with data scientists, analysts, and business stakeholders while diving deep into cloud infrastructure primitives. Responsibilities: Within Data Infra, you may be matched to critical business areas including: Data Governance & Access Control: Design and implement robust access control systems ensuring only authorized users can access sensitive data. Build infrastructure for permission management, audit logging, and compliance requirements. Work on IAM policies, ACLs, and security controls that scale across thousands of users and systems. Financial Data Infrastructure: Build and maintain data pipelines and warehouses powering business-critical reporting. Ensure data integrity, accuracy, and availability for complex financial systems, including third party revenue ingestion pipelines; manage the external relationships as needed to drive upstream dependencies. Own the reliability of systems processing revenue, usage, and business metrics. Cloud Storage & Reliability: Architect disaster recovery, backup, and replication systems for petabyte-scale data. Ensure high availability and durability of data stored in cloud object storage (GCS, S3). Build systems that protect against data loss and enable rapid recovery. Data Platform & Tooling: Scale data processing infrastructure using technologies like BigQuery, BigTable, Airflow, dbt, and Spark. Optimize query performance, manage costs, and enable self-service analytics across the organization. You might be a good fit if you: Have 10+ years (not including internships or co-ops) of experience in a Software Engineer role, building data infrastructure, storage systems, or related distributed systems Have 3+ years (not including internships or co-ops) of experience leading large scale, complex projects or teams as an engineer or tech lead Can set technical direction for a team, not just execute within it Have deep experience with at least one of: Strong proficiency in programming languages like Python, Go, Java, or similar Experience with infrastructure-as-code (Terraform, Pulumi) and cloud platforms (GCP, AWS) Can navigate complex technical tradeoffs between performance, cost, security, and maintainability Have excellent collaboration skills - you work well with both technical and non-technical stakeholders Strong candidates may also have: Experience with security and compliance requirements (ITGC, GDPR, financial controls) Background in data warehousing, ETL/ELT pipelines, or analytics infrastructure Experience with Kubernetes, containerization, and cloud-native architectures Track record of improving data reliability, availability, or cost efficiency at scale Knowledge of column-oriented databases, OLAP systems, or big data processing frameworks Experience working in fintech, financial services, or highly regulated environments Security engineering background with focus on data protection and access controls Technologies We Use: Data: BigQuery, BigTable, Airflow, Cloud Composer, dbt, Spark, Segment, Fivetran Storage: GCS, S3 Infrastructure: Terraform, Kubernetes, GCP, AWS Languages: Python, Go, SQL Deadline to apply: None. Applications will be reviewed on a rolling basis.The annual compensation range for this role is listed below. For sales roles, the range provided is the role’s On Target Earnings ("OTE") range, meaning that the range includes both the sales commissions/sales bonuses target and annual base salary for the role.Annual Salary:$405,000—$485,000 USDLogistics Minimum education: Bachelor’s degree or an equivalent combination of education, training, and/or experience Required field of study: A field relevant to the role as demonstrated through coursework, training, or professional experience Minimum years of experience: Years of experience required will correlate with the internal job level requirements for the position Location-based hybrid policy: Currently, we expect all staff to be in one of our offices at least 25% of the time. However, some roles may require more time in our offices. Visa sponsorship: We do sponsor visas! However, we aren't able to successfully sponsor visas for every role and every candidate. But if we make you an offer, we will make every reasonable effort to get you a visa, and we retain an immigration lawyer to help with this. We encourage you to apply even if you do not believe you meet every single qualification. Not all strong candidates will meet every single qualification as listed. Research shows that people who identify as being from underrepresented groups are more prone to experiencing imposter syndrome and doubting the strength of their candidacy, so we urge you not to exclude yourself prematurely and to submit an application if you're interested in this work. We think AI systems like the ones we're building have enormous social and ethical implications. We think this makes representation even more important, and we strive to include a range of diverse perspectives on our team.Your safety matters to us. To protect yourself from potential scams, remember that Anthropic recruiters only contact you from @anthropic.com email addresses. In some cases, we may partner with vetted recruiting agencies who will identify themselves as working on behalf of Anthropic. Be cautious of emails from other domains. Legitimate Anthropic recruiters will never ask for money, fees, or banking information before your first day. If you're ever unsure about a communication, don't click any links—visit anthropic.com/careers directly for confirmed position openings. How we're different We believe that the highest-impact AI research will be big science. At Anthropic we work as a single cohesive team on just a few large-scale research efforts. And we value impact — advancing our long-term goals of steerable, trustworthy AI — rather than work on smaller and more specific puzzles. We view AI research as an empirical science, which has as much in common with physics and biology as with traditional efforts in computer science. We're an extremely collaborative group, and we host frequent research discussions to ensure that we are pursuing the highest-impact work at any given time. As such, we greatly value communication skills. The easiest way to understand our research directions is to read our recent research. This research continues many of the directions our team worked on prior to Anthropic, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute, Concrete Problems in AI Safety, and Learning from Human Preferences. Come work with us! Anthropic is a public benefit corporation headquartered in San Francisco. We offer competitive compensation and benefits, optional equity donation matching, generous vacation and parental leave, flexible working hours, and a lovely office space in which to collaborate with colleagues. Guidance on Candidates' AI Usage: Learn about our policy for using AI in our application process.
San Francisco, CA | Seattle, WA
Posted: 27/07/2026 | Deadline: 26/08/2026
About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems.About the Role Anthropic's Infrastructure organization is foundational to our mission of developing AI systems that are reliable, interpretable, and steerable. The systems we build determine how quickly we can train new models, how reliably we can run safety experiments, and how effectively we can scale Claude to millions of users — demonstrating that safe, reliable infrastructure and frontier capabilities can go hand in hand. Developer Productivity owns the end-to-end experience of how engineers and researchers at Anthropic develop, build, test, and ship code at scale — from the source control and language ecosystems that underpin our monorepo, to the build and CI infrastructure that keeps thousands of daily builds running reliably across multiple cloud providers, to the developer acceleration tooling that deeply integrates Claude into engineering workflows. Team Matching: Team matching is determined after the interview process based on interview performance, interests, and business priorities. Please note we may also consider you for different Infrastructure teams. Responsibilities: Own the technical strategy and roadmap for your area, translating team-level goals into concrete execution plans Define infrastructure architecture, ensuring the hardest problems get solved — whether by you directly or by working through others Design and build scalable, reliable distributed infrastructure and shared libraries that support high-volume workloads across all engineering teams Own and evolve build environments, package management, and dependency systems to enable fast, reproducible builds Define and implement language ecosystem standards, tooling, and frameworks that drive developer productivity across research and production workloads You may be a good fit if you: Have deep experience with build systems, CI/CD pipelines, and/or developer tooling in a large monorepo environment Have strong proficiency in Python, Rust and/or Go Are obsessed with developer productivity and reducing friction in the software development lifecycle Have experience with container orchestration and infrastructure at scale Have excellent communication skills and enjoy supporting internal partners to improve their development experience Are excited about designing foundational systems and are comfortable working independently on ambiguous, high-impact technical challenges Strong candidates may have: 15+ years (not including internships or co-ops) of experience in a Software Engineer role, building and operating large-scale developer infrastructure 3+ years (not including internships or co-ops) of experience leading large scale complex projects or teams as a tech lead Experience with CI orchestration tools (Buildkite, Jenkins, GitHub Actions, or similar) and merge queue management at scale Experience building or operating remote build execution systems (Bazel Remote Execution API, BuildBarn, BuildBuddy, or similar) Experience with Nix/NixOS/Docker and managing large image / package sets at scale Experience building CLI tools, developer-facing services, and GitHub API and automation workflows Deadline to apply: None. Applications will be reviewed on a rolling basis.The annual compensation range for this role is listed below. For sales roles, the range provided is the role’s On Target Earnings ("OTE") range, meaning that the range includes both the sales commissions/sales bonuses target and annual base salary for the role.Annual Salary:$405,000—$625,000 USDLogistics Minimum education: Bachelor’s degree or an equivalent combination of education, training, and/or experience Required field of study: A field relevant to the role as demonstrated through coursework, training, or professional experience Minimum years of experience: Years of experience required will correlate with the internal job level requirements for the position Location-based hybrid policy: Currently, we expect all staff to be in one of our offices at least 25% of the time. However, some roles may require more time in our offices. Visa sponsorship: We do sponsor visas! However, we aren't able to successfully sponsor visas for every role and every candidate. But if we make you an offer, we will make every reasonable effort to get you a visa, and we retain an immigration lawyer to help with this. We encourage you to apply even if you do not believe you meet every single qualification. Not all strong candidates will meet every single qualification as listed. Research shows that people who identify as being from underrepresented groups are more prone to experiencing imposter syndrome and doubting the strength of their candidacy, so we urge you not to exclude yourself prematurely and to submit an application if you're interested in this work. We think AI systems like the ones we're building have enormous social and ethical implications. We think this makes representation even more important, and we strive to include a range of diverse perspectives on our team.Your safety matters to us. To protect yourself from potential scams, remember that Anthropic recruiters only contact you from @anthropic.com email addresses. In some cases, we may partner with vetted recruiting agencies who will identify themselves as working on behalf of Anthropic. Be cautious of emails from other domains. Legitimate Anthropic recruiters will never ask for money, fees, or banking information before your first day. If you're ever unsure about a communication, don't click any links—visit anthropic.com/careers directly for confirmed position openings. How we're different We believe that the highest-impact AI research will be big science. At Anthropic we work as a single cohesive team on just a few large-scale research efforts. And we value impact — advancing our long-term goals of steerable, trustworthy AI — rather than work on smaller and more specific puzzles. We view AI research as an empirical science, which has as much in common with physics and biology as with traditional efforts in computer science. We're an extremely collaborative group, and we host frequent research discussions to ensure that we are pursuing the highest-impact work at any given time. As such, we greatly value communication skills. The easiest way to understand our research directions is to read our recent research. This research continues many of the directions our team worked on prior to Anthropic, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute, Concrete Problems in AI Safety, and Learning from Human Preferences. Come work with us! Anthropic is a public benefit corporation headquartered in San Francisco. We offer competitive compensation and benefits, optional equity donation matching, generous vacation and parental leave, flexible working hours, and a lovely office space in which to collaborate with colleagues. Guidance on Candidates' AI Usage: Learn about our policy for using AI in our application process.
San Francisco, California, United States
Posted: 27/07/2026 | Deadline: 26/08/2026