I'm looking for a job. I exported this list from FewerJobs.com - a curated job board. Please: 1. Rank these jobs by fit for me, given my resume / skills. 2. Highlight the top 5 with a one-sentence rationale each. 3. Flag any concerns, including benefit values without source-backed evidence. 4. Suggest one or two filter changes I could make on FewerJobs to find more good matches. Filters I applied: - q: Prometheus - quality_floor: all - match_401k_strict: true - parental_strict: true - non_birth_strict: true - pto_strict: true - include_older: false - apply_url_verified: false - page: 1 - per_page: 100 - sort: relevance Jobs (100 total): --- TITLE: Manager, Product Management - Prometheus Financial Core (PFC) EMPLOYER: Capital One Financial Corporation LOCATION: 2 Locations (unspecified) SALARY: Not disclosed POSTED: 2026-05-13 K401_MATCH: yes (not source-backed) APPLY_URL: https://capitalone.wd12.myworkdayjobs.com/Capital_One/job/McLean-VA/Manager--Product-Management---Prometheus-Financial-Core--PFC-_R242443-1 EXCERPT: Manager, Product Management - Prometheus Financial Core (PFC) 2 Locations posted: Posted 30 Days Ago --- TITLE: MTS - Opportunistic Test EMPLOYER: Prometheus LOCATION: Zurich | OnSite | London (onsite) SALARY: Not disclosed POSTED: 2026-06-12 APPLY_URL: https://jobs.ashbyhq.com/prometheus/d5778f9e-c962-456d-a37a-7f17e26083ac EXCERPT: MTS - Opportunistic Test Zurich | OnSite | London TEST --- TITLE: Infrastructure (Rust, Devops) EMPLOYER: Backpack LOCATION: Remote (remote) SALARY: Not disclosed POSTED: 2026-06-10 APPLY_URL: https://jobs.ashbyhq.com/backpack/9a92dd28-3e3f-47c0-ab53-c0eeb79ecdae EXCERPT: Infrastructure (Rust, Devops) Remote YOU WILL: - Design, implement, and maintain infrastructure tools and services in Rust - Own CI/CD pipelines, monitoring, and deployment automation - Collaborate with engineering and security teams to ensure scalable and resilient systems - Optimize infrastructure for performance, cost, and security - Help shape internal developer experience and workflows REQUIREMENTS: - At least 3+ years of experience in DevOps, SRE, or infrastructure engineering - Strong proficiency in Rust - Experience managing cloud infrastructure (AWS, GCP, or similar) - Proficiency with containerization, orchestration, and observability tools (e.g., Docker, Kubernetes, Prometheus, etc.) - Obsessive about reliability, automation, and elegant internal tooling - Very fluent in English WHAT YOU'LL GET: - Generous ownership & stake in the company. We're looking for owners. - A fast-paced environment. - Flexible working hours & hybrid schedule. LOCATION: We have offices in Chicago, Dubai, and Tokyo. You're welcome to join us there and we'll handle your visa. Remote is also acceptable for the right candidate. --- TITLE: Advanced Technology: AI/ML Research Scientist EMPLOYER: Cerebras Systems LOCATION: Sunnyvale, CA; Toronto, Ontario, Canada; Vancouver, British Columbia, Canada (unspecified) SALARY: Not disclosed POSTED: 2026-04-06 K401_MATCH: yes (not source-backed) APPLY_URL: https://job-boards.greenhouse.io/cerebrassystems/jobs/7691353003 EXCERPT: Advanced Technology: AI/ML Research Scientist Sunnyvale, CA; Toronto, Ontario, Canada; Vancouver, British Columbia, Canada Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. Our novel wafer-scale architecture provides the AI compute power of dozens of GPUs on a single chip, with the programming simplicity of a single device. This approach allows Cerebras to deliver industry-leading training and inference speeds and empowers machine learning users to effortlessly run large-scale ML applications, without the hassle of managing hundreds of GPUs or TPUs. Cerebras' current customers include top model labs, global enterprises, and cutting-edge AI-native startups. OpenAI recently announced a multi-year partnership with Cerebras , to deploy 750 megawatts of scale, transforming key workloads with ultra high-speed inference. Thanks to the groundbreaking wafer-scale architecture, Cerebras Inference offers the fastest Generative AI inference solution in the world, over 10 times faster than GPU-based hyperscale cloud inference services. This order of magnitude increase in speed is transforming the user experience of AI applications, unlocking real-time iteration and increasing intelligence via additional agentic computation. About The Team Cerebras builds wafer-scale AI processors-single chips delivering tens of PB/s of memory bandwidth and a dataflow architecture that accelerates at a granularity no multi-device system can match. The Advanced Technology Group (ATG) is Cerebras ' pathfinding organization. We work ahead of product to explore new architectures, demonstrate breakthrough performance on scientific and AI workloads, and shape the technical roadmap for future Cerebras hardware and software. Our work regularly appears at top-tier venues (Supercomputing, SIAM, IEEE, --- TITLE: Research Scientist - LTX Model Applications EMPLOYER: Lightricks LOCATION: Jerusalem (unspecified) SALARY: Not disclosed POSTED: 2026-03-01 APPLY_URL: https://careers.lightricks.com/position?gh_jid=8439227002&gh_jid=8439227002 EXCERPT: Research Scientist - LTX Model Applications Jerusalem Who we are Lightricks is an AI-first company creating next-generation content creation technology for businesses, enterprises, and studios with a mission to bridge the gap between imagination and creation. At our core is LTX-2, an open-source generative video model, built to deliver expressive, high-fidelity video at unmatched speed. It powers both our own products and a growing ecosystem of partners through API integration. The company is also known globally for pioneering consumer creativity through products like Facetune, one of the world's most recognized creative brands, which helped introduce AI-powered visual expression to hundreds of millions of users worldwide. We combine deep research, user-first design, and end-to-end execution from concept to final render to bring the future of expression to all. The role We are seeking a researcher to join the LTX Model Applications team, focusing on advancing and extending the capabilities of our LTX-Video models. You will design and explore new use cases, train novel ways to control the generation process, and push the boundaries of generative video technology. Your work will directly influence millions of users through our apps and contribute to the broader research and open-source communities, shaping the next generation of creative tools. What you will be doing - Advance and extend the capabilities of our LTX-Video models through new controls, features, and use cases. - Design and implement machine learning models for video, image, and audio generation, as well as multimodal applications. - Conduct experiments, prototype concepts, and validate --- TITLE: Developer Evangelist EMPLOYER: LiteLLM LOCATION: San Francisco | OnSite (onsite) SALARY: Not disclosed POSTED: 2026-06-10 APPLY_URL: https://jobs.ashbyhq.com/litellm/eb38a78c-c374-4f87-a860-7442e7824b36 EXCERPT: Developer Evangelist San Francisco | OnSite LiteLLM is the world's most popular AI Gateway used by the largest companies (Adobe, Netflix, NASA, etc.) in the world to give their developers access to LLMs and adjacent services (MCP's, Vector Stores, etc.). WHY DO COMPANIES USE LITELLM ENTERPRISE Companies use LiteLLM Enterprise once they put LiteLLM into production and need enterprise features like Prometheus metrics (production monitoring) and need to give LLM access to a large number of people with SSO (secure sign on) or JWT (JSON Web Tokens). WHAT YOU WILL BE WORKING ON We are looking for someone who is passionate about AI and developer infrastructure that can cultivate a thriving developer community around our product. You will define and drive the strategy for building the developer community. You will work regularly with customers and business team to define requirements and articulate them to our technical team. You will engage with our customers in both technical pre-sales and post-sales capacities, ensuring they are happy and engaged with our product. WHAT YOU WILL DO - Engage in technical conversations in our development community - with both existing and prospective customers. - Generate blog posts and other content to build our user community to get developers to discover and use our product - Provide world-class customer support for our enterprise customers and for our open source customers via forums, email, social networks and wherever else technical questions arise. - Meet with developers at hackathons, meetups, conferences and other events, both local and --- TITLE: Senior Applied Scientist , Alexa AI Aurora EMPLOYER: Amazon LOCATION: Bellevue, Washington, USA (unspecified) SALARY: $167K-$226K POSTED: 2026-03-19 PARENTAL_LEAVE_WEEKS: 6 (not source-backed) NON_BIRTH_PARENT_LEAVE_WEEKS: 6 (not source-backed) K401_MATCH: yes (not source-backed) FERTILITY_FAMILY_BUILDING_BENEFITS: yes (not source-backed) ADOPTION_ASSISTANCE_OFFERED: yes (not source-backed) MENTAL_HEALTH_SUPPORT: yes (not source-backed) CHILDCARE_SUBSIDY: yes (not source-backed) APPLY_URL: https://www.amazon.jobs/en/jobs/10371758/senior-applied-scientist-alexa-ai-aurora EXCERPT: Senior Applied Scientist , Alexa AI Aurora Bellevue, Washington, USA The Alexa AI, AURORA: Alexa Understanding, Runtime, ORchestration, and Applied sciences org is seeking a passionate, talented, and resourceful Senior Applied Scientist to invent and build scalable solutions for sate-of-the-art conversational AI. You will be working on advancing technologies in the fields of LLM, Artificial Intelligence (AI), Natural Language Processing (NLP). As part of this role, you will collaborate with talented peers, have significant influence on our overall strategy as you guide and mentor Applied Scientists to create innovative and scalable solutions that touch millions of Alexa customers. Creating reliable, scalable, and high performance products requires exceptional technical expertise, a sound understanding of the fundamentals of AI, and practical experience building large-scale machine learning systems. The ideal candidate will be a self-starter who can dive into a project with limited guidance and is able to design and implement inventive, simple solutions to complex problems. They will be passionate about new technologies and have a track record of delivering valuable software features and products in a fast-paced, highly iterative environment. A commitment to team work, hustle, and strong communication skills (to both business and technical partners) are absolute requirements. Join us in our mission to shape the space of Generative AI and provide unparalleled experiences for our customers. Key job responsibilities • Define the science roadmap and advance core science primitives for conversation modelling, content generation, and automated quality assurance via state-of-the-art computer vision, machine learning, and generative AI • Architect --- TITLE: DevOps Engineer EMPLOYER: Bank of Nova Scotia LOCATION: Toronto (unspecified) SALARY: Not disclosed POSTED: 2026-06-12 K401_MATCH: yes (not source-backed) APPLY_URL: https://jobs.scotiabank.com/job/Toronto-DevOps-Engineer-ON-M5H-1H1/600857917/ EXCERPT: DevOps Engineer Toronto Requisition ID: 255193 Join a purpose driven winning team, committed to results, in an inclusive and high-performing culture. The Role This Devops Engineer will contribute in ensuring specific individual goals, plans, initiatives are executed / delivered in support of the team's business strategies and objectives. Ensures all activities conducted are in compliance with governing regulations, internal policies and procedures. Is this role right for you? In this role you will: Build, maintain, and optimize Continuous Integration and Continuous Deployment pipelines. Automate build, test, and deployment workflows. Use tools like Terraform, Ansible, to provision and manage infrastructure. Enforce version‑controlled, repeatable infrastructure deployments. Manage cloud environments (On-Prem, GCP), ensuring scalability, cost efficiency, and resilience. Implement container platforms (Kubernetes, GKE, RKE2) and container‑based workflows. Implement observability stacks (Dynatrace, Prometheus, Grafana, Loki). Optimize system performance and drive SRE‑aligned practices such as SLIs/SLOs. Reduce manual work through automation across build, deployment, infra, configuration, and validation workflows. Evaluate and integrate DevOps tools to improve developer productivity and operational excellence. Ensure compliance with organizational governance, audit requirements, and cloud security posture. Partner with development, architecture, cloud, and security teams. Support developers by improving b --- TITLE: Senior/Staff Engineer : Post Silicon- Bring Up EMPLOYER: Cerebras Systems LOCATION: Bengaluru, Karnataka, India; Sunnyvale, CA; Toronto, Ontario, Canada (unspecified) SALARY: $175K-$275K POSTED: 2026-02-16 K401_MATCH: yes (not source-backed) APPLY_URL: https://job-boards.greenhouse.io/cerebrassystems/jobs/7628233003 EXCERPT: Senior/Staff Engineer : Post Silicon- Bring Up Bengaluru, Karnataka, India; Sunnyvale, CA; Toronto, Ontario, Canada Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. Our novel wafer-scale architecture provides the AI compute power of dozens of GPUs on a single chip, with the programming simplicity of a single device. This approach allows Cerebras to deliver industry-leading training and inference speeds and empowers machine learning users to effortlessly run large-scale ML applications, without the hassle of managing hundreds of GPUs or TPUs. Cerebras' current customers include top model labs, global enterprises, and cutting-edge AI-native startups. OpenAI recently announced a multi-year partnership with Cerebras , to deploy 750 megawatts of scale, transforming key workloads with ultra high-speed inference. Thanks to the groundbreaking wafer-scale architecture, Cerebras Inference offers the fastest Generative AI inference solution in the world, over 10 times faster than GPU-based hyperscale cloud inference services. This order of magnitude increase in speed is transforming the user experience of AI applications, unlocking real-time iteration and increasing intelligence via additional agentic computation. The Role: In this exciting role, you will be responsible for bring up and optimizations of Cerebras's Wafer Scale Engine (WSE). Suitable candidate will have experience delivering end to end solutions working closely with teams across chip design, system performance, software development and productization. Responsibilities: - On Wafer Scale Engines, develop and debug flows that embed well tested and deployable optimizations in production processes to reduce time and costs - Work on refining AI Systems across H/W-S/W --- TITLE: AI DevOps Engineer EMPLOYER: Moody's Corporation LOCATION: Location not specified (unspecified) SALARY: Not disclosed POSTED: 2026-04-27 LEARNING_BUDGET_OFFERED: yes (not source-backed) APPLY_URL: https://careers.moodys.com/en EXCERPT: AI DevOps Engineer At Moody's, we unite the brightest minds to turn today's risks into tomorrow's opportunities. We do this by striving to create an inclusive environment where everyone feels welcome to be who they are-with the freedom to exchange ideas, think innovatively, and listen to each other and customers in meaningful ways. Moody's is transforming how the world sees risk. As a global leader in ratings and integrated risk assessment, we're advancing AI to move from insight to action-enabling intelligence that not only understands complexity but responds to it. We decode risk to unlock opportunity, helping our clients navigate uncertainty with clarity, speed, and confidence. If you are excited about this opportunity but do not meet every single requirement, please apply! You still may be a great fit for this role or other open roles. We are seeking candidates who model our values: invest in every relationship, lead with curiosity, champion diverse perspectives, turn inputs into actions, and uphold trust through integrity. Skills & Competencies 4+ years of experience with Terraform for provisioning and managing cloud resources. Hands-on experience with AWS and Azure; familiarity with GCP is a plus. Proficiency in creating and maintaining CI/CD pipelines using tools such as GitHub Actions, Jenkins, or Azure DevOps. Experience with Docker and container orchestration. Strong skills in scripting languages (Python, Bash) and configuration management tools (Ansible, Chef, or Puppet). Familiarity with tools like Splunk, Prometheus, Grafana, ELK stack, or equivalent. Understanding of secure infrastructure practices, identity management, and encryption. Proficient with --- TITLE: Applied Scientist - Perception (SLAM/VIO), Fauna EMPLOYER: Amazon LOCATION: New York, New York, USA (unspecified) SALARY: $172K-$223K POSTED: 2026-05-08 PARENTAL_LEAVE_WEEKS: 6 (not source-backed) NON_BIRTH_PARENT_LEAVE_WEEKS: 6 (not source-backed) K401_MATCH: yes (not source-backed) FERTILITY_FAMILY_BUILDING_BENEFITS: yes (not source-backed) ADOPTION_ASSISTANCE_OFFERED: yes (not source-backed) MENTAL_HEALTH_SUPPORT: yes (not source-backed) CHILDCARE_SUBSIDY: yes (not source-backed) APPLY_URL: https://www.amazon.jobs/en/jobs/10415069/applied-scientist-perception-slam-vio-fauna EXCERPT: Applied Scientist - Perception (SLAM/VIO), Fauna New York, New York, USA We are seeking an Applied Scientist to develop and optimize Visual Inertial Odometry (VIO) and sensor fusion systems for our intelligent robots. In this role, you will design, implement, and deploy state estimation and tracking algorithms that enable robots to understand their position and motion in real time, even in challenging and dynamic environments. You will own the full pipeline from algorithm development through embedded deployment, ensuring that perception systems run efficiently on resource-constrained robotic hardware. You will also leverage modern machine learning approaches to push the boundaries of classical perception methods, combining learned representations with geometric techniques to achieve robust, real-time performance. This is a deeply hands-on role. You will work directly with sensors, hardware, and real-world data, while prototyping, testing, and iterating in physical environments. The ideal candidate has strong foundations in VIO and sensor fusion, practical experience optimizing algorithms for embedded platforms, and familiarity with how modern deep learning is transforming perception. Key job responsibilities - Design and implement Visual Inertial Odometry algorithms for robust real-time state estimation on robotic platforms like Sprout - Develop multi-sensor fusion pipelines integrating cameras, IMUs, and other sensing modalities for accurate pose tracking - Optimize perception and tracking algorithms for deployment on embedded hardware (e.g., ARM, GPU-accelerated edge devices) under strict latency and power constraints - Apply modern ML-based perception techniques (learned features, depth estimation, neural odometry) to complement and improve classical geometric approaches - Build and maintain calibration, evaluation, --- TITLE: Project Coordinator, Enterprise Projects Office EMPLOYER: Voyager Technologies INC LOCATION: Remote - United States (remote) SALARY: $65K-$85K POSTED: 2026-06-12 APPLY_URL: https://job-boards.greenhouse.io/voyagertechnologiesinc/jobs/4278858009 EXCERPT: Project Coordinator, Enterprise Projects Office Remote - United States Voyager is an innovative space, defense, and national security technology company committed to advancing and delivering transformative, mission-critical solutions. We tackle the most complex challenges to unlock new frontiers for human progress, fortify national security, and protect critical assets to lead in the race for technological and operational superiority from ground to space. Forge the Future: Join Voyager Technologies The future belongs to those who build it. At Voyager Technologies, we're building technologies that protect lives, expand frontiers and prepare us for what's next. And we're doing that with people who are wired to solve, build, adapt and lead. These roles are not for the faint of heart. You'll help lay the foundation for humanity's future. Join a culture where innovation thrives, curiosity is rewarded, and impact is real. We're a company of doers, thinkers and builders, united by purpose and grounded in reality. If you want to put your skills to work where the stakes are real and the mission is bigger than any one person, forge the future with Voyager. The Project Coordinator, Enterprise Projects Office (EPO) , is responsible for executing the communications, change management content, and operational infrastructure that keep Voyager's highest-priority enterprise programs moving. This is a hands-on execution role that turns strategy into action - publishing program updates, building training materials, supporting enterprise planning cycles, coordinating project intake logistics, and maintaining the EPO's project portfolio infrastructure so senior program leaders can stay focused on strategy --- TITLE: Senior Mechanical Engineer EMPLOYER: Cerebras Systems LOCATION: Sunnyvale, CA (unspecified) SALARY: $190K-$230K POSTED: 2026-01-08 K401_MATCH: yes (not source-backed) APPLY_URL: https://job-boards.greenhouse.io/cerebrassystems/jobs/7522557003 EXCERPT: Senior Mechanical Engineer Sunnyvale, CA Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. Our novel wafer-scale architecture provides the AI compute power of dozens of GPUs on a single chip, with the programming simplicity of a single device. This approach allows Cerebras to deliver industry-leading training and inference speeds and empowers machine learning users to effortlessly run large-scale ML applications, without the hassle of managing hundreds of GPUs or TPUs. Cerebras' current customers include top model labs, global enterprises, and cutting-edge AI-native startups. OpenAI recently announced a multi-year partnership with Cerebras , to deploy 750 megawatts of scale, transforming key workloads with ultra high-speed inference. Thanks to the groundbreaking wafer-scale architecture, Cerebras Inference offers the fastest Generative AI inference solution in the world, over 10 times faster than GPU-based hyperscale cloud inference services. This order of magnitude increase in speed is transforming the user experience of AI applications, unlocking real-time iteration and increasing intelligence via additional agentic computation. About The Role As a Senior Mechanical Engineer at Cerebras, you will lead the design of mechanical systems for our next-generation wafer-scale engine. Your responsibilities will include ensuring compliance with specifications, validating manufacturability, and delivering a high-quality product in a fast-paced environment-tackling some of the most challenging problems in the rapidly evolving AI space. In this role, you will develop mechanical infrastructure for Cerebras' custom hardware system. - Rapidly iterate on designs and analysis to inform high level systems trades and steer overall product direction. - Provide --- TITLE: (Senior) Backend Engineer (Golang) - Balance Management EMPLOYER: SumUp LOCATION: Vilnius, Lithuania (unspecified) SALARY: EUR 4K-7K/mo POSTED: 2025-10-10 K401_MATCH: yes (not source-backed) LEARNING_BUDGET_OFFERED: yes (not source-backed) APPLY_URL: https://sumup.com/careers/positions/8207386002?gh_jid=8207386002 EXCERPT: (Senior) Backend Engineer (Golang) - Balance Management Vilnius, Lithuania We believe in the everyday hero. Those who have the courage to follow their passion and who have the strength and determination to realise their dreams. Small business owners are at the heart of all we do, so we're creating powerful, easy-to-use financial solutions to help them run their businesses. With a founder's mentality and a 'team-first' attitude, our diverse teams across Europe, South America and the United States work together to ensure that small business owners can be successful doing what they love. At SumUp, the Global Bank team builds the core infrastructure and services that give merchants a digital bank account, helping small businesses manage their money easily and reliably. The Balance Management Squad sits at the core of the Global Bank Platform. Right now, our team is completing a major milestone: consolidating and modernizing our balance management system in Europe. The next big step is creating a global transaction history service; a single, shared platform used by all regions to give merchants consistent and transparent views of their financial activity. As a (Senior) Backend Engineer , you'll help modernize existing systems while balancing innovation with stability, collaborate with teams across regions, and build reliable services that manage diverse markets with different regulations. Our tech stack includes Go, AWS, Kafka, PostgreSQL, and Kubernetes, supported by a strong observability toolchain with Prometheus, Grafana, and Honeycomb. We also actively use AI‑assisted development tools such as Cursor, GitHub Copilot, and others. SumUp's --- TITLE: Senior Applied Scientist, Amazon AWS Agentic AI, AWS AI Fundamental Research EMPLOYER: Amazon LOCATION: Santa Clara, California, USA (unspecified) SALARY: $192K-$260K POSTED: 2026-06-03 PARENTAL_LEAVE_WEEKS: 6 (not source-backed) NON_BIRTH_PARENT_LEAVE_WEEKS: 6 (not source-backed) K401_MATCH: yes (not source-backed) FERTILITY_FAMILY_BUILDING_BENEFITS: yes (not source-backed) ADOPTION_ASSISTANCE_OFFERED: yes (not source-backed) MENTAL_HEALTH_SUPPORT: yes (not source-backed) CHILDCARE_SUBSIDY: yes (not source-backed) APPLY_URL: https://www.amazon.jobs/en/jobs/10438803/senior-applied-scientist-amazon-aws-agentic-ai-aws-ai-fundamental-research EXCERPT: Senior Applied Scientist, Amazon AWS Agentic AI, AWS AI Fundamental Research Santa Clara, California, USA Amazon is looking for a passionate, talented, and inventive Applied Scientist with a strong machine learning background to help build industry-leading technology in generative AI and foundational models. As part of our AI team in Amazon AWS, you will work alongside internationally recognized experts to develop novel algorithms and modeling techniques to advance the state-of-the-art in generative AI. Your work will directly impact millions of our customers in the form of products and services that make use of speech, vision and language technology. You will gain hands on experience with Amazon's heterogeneous speech, text, image and structured data sources, and large-scale computing resources to accelerate advances in machine learning and foundation models. More specifically, you will have the opportunity to impact millions of our customers by researching and building innovative solutions using Agentic AI. Agentic AI drives innovation at the forefront of artificial intelligence, enabling customers to transform their businesses through generative AI solutions. We build and deliver the foundational AI services that power the future of cloud computing, helping organizations harness the potential of AI to solve their most complex challenges. Join our dynamic team of AI/ML practitioners and applied scientists who work backwards from customer needs to create novel technologies. If you're passionate about shaping the future of AI while making a meaningful impact for customers worldwide, we want to hear from you. Key job responsibilities The Senior Applied Scientist will lead the --- TITLE: AI/ML Scientist, Planetary Science EMPLOYER: Relativity LOCATION: Long Beach, California (unspecified) SALARY: $115K-$158K POSTED: 2026-05-22 APPLY_URL: https://boards.greenhouse.io/relativity/jobs/8561541002?gh_jid=8561541002 EXCERPT: AI/ML Scientist, Planetary Science Long Beach, California At Relativity Space, we're building rockets to serve today's needs and tomorrow's breakthroughs. Our Terran R vehicle will deliver customer payloads to orbit, meeting the growing demand for launch capacity. But that's just the start. Achieving commercial success with Terran R will unlock new opportunities to advance science, exploration, and innovation, pioneering progress that reaches beyond the known. Joining Relativity means becoming part of something where autonomy, ownership, and impact exist at every level. Here, you're not just executing tasks; you're solving problems that haven't been solved before, helping develop a rocket, a factory, and a business from the ground up. Whether you're in propulsion, manufacturing, software, avionics, or a corporate function, you'll collaborate across teams, shape decisions, and see your work come to life in record time. Relativity is a place where creativity and technical rigor go hand in hand, and your voice will help define the stories we're writing together. Now is a unique moment in time where it's early enough to leave your mark on the product, the process, and the culture, but far enough along that Terran R is tangible and picking up momentum. The most meaningful work of your career is waiting. Join us. About the Team: The Interplanetary Sciences Program was established to expand access to scientific exploration across our Solar System, with the mission to push the boundaries of how planetary science is done, and make planetary research faster, more affordable, and more capable than ever --- TITLE: Senior Software Engineer (Monitor team) EMPLOYER: Sysdig LOCATION: Flexible - Serbia (unspecified) SALARY: Not disclosed POSTED: 2026-02-26 K401_MATCH: yes (not source-backed) MENTAL_HEALTH_SUPPORT: yes (not source-backed) APPLY_URL: https://jobs.lever.co/sysdig/6e819a77-9585-4508-80c3-d0f130263ec4 EXCERPT: Senior Software Engineer (Monitor team) Flexible - Serbia What you will do - Reporting to the Engineering Manager, you will be part of a team of expert developers building the core distributed systems that power the Sysdig platform at scale. - Collaborate with colleagues in product and design to define feature specifications and timelines. - Be part of an amazing team where the primary goal for everybody is to work together to build the best product we can. We enjoy what we do and we always aim to get better. - Participate in an on-call rotation to address any urgent issues. What you will bring with you - Advanced coding skills, with 5+ years of experience in Java or Golang Backend, such as RESTful API design and microservice-based architectures. - Hands-on experience building horizontally scalable components and event-driven architectures using messaging and streaming platforms like Redis, Kafka, NATS, Apache Flink, or similar. - Proficiency with testing methodologies, such as unit testing and integration testing. - We appreciate experience observability topics (metrics, logging, tracing) and tools (e.g., Prometheus, Grafana, Jaeger, Elastic Stack, or similar). - Experience operating in Kubernetes and Cloud providers (AWS, GCP, Azure) is desirable. What we look for - You are willing to participate in professional development activities to stay current with industry trends and improve. - You are fluent in English with the proficiency to collaborate with a global team while expressing your thoughts and ideas, contributing to both technical and product discussions and seeing them come --- TITLE: Operations Engineer, Fleet Reliability EMPLOYER: Fal LOCATION: Remote (remote) SALARY: Not disclosed POSTED: 2026-05-14 APPLY_URL: https://job-boards.greenhouse.io/fal/jobs/4248332009 EXCERPT: Operations Engineer, Fleet Reliability Remote fal is the generative media ecosystem powering the next generation of AI products. We build the infrastructure, tools, and model access that teams need to move from idea to production, and do it at scale without compromise. For developers and enterprises, fal is the foundation that makes generative media not just possible, but practical: a unified platform where high-performance inference, orchestration, and observability come together to unlock new categories of AI-native products. As generative media reshapes industries across a market projected to grow by hundreds of billions over the next decade, fal is becoming the ecosystem that ambitious teams build on. About the role As we bring up owned clusters alongside our cloud capacity, we're hiring Operations Engineers to keep the fleet alive. This is a hands-on role. You're first in line when nodes go bad, GPUs throw ECC errors, IB links flap, or a rack stops responding. You'll provision new nodes, validate them, ship them to production, and troubleshoot whatever entropy throws at them. You'll be on-call. You'll be in the weeds. You're a fit if you've: - Administered Linux Systems in the critical path before - Troubleshooted GPU node issues: NVLink, NCCL, IB, driver and firmware bugs - Has experience in observability systems like Grafana and Prometheus - Scripted your way out of repetitive work (bash, python, go, whatever) Who you are: - Curious. You don't accept "it's flaky" as a root cause - Comfortable with ambiguity. The runbook doesn't exist yet for --- TITLE: Site Reliability Engineer, Metal EMPLOYER: TensTorrent LOCATION: Toronto, Ontario, Canada (unspecified) SALARY: Not disclosed POSTED: 2026-04-10 APPLY_URL: https://job-boards.greenhouse.io/tenstorrent/jobs/5105302007 EXCERPT: Site Reliability Engineer, Metal Toronto, Ontario, Canada Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify innovations in software models, compilers, platforms, networking, and semiconductors. Our diverse team of technologists have developed a high performance RISC-V CPU from scratch, and share a passion for AI and a deep desire to build the best AI platform possible. We value collaboration, curiosity, and a commitment to solving hard problems. We are growing our team and looking for contributors of all seniorities. Tenstorrent is building large-scale AI systems across internal clusters and customer deployments. This role sits at the intersection of site reliability, infrastructure operations, and customer engineering, ensuring our systems are reliable, observable, and production-ready. This role is hybrid, based out of Toronto, ON; Austin, TX; or Santa Clara, CA. We welcome candidates at various experience levels for this role. During the interview process, candidates will be assessed for the appropriate level, and offers will align with that level, which may differ from the one in this posting. Who You Are - Experienced in site reliability, infrastructure, or systems engineering in distributed environments. - Strong Linux systems knowledge with the ability to troubleshoot complex multi-layer issues. - Proficient with observability tools such as Prometheus, Grafana, and alerting systems. - Comfortable with scripting and automation using Python, Go, or similar languages. - Solid understanding of networking fundamentals and how systems behave at scale. --- TITLE: Senior Avionics Hardware Engineer, Navigation EMPLOYER: Relativity LOCATION: Long Beach, California (unspecified) SALARY: $148K-$204K POSTED: 2026-02-26 APPLY_URL: https://boards.greenhouse.io/relativity/jobs/8387902002?gh_jid=8387902002 EXCERPT: Senior Avionics Hardware Engineer, Navigation Long Beach, California At Relativity Space, we're building rockets to serve today's needs and tomorrow's breakthroughs. Our Terran R vehicle will deliver customer payloads to orbit, meeting the growing demand for launch capacity. But that's just the start. Achieving commercial success with Terran R will unlock new opportunities to advance science, exploration, and innovation, pioneering progress that reaches beyond the known. Joining Relativity means becoming part of something where autonomy, ownership, and impact exist at every level. Here, you're not just executing tasks; you're solving problems that haven't been solved before, helping develop a rocket, a factory, and a business from the ground up. Whether you're in propulsion, manufacturing, software, avionics, or a corporate function, you'll collaborate across teams, shape decisions, and see your work come to life in record time. Relativity is a place where creativity and technical rigor go hand in hand, and your voice will help define the stories we're writing together. Now is a unique moment in time where it's early enough to leave your mark on the product, the process, and the culture, but far enough along that Terran R is tangible and picking up momentum. The most meaningful work of your career is waiting. Join us. About the Team: The Avionics team is responsible for the full lifecycle of Terran R's nervous system, designing, building, testing, installing, and operating the hardware that connects and controls every major electrical system on the vehicle and ground. The team leverages close partnership --- TITLE: DevOps Engineer | AWS Cloud & Automation (Español + Portugués) EMPLOYER: SkydropX LOCATION: México | Remote (remote) SALARY: Not disclosed POSTED: 2026-06-10 APPLY_URL: https://jobs.ashbyhq.com/skydropx/043671b8-7b52-49f2-bfa4-fbfec7f9fa62 EXCERPT: DevOps Engineer | AWS Cloud & Automation (Español + Portugués) México | Remote RESPONSABILIDADES: - Mantener y desarrollar la infraestructura en la nube de AWS (ECS, EC2, RDS, S3, VPC, etc.). - Garantizar las prácticas de observabilidad (monitoreo, alertas y logs) utilizando herramientas como Prometheus, Grafana y CloudWatch. - Mantener y desarrollar los pipelines de CI/CD en Bitbucket para aplicaciones en contenedores .NET, Python y Docker. - Apoyar y mejorar los procesos de automatización de pruebas (integración, regresión y desempeño). - Implementar estrategias de seguridad en la nube (IAM, roles, compliance, hardening). - Apoyar al equipo de ingeniería en el diagnóstico y resolución de incidentes críticos. - Proponer mejoras de arquitectura, automatización y gobernanza en la nube. - Trabajar en colaboración con los equipos de ingeniería y calidad para garantizar entregas rápidas y seguras. REQUISITOS: - Indispensable: Portugués C1 o superior. - Experiencia en entornos con ASP.NET http://ASP.NET. - Más de 4 años de experiencia trabajando con AWS, utilizando servicios como EKS, ECS + Fargate, SQL Server + RDS, SQS, EC2, VPN, VPC, Transit Gateway, NAT Gateway e IAM. - Experiencia en CI/CD con Bitbucket Pipelines e integración con contenedores Docker. - Conocimiento en Infraestructura como Código (Terraform o AWS CDK). - Experiencia con Linux y automatización con Shell/Bash. - Conocimiento en contenedores y orquestación (Docker y ECS/Fargate). - Experiencia con monitoreo y observabilidad (Prometheus, Grafana, CloudWatch). - Experiencia con herramientas de gestión y calidad: Jira (gestión ágil) y SonarQube (calidad de código). - Conocimientos básicos de bases de datos --- TITLE: Senior AI/ML Scientist, Planetary Science EMPLOYER: Relativity LOCATION: Long Beach, California (unspecified) SALARY: $154K-$211K POSTED: 2026-05-22 APPLY_URL: https://boards.greenhouse.io/relativity/jobs/8561397002?gh_jid=8561397002 EXCERPT: Senior AI/ML Scientist, Planetary Science Long Beach, California At Relativity Space, we're building rockets to serve today's needs and tomorrow's breakthroughs. Our Terran R vehicle will deliver customer payloads to orbit, meeting the growing demand for launch capacity. But that's just the start. Achieving commercial success with Terran R will unlock new opportunities to advance science, exploration, and innovation, pioneering progress that reaches beyond the known. Joining Relativity means becoming part of something where autonomy, ownership, and impact exist at every level. Here, you're not just executing tasks; you're solving problems that haven't been solved before, helping develop a rocket, a factory, and a business from the ground up. Whether you're in propulsion, manufacturing, software, avionics, or a corporate function, you'll collaborate across teams, shape decisions, and see your work come to life in record time. Relativity is a place where creativity and technical rigor go hand in hand, and your voice will help define the stories we're writing together. Now is a unique moment in time where it's early enough to leave your mark on the product, the process, and the culture, but far enough along that Terran R is tangible and picking up momentum. The most meaningful work of your career is waiting. Join us. About the Team: The Interplanetary Sciences Program was established to expand access to scientific exploration across our Solar System, with the mission to push the boundaries of how planetary science is done, and make planetary research faster, more affordable, and more capable than --- TITLE: Senior Applied Scientist - SLAM and Perception, Amazon Robotics EMPLOYER: Amazon LOCATION: North Reading, Massachusetts, USA (unspecified) SALARY: $167K-$226K POSTED: 2026-05-21 PARENTAL_LEAVE_WEEKS: 6 (not source-backed) NON_BIRTH_PARENT_LEAVE_WEEKS: 6 (not source-backed) K401_MATCH: yes (not source-backed) FERTILITY_FAMILY_BUILDING_BENEFITS: yes (not source-backed) ADOPTION_ASSISTANCE_OFFERED: yes (not source-backed) MENTAL_HEALTH_SUPPORT: yes (not source-backed) CHILDCARE_SUBSIDY: yes (not source-backed) APPLY_URL: https://www.amazon.jobs/en/jobs/10426858/senior-applied-scientist-slam-and-perception-amazon-robotics EXCERPT: Senior Applied Scientist - SLAM and Perception, Amazon Robotics North Reading, Massachusetts, USA Application deadline: Jun 15, 2026 The Amazon Robotics Autonomous Mobility software team develops the autonomy software that powers Proteus, Amazon's first fully autonomous warehouse robot. We are looking for a Senior Applied Scientist to join our perception and localization team. You will be part of an exceptional group of engineers building the platform and architecture for our onboard mapping, perception, motion planning and control software. In addition, you will collaborate with engineers and scientists across the team to develop and enhance our real-time multi-robot simulation capabilities. Key job responsibilities Specific areas of focus may include (but are not limited too): - Researching, designing, and implementing scientific approaches for sensor calibration, including cameras and LIDARs. - Researching, designing, and implementing scientific approaches for SLAM (simultaneous localization and mapping). - Supporting and improving the large-scale deployment of a SLAM system shared across an autonomous robot fleet. - Evaluating and integrating new sensing modalities into an existing autonomy software stack. - Optimizing runtime performance of autonomy algorithms by exploiting underlying hardware acceleration capabilities. - Building frameworks for large-scale replay and analysis of events in pre-recorded sensor data. - Deliver high quality production level code (C++ or Python) and support systems in production. - Collaborate with other functional teams in a robotics organization. - Collaborate with hardware engineering team members on developing systems from prototyping to production level. - Work with stakeholders across hardware and operations teams to iterate on system --- TITLE: Senior Applied Scientist, Fauna EMPLOYER: Amazon LOCATION: New York, New York, USA (unspecified) SALARY: $184K-$249K POSTED: 2026-05-18 PARENTAL_LEAVE_WEEKS: 6 (not source-backed) NON_BIRTH_PARENT_LEAVE_WEEKS: 6 (not source-backed) K401_MATCH: yes (not source-backed) FERTILITY_FAMILY_BUILDING_BENEFITS: yes (not source-backed) ADOPTION_ASSISTANCE_OFFERED: yes (not source-backed) MENTAL_HEALTH_SUPPORT: yes (not source-backed) CHILDCARE_SUBSIDY: yes (not source-backed) APPLY_URL: https://www.amazon.jobs/en/jobs/10423168/senior-applied-scientist-fauna EXCERPT: Senior Applied Scientist, Fauna New York, New York, USA We are seeking a Sr. Applied Scientist to develop cutting-edge machine learning algorithms for motor control systems in robots. In this role, you will focus on creating and optimizing intelligent motor control strategies to enable robots to perform complex, whole-body tasks. Your contributions will be essential in advancing robotics by enabling fluid, reliable, and safe interactions between robots and their environments. Key job responsibilities - Develop controllers that leverage reinforcement learning, imitation learning, or other advanced AI techniques to achieve natural, robust, and adaptive motor behaviors - Collaborate with multi-disciplinary teams to integrate motor control systems with robotic hardware, ensuring alignment with real-world constraints such as actuator dynamics and energy efficiency - Use simulation and real-world testing to refine and validate control algorithms - Stay updated on advancements in robotics, AI, and control systems to apply advanced techniques to robotic motion challenges - Lead technical projects from conception through production deployment - Mentor junior scientists and engineers - Bridge research initiatives with practical engineering implementation About the team Fauna Robotics, an Amazon company, is building capable, safe, and genuinely delightful robots for everyday life. Our goal is simple: make robots people actually want to live and interact with in everyday human spaces. We believe that future won't arrive until building for robotics becomes far more accessible. Today, too much effort is spent reinventing the fundamentals. We're changing that by developing tightly integrated hardware and software systems that make it faster, safer, --- TITLE: Senior Software Engineer EMPLOYER: Nubank LOCATION: USA, Miami (unspecified) SALARY: Not disclosed POSTED: 2025-07-18 APPLY_URL: https://job-boards.greenhouse.io/nubank/jobs/7081018 EXCERPT: Senior Software Engineer USA, Miami About Us Nu is one of the largest digital financial platforms in the world, with more than 127 million customers across Brazil, Mexico, and Colombia. Guided by our mission to fight complexity and empower people, we are redefining financial services in Latin America and this is still just the beginning of the purple future we're building. Listed on the New York Stock Exchange (NYSE: NU), we combine proprietary technology, data intelligence, and an efficient operating model to deliver financial products that are simple, accessible, and human. Our impact has been recognized by global rankings such as Time 100 Companies, Fast Company's Most Innovative Companies, and Forbes World's Best Bank. Visit our institutional page https://international.nubank.com.br/careers/ Engineering at Nubank We strive for state-of-the-art software development practices, that currently includes a variety of technologies. While we value candidates that are familiar with them, we are also confident that software engineers who are interested in joining Nubank will be able to learn from our team. - Horizontally scalable microservices written mostly in Clojure, using Finagle and leveraging upon functional programming techniques and hexagonal architecture - High throughput jobs and inter-service communication using Kafka - Continuous Integration and Deployment into AWS - Storing data in Datomic and DynamoDB - Monitoring and observability with Prometheus - Running as much as possible in Kubernetes We are a process-light organization that values human interactions. We value working in small, independent teams that feel like small startups within the company, and eschew coupling and --- TITLE: Senior Hardware Engineer, Digital EMPLOYER: Relativity LOCATION: Long Beach, California (unspecified) SALARY: $148K-$204K POSTED: 2026-05-04 APPLY_URL: https://boards.greenhouse.io/relativity/jobs/8533748002?gh_jid=8533748002 EXCERPT: Senior Hardware Engineer, Digital Long Beach, California At Relativity Space, we're building rockets to serve today's needs and tomorrow's breakthroughs. Our Terran R vehicle will deliver customer payloads to orbit, meeting the growing demand for launch capacity. But that's just the start. Achieving commercial success with Terran R will unlock new opportunities to advance science, exploration, and innovation, pioneering progress that reaches beyond the known. Joining Relativity means becoming part of something where autonomy, ownership, and impact exist at every level. Here, you're not just executing tasks; you're solving problems that haven't been solved before, helping develop a rocket, a factory, and a business from the ground up. Whether you're in propulsion, manufacturing, software, avionics, or a corporate function, you'll collaborate across teams, shape decisions, and see your work come to life in record time. Relativity is a place where creativity and technical rigor go hand in hand, and your voice will help define the stories we're writing together. Now is a unique moment in time where it's early enough to leave your mark on the product, the process, and the culture, but far enough along that Terran R is tangible and picking up momentum. The most meaningful work of your career is waiting. Join us. About the Team: The Avionics team is responsible for the full lifecycle of Terran R's nervous system, designing, building, testing, installing, and operating the hardware that connects and controls every major electrical system on the vehicle and ground. The team leverages close partnership between --- TITLE: Software Engineer EMPLOYER: Stripe LOCATION: Seattle (unspecified) SALARY: $157K-$235K POSTED: 2026-06-10 NON_BIRTH_PARENT_LEAVE_WEEKS: 16 (not source-backed) K401_MATCH: yes (not source-backed) FERTILITY_FAMILY_BUILDING_BENEFITS: yes (not source-backed) CHILDCARE_SUBSIDY: yes (not source-backed) APPLY_URL: https://stripe.com/jobs/search?gh_jid=7991636 EXCERPT: Software Engineer Seattle Who we are About Stripe Stripe, LLC. is a financial infrastructure platform for businesses. Millions of companies - from the world's largest enterprises to the most ambitious startups - use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone's reach while doing the most important work of your career. What you'll do Responsibilities Design, build, and maintain APIs, services, and systems across Stripe's engineering teams using Java, Ruby, Scala, and Go Build software infrastructure, including developing, testing, and deploying it Design and develop software systems, using scientific analysis and mathematical models to predict and measure outcome and consequences of design Engineer payments integration with various financial partners software systems Design APIs and underlying data models to support complex financial abstractions and multi-party integrations, enabling flexible billing and settlement configurations across international markets Develop and direct software system testing and validation procedures, programming, and documentation Debug production issues across services and multiple levels of the stack Analyze user needs and software requirements to determine the feasibility of design within time and cost constraints Work with engineers across the company to build new features at large-scale Build new systems to securely store sensitive data; Improve engineering standards, tooling, and processes Integrate observability tools and alerting mechanisms (Datadog, Prometheus) into high-traffic production systems, defining --- TITLE: Software Engineer 2 EMPLOYER: Abnormal Security LOCATION: Hybrid - Bangalore, India (hybrid) SALARY: Not disclosed POSTED: 2026-05-26 K401_MATCH: yes (not source-backed) APPLY_URL: https://abnormal.ai/careers/jobs/7741028003?gh_jid=7741028003 EXCERPT: Software Engineer 2 Hybrid - Bangalore, India About the Role Abnormal AI is looking for an experienced and driven Platform & Infra software engineer to join the PI team. Join us and help build the platforms that power Abnormal's growth - Observability Platform - Own and evolve the monitoring, metrics, and alerting infrastructure that every engineering team at Abnormal depends on. You'll work across the Prometheus, Chronosphere, and Grafana stack to ensure engineers can see what their systems are doing in real time - building dashboards, managing metric pipelines at scale, operating the PagerDuty alerting pipeline, and driving cost-efficient observability across all production environments (US, EU, and GovCloud). Your Impact - Own the observability stack (Prometheus, Chronosphere, Grafana, PagerDuty) that every team relies on to detect, diagnose, and resolve production issues - when you make it better, every engineer at Abnormal gets faster. - Design platforms and developer tooling that remove friction - reducing deployment times, simplifying pipeline authoring, and letting product teams focus on building rather than firefighting. - Drive SLAs and SLOs for critical shared infrastructure ensuring the systems behind our products are resilient and cost-efficient. - Your architectural decisions on alerting pipelines and cross-environment deployments will define what products we can build and how quickly we deliver them to customers. What you will do - Work with the Tech Lead, Engineering Manager, and Product Manager to design, develop, and deliver key platform features - from technical design docs through production rollout - Own features end-to-end: scoping, implementation, --- TITLE: Site Reliability Engineer EMPLOYER: GoDaddy LOCATION: Bulgaria (hybrid) SALARY: Not disclosed POSTED: 2026-02-19 APPLY_URL: https://careers.godaddy/jobs?gh_jid=7623136003 EXCERPT: Site Reliability Engineer Bulgaria Location Details: At GoDaddy the future of work looks different for each team. Some teams work in the office full-time, others have a hybrid arrangement (they work remotely some days and in the office some days) and some work entirely remotely. This is a remote position, so you'll be working remotely from your home. You may occasionally visit a GoDaddy office to meet with your team for events or meetings. Join Our Team GoDaddy is looking for a Site Reliability Engineer to join our Monitoring and Observability team. In this role, you'll help ensure the reliability, performance, and availability of infrastructure that serves millions of customers worldwide. You'll work at the intersection of development and operations to build and maintain observability solutions that enable proactive monitoring and rapid incident response across cloud and on-prem systems. What you'll get to do... Design, deploy, and maintain observability and monitoring platforms using Python , including metrics, logging, tracing, and visualization tools (e.g., Prometheus, Grafana). Automate operational processes and build self-service tooling to improve reliability and reduce manual effort. Respond to production incidents, participate in on-call rotations, and collaborate with teams to resolve performance, availability, and security issues. Support CI/CD pipelines and configuration management for monitoring and observability infrastructure. Your experience should include... 3+ years of professional experience designing, building, and operating large-scale infrastructure as a Site Reliability Engineer, DevOps, or similar role. 3+ years of deep expertise in Linux/Unix systems, including performance tuning, kernel-level troubleshooting, and systems optimization. 3+ --- TITLE: Director of Software Engineering, Data & AI EMPLOYER: Relativity LOCATION: Long Beach, California (unspecified) SALARY: $245K-$336K POSTED: 2026-02-12 APPLY_URL: https://boards.greenhouse.io/relativity/jobs/8419838002?gh_jid=8419838002 EXCERPT: Director of Software Engineering, Data & AI Long Beach, California At Relativity Space, we're building rockets to serve today's needs and tomorrow's breakthroughs. Our Terran R vehicle will deliver customer payloads to orbit, meeting the growing demand for launch capacity. But that's just the start. Achieving commercial success with Terran R will unlock new opportunities to advance science, exploration, and innovation, pioneering progress that reaches beyond the known. Joining Relativity means becoming part of something where autonomy, ownership, and impact exist at every level. Here, you're not just executing tasks; you're solving problems that haven't been solved before, helping develop a rocket, a factory, and a business from the ground up. Whether you're in propulsion, manufacturing, software, avionics, or a corporate function, you'll collaborate across teams, shape decisions, and see your work come to life in record time. Relativity is a place where creativity and technical rigor go hand in hand, and your voice will help define the stories we're writing together. Now is a unique moment in time where it's early enough to leave your mark on the product, the process, and the culture, but far enough along that Terran R is tangible and picking up momentum. The most meaningful work of your career is waiting. Join us. About the Team: Relativity Space is on a mission to better connect humanity to space and the universe beyond our planet. With decades of experience scaling world-class technology organizations like Google, CEO Eric Schmidt is guiding Relativity into its next phase: --- TITLE: Senior Site Reliability Engineer I EMPLOYER: Axon Enterprise LOCATION: Boston, Massachusetts, United States (unspecified) SALARY: $134K-$215K POSTED: 2026-03-13 PARENTAL_LEAVE_WEEKS: 10 (not source-backed) NON_BIRTH_PARENT_LEAVE_WEEKS: 10 (not source-backed) K401_MATCH: yes (not source-backed) CHILDCARE_SUBSIDY: yes (not source-backed) APPLY_URL: https://job-boards.greenhouse.io/axon/jobs/7660088003 EXCERPT: Senior Site Reliability Engineer I Boston, Massachusetts, United States Join Axon and be a Force for Good. At Axon, we're on a mission to Protect Life. We're explorers, pursuing society's most critical safety and justice issues with our ecosystem of devices and cloud software. Like our products, we work better together. We connect with candor and care, seeking out diverse perspectives from our customers, communities and each other. Life at Axon is fast-paced, challenging and meaningful. Here, you'll take ownership and drive real change. Constantly grow as you work hard for a mission that matters at a company where you matter. Your Impact Are you an engineer who gets excited about the challenge of making complex distributed systems observable - not just instrumenting them, but designing the infrastructure that makes traces, metrics, and logs useful at scale? In this role, you will help build and evolve Axon's next-generation observability platform, enabling the entire engineering organization to understand and operate their services with confidence. You'll work across the full observability stack: from distributed tracing adoption (OpenTelemetry, Jaeger) to log infrastructure (Loki, Alloy) to metrics (Cortex, Prometheus, Grafana). You'll partner directly with Axon's engineering teams to drive adoption of modern observability practices and build the tooling that makes our platform self-service for the teams that depend on it. You will be part of the Observability team within Axon's Site Reliability organization - a focused team responsible for Axon's metrics, logging, tracing, and alerting infrastructure across dozens of environments globally. The ideal candidate --- TITLE: Senior End to End Performance Engineer EMPLOYER: Apple LOCATION: Seattle, United States of America (unspecified) SALARY: $172K-$302K POSTED: 2026-03-13 K401_MATCH: yes (not source-backed) APPLY_URL: https://jobs.apple.com/en-us/details/200651561/senior-end-to-end-performance-engineer?team=SFTWR EXCERPT: Senior End to End Performance Engineer Seattle, United States of America Join the Apple Service Engineering (ASE) team and drive innovation that matters! The ASE team builds and provides systems and infrastructure that fuel Apple's services. As part of this team, you will be responsible for building and integrating technologies that enhance people's lives. We are also tasked with enabling Apple Intelligence and Private Cloud Compute in the cloud. We are seeking a System Performance Engineer with a proven track record of developing technology that has made a significant impact. Collaborate cross-functionally with architecture, platform design, SOC architects, and software teams to deliver exceptional end-to-end performance that delights our customers. In this role, you'll be responsible for the performance of one or more products. Specifically, you'll be tasked with ensuring the end-to-end performance of Private Cloud Compute running Apple Intelligence in the cloud. This position demands technical expertise and depth to work on various elements of the stack, from SOC to the platform, software stack, and communication with the external world. You'll collaborate with architecture teams to define the product and set its performance targets, and with the SRE team to assess end-to-end performance and develop automation and infrastructure tools to evaluate the performance of our Apple Intelligence products. Strong problem-solving skills and effective communication are crucial for success in this role. Minimum Qualifications: 5+ years of experience as a Performance Engineer In-depth professional experience with cloud operations Experience with monitoring and observability tools such as Splunk, Grafana, and Prometheus --- TITLE: Senior Hardware Engineer, Power Electronics EMPLOYER: Relativity LOCATION: Long Beach, California (unspecified) SALARY: $148K-$204K POSTED: 2026-05-01 APPLY_URL: https://boards.greenhouse.io/relativity/jobs/8533802002?gh_jid=8533802002 EXCERPT: Senior Hardware Engineer, Power Electronics Long Beach, California At Relativity Space, we're building rockets to serve today's needs and tomorrow's breakthroughs. Our Terran R vehicle will deliver customer payloads to orbit, meeting the growing demand for launch capacity. But that's just the start. Achieving commercial success with Terran R will unlock new opportunities to advance science, exploration, and innovation, pioneering progress that reaches beyond the known. Joining Relativity means becoming part of something where autonomy, ownership, and impact exist at every level. Here, you're not just executing tasks; you're solving problems that haven't been solved before, helping develop a rocket, a factory, and a business from the ground up. Whether you're in propulsion, manufacturing, software, avionics, or a corporate function, you'll collaborate across teams, shape decisions, and see your work come to life in record time. Relativity is a place where creativity and technical rigor go hand in hand, and your voice will help define the stories we're writing together. Now is a unique moment in time where it's early enough to leave your mark on the product, the process, and the culture, but far enough along that Terran R is tangible and picking up momentum. The most meaningful work of your career is waiting. Join us. About the Team: The Avionics team is responsible for the full lifecycle of Terran R's nervous system, designing, building, testing, installing, and operating the hardware that connects and controls every major electrical system on the vehicle and ground. The team leverages close partnership --- TITLE: Advanced Technology: R&D Engineer - AI/ML, HPC EMPLOYER: Cerebras Systems LOCATION: Sunnyvale, CA; Toronto, Ontario, Canada; Vancouver, British Columbia, Canada (unspecified) SALARY: Not disclosed POSTED: 2026-04-06 K401_MATCH: yes (not source-backed) APPLY_URL: https://job-boards.greenhouse.io/cerebrassystems/jobs/7691343003 EXCERPT: Advanced Technology: R&D Engineer - AI/ML, HPC Sunnyvale, CA; Toronto, Ontario, Canada; Vancouver, British Columbia, Canada Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. Our novel wafer-scale architecture provides the AI compute power of dozens of GPUs on a single chip, with the programming simplicity of a single device. This approach allows Cerebras to deliver industry-leading training and inference speeds and empowers machine learning users to effortlessly run large-scale ML applications, without the hassle of managing hundreds of GPUs or TPUs. Cerebras' current customers include top model labs, global enterprises, and cutting-edge AI-native startups. OpenAI recently announced a multi-year partnership with Cerebras , to deploy 750 megawatts of scale, transforming key workloads with ultra high-speed inference. Thanks to the groundbreaking wafer-scale architecture, Cerebras Inference offers the fastest Generative AI inference solution in the world, over 10 times faster than GPU-based hyperscale cloud inference services. This order of magnitude increase in speed is transforming the user experience of AI applications, unlocking real-time iteration and increasing intelligence via additional agentic computation. About The Team Cerebras builds wafer-scale AI processors-single chips delivering tens of PB/s of memory bandwidth and a dataflow architecture that accelerates at a granularity no multi-device system can match. The Advanced Technology Group (ATG) is Cerebras ' pathfinding organization. We work ahead of product to explore new architectures, demonstrate breakthrough performance on scientific and AI workloads, and shape the technical roadmap for future Cerebras hardware and software. Our work regularly appears at top-tier venues (Supercomputing, --- TITLE: Trading Engineer - Strategy EMPLOYER: IMC Trading LOCATION: Chicago, United States (unspecified) SALARY: $175K-$225K POSTED: 2025-01-15 APPLY_URL: https://job-boards.eu.greenhouse.io/imc/jobs/4439286101 EXCERPT: Trading Engineer - Strategy Chicago, United States We're looking for a hands-on, entrepreneurial Trading Engineer to join our Strategy Development team. This role sits at the heart of IMC's trading operation, where technology, data, and market strategy meet. You'll own systems that power our trading strategies, build automations that make us faster and smarter, and keep our trading running smoothly under pressure. This is an ideal opportunity for someone with an SRE or operations engineering background who wants to get closer to the action, where milliseconds matter and creativity drives impact. What You'll Do: - Be on the front lines of trading: Monitor system health and performance in real time, respond to market-driven incidents, and ensure our systems stay up and fast during trading hours. Dive deep into distributed architectures to identify root causes and prevent recurrences. - Engineer away the pain: Build tools and automations that eliminate manual processes, scale our reliability, and free time for innovation. - Collaborate across disciplines: Partner with traders, quants, and developers to deliver high-impact improvements to our trading stack. - Continuously improve: Challenge existing processes, propose creative solutions, and drive reliability and performance through engineering excellence. What You Bring: - Bachelor's degree in Computer Science, Information Technology, or related field - 3+ years in site reliability, systems engineering, or technical operations, ideally supporting high-performance or real-time systems - A software-development mindset applied to automating operational problems (Python required) - Hands-on experience managing production Linux and Kubernetes environments - Familiarity with Prometheus, SQL/MongoDB, InfluxDB, --- TITLE: Member of Technical Staff (Software Engineer) EMPLOYER: Cerebras Systems LOCATION: Sunnyvale, CA (unspecified) SALARY: $170K-$175K POSTED: 2026-05-08 K401_MATCH: yes (not source-backed) APPLY_URL: https://job-boards.greenhouse.io/cerebrassystems/jobs/7728798003 EXCERPT: Member of Technical Staff (Software Engineer) Sunnyvale, CA Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. Our novel wafer-scale architecture provides the AI compute power of dozens of GPUs on a single chip, with the programming simplicity of a single device. This approach allows Cerebras to deliver industry-leading training and inference speeds and empowers machine learning users to effortlessly run large-scale ML applications, without the hassle of managing hundreds of GPUs or TPUs. Cerebras' current customers include top model labs, global enterprises, and cutting-edge AI-native startups. OpenAI recently announced a multi-year partnership with Cerebras , to deploy 750 megawatts of scale, transforming key workloads with ultra high-speed inference. Thanks to the groundbreaking wafer-scale architecture, Cerebras Inference offers the fastest Generative AI inference solution in the world, over 10 times faster than GPU-based hyperscale cloud inference services. This order of magnitude increase in speed is transforming the user experience of AI applications, unlocking real-time iteration and increasing intelligence via additional agentic computation. Cerebras Systems Inc. has multiple openings for Member of Technical Staff (Software Engineer) Title : Member of Technical Staff (Software Engineer) Job Duties - Implement infrastructure to support high-performance, low-latency inference service. - Deploy and configure Kubernetes services to ensure scalability and reliability of inference workloads. - Optimize resource allocation and auto-scaling policies to handle variable inference demand while minimizing operational costs. - Integrate inference services with containerized environments using Docker and Kubernetes for orchestration. - Ensure high availability and fault tolerance by implementing --- TITLE: Senior Platform Engineer EMPLOYER: MongoDB, Inc. LOCATION: Gurugram (unspecified) SALARY: Not disclosed POSTED: 2026-05-22 NON_BIRTH_PARENT_LEAVE_WEEKS: 20 (not source-backed) FERTILITY_FAMILY_BUILDING_BENEFITS: yes (not source-backed) MENTAL_HEALTH_SUPPORT: yes (not source-backed) CHILDCARE_SUBSIDY: yes (not source-backed) APPLY_URL: https://www.mongodb.com/careers/job/?gh_jid=7924987 EXCERPT: Senior Platform Engineer Gurugram The Infrastructure Engineering team is responsible for building and maintaining a self-service internal development platform that enables MongoDB engineering teams to reliably deploy and operate their own production services and products. We work with numerous engineering teams across the company to understand their infrastructure requirements and development workflows, develop broadly applicable self-service platform services and tooling, continuously monitor how platform services are being utilized, and look for ways to improve developer productivity through automation and education. We are big open source enthusiasts and use a number of open source tools in our stack (contributing upstream whenever possible). Some of the tools we use regularly include Go, AWS, Kubernetes, Crossplane, Terraform, Helm, Drone, Prometheus, and Grafana. However, technology is nothing without a stellar team of engineers that are focused on doing high quality work and working as a team to solve complex distributed computing and platform engineering problems. This is where you come in! We are looking to speak to candidates who are based in Gurugram for our hybrid working model. Our ideal candidate - Has built and operated large-scale distributed systems in cloud providers (AWS strongly preferred) - Has a strong backend programming background. Fluency in Go is strongly preferred; deep experience with another compiled or strongly-typed backend language is acceptable - Has experience working with AI coding agents and can demonstrate building high quality context to yield high quality outputs - Has experience designing and implementing medium-to-large software projects, including driving design reviews and mentoring --- TITLE: Backend Engineer - Transfers Gateway EMPLOYER: SumUp LOCATION: Vilnius, Lithuania (unspecified) SALARY: EUR 3K-5K/mo POSTED: 2026-01-23 K401_MATCH: yes (not source-backed) LEARNING_BUDGET_OFFERED: yes (not source-backed) APPLY_URL: https://sumup.com/careers/positions/8388972002?gh_jid=8388972002 EXCERPT: Backend Engineer - Transfers Gateway Vilnius, Lithuania At SumUp, the Global Bank tribe builds the core infrastructure and services that give merchants a digital business account, helping small businesses manage their banking needs easily and reliably. The Transfers Gateway squad connects SumUp's internal systems to external financial networks across the EU and the UK. We build and run the services that move money through payment schemes every day. The team owns end‑to‑end scheme integrations and operates production services that process real money at scale. You'll help modernise existing systems while balancing reliability, correctness, and change, making architectural decisions that directly affect uptime and merchant experience across multiple markets. Our tech stack includes Go, AWS, Kafka, PostgreSQL, and Kubernetes, supported by a strong observability toolchain with Prometheus, Grafana, and Honeycomb. We also actively use AI‑assisted development tools such as Cursor, GitHub Copilot, and others. SumUp's engineering culture is built around ownership. We follow a “build it, run it, improve it” model, with strong observability and a focus on learning. Engineers have clear scope and autonomy, with decisions documented, reviewed, and made collaboratively. The Vilnius team is a core part of the Global Bank domain. We own production systems, shape technical direction, and work closely with colleagues across SumUp's global offices on critical merchant banking capabilities. Join our Vilnius Team today ( Vilnius office ). What You'll Do - Build and maintain production money movement backend services, contributing to their delivery, observability, and ongoing maintenance. - Own and deliver well‑defined tasks and --- TITLE: Senior Embedded Engineer - Space Systems EMPLOYER: Hubble Network LOCATION: Seattle, Washington, United States (unspecified) SALARY: $126K-$249K POSTED: 2026-06-07 APPLY_URL: https://hubble.com/careers?gh_jid=5122786008 EXCERPT: Senior Embedded Engineer - Space Systems Seattle, Washington, United States Hubble Network was founded with the intention of delivering on the promise of what Internet-of-Things (IoT) was supposed to be. We're building a global Bluetooth® network dedicated to machine-to-machine connectivity. We differentiate ourselves as the first modem-less and gateway-less, direct-to-satellite network from off-the-shelf Bluetooth® Low Energy chips. Hubble is ideal for applications in logistics, AgTech, and maritime where economies of scale for volume consumer and enterprise asset tracking is a priority. Our goal is to be the first billion-endpoint-connected network in the world. Hubble is an early-stage, venture-backed startup supported by some of the best investors in the world. In their previous lives, the founding team has been successful in raising $100s of millions in venture funding, developing the Amazon Sidewalk network, launching billions of dollars of space assets, and leading their teams to successful exits, both through acquisition and IPO. We are now looking to bring on talented team members who are the best at what they do to help us make Hubble a reality for the world. We are seeking an experienced Senior Embedded Engineer to architect and implement software for our satellite systems. This is a hands-on role where you will work closely with our hardware, RF, and operations teams to develop high-reliability embedded software that powers our constellation. You will act as a technical leader on our engineering team, taking system-level requirements and translating them into robust, efficient embedded solutions that operate in the harsh space --- TITLE: Applied AI/ML Scientist EMPLOYER: Cerebras Systems LOCATION: UAE (unspecified) SALARY: Not disclosed POSTED: 2026-01-14 K401_MATCH: yes (not source-backed) APPLY_URL: https://job-boards.greenhouse.io/cerebrassystems/jobs/7588502003 EXCERPT: Applied AI/ML Scientist UAE Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. Our novel wafer-scale architecture provides the AI compute power of dozens of GPUs on a single chip, with the programming simplicity of a single device. This approach allows Cerebras to deliver industry-leading training and inference speeds and empowers machine learning users to effortlessly run large-scale ML applications, without the hassle of managing hundreds of GPUs or TPUs. Cerebras' current customers include top model labs, global enterprises, and cutting-edge AI-native startups. OpenAI recently announced a multi-year partnership with Cerebras , to deploy 750 megawatts of scale, transforming key workloads with ultra high-speed inference. Thanks to the groundbreaking wafer-scale architecture, Cerebras Inference offers the fastest Generative AI inference solution in the world, over 10 times faster than GPU-based hyperscale cloud inference services. This order of magnitude increase in speed is transforming the user experience of AI applications, unlocking real-time iteration and increasing intelligence via additional agentic computation. About The Role As an Applied AI Scientist in the FieldML team, you will be responsible for developing and customizing large language models and more broadly large-scale deep learning models to solve specific customer problems. You won't just advise; you will build. You will bridge the gap between state-of-the-art research and real-world applications by helping customers harness the power of the Cerebras Wafer-Scale Engine (WSE) for their AI initiatives. We are looking for experienced AI Scientists who are passionate about the "applied" side of machine learning - those --- TITLE: Site Reliability Engineer EMPLOYER: Accenture Federal Services LOCATION: Arlington, VA (unspecified) SALARY: $100K-$203K POSTED: 2026-05-26 APPLY_URL: https://boards.greenhouse.io/accenturefederalservices/jobs/4683281006?gh_jid=4683281006 EXCERPT: Site Reliability Engineer Arlington, VA At Accenture Federal Services, nothing matters more than helping the US federal government make the nation stronger and safer and life better for people. Our 13,000+ people are united in a shared purpose to pursue the limitless potential of technology and ingenuity for clients across defense, national security, public safety, civilian, and military health organizations. Join Accenture Federal Services, a technology company within global Accenture. Recognized as a Glassdoor Top 100 Best Place to Work, we offer a collaborative and caring community where you feel like you belong and are empowered to grow, learn and thrive through hands-on experience, certifications, industry training and more. Join us to drive positive, lasting change that moves missions and the government forward! The work As a Site Reliability Engineer, you will play a pivotal role in advancing operational AI adoption within a cutting-edge Hub-and-Spoke architecture. Your primary focus will be on ensuring the reliability, scalability, and continuous monitoring of enterprise AI systems that support mission-critical applications and enterprise AI governance Key responsibilities: - Ensure the reliability, scalability, and performance of enterprise AI systems within a modern Hub-and-Spoke architecture - Lead incident response efforts to minimize downtime and maintain service continuity - Implement and manage SLOs/SLAs, capacity planning, and performance optimization strategies - Operate and enhance observability platforms using OpenTelemetry, Prometheus, Grafana, Loki, and Tempo Drive FinOps practices to optimize operational costs and resource utilization - Collaborate with cross-functional teams in AI, DevSecOps, data engineering, platform engineering, and cybersecurity - --- TITLE: Postdoctoral Scientist, Amazon Robotics Research and AI Development EMPLOYER: Amazon LOCATION: North Reading, Massachusetts, USA (unspecified) SALARY: $143K-$193K POSTED: 2026-04-16 PARENTAL_LEAVE_WEEKS: 6 (not source-backed) NON_BIRTH_PARENT_LEAVE_WEEKS: 6 (not source-backed) K401_MATCH: yes (not source-backed) FERTILITY_FAMILY_BUILDING_BENEFITS: yes (not source-backed) ADOPTION_ASSISTANCE_OFFERED: yes (not source-backed) MENTAL_HEALTH_SUPPORT: yes (not source-backed) CHILDCARE_SUBSIDY: yes (not source-backed) APPLY_URL: https://www.amazon.jobs/en/jobs/3117815/postdoctoral-scientist-amazon-robotics-research-and-ai-development EXCERPT: Postdoctoral Scientist, Amazon Robotics Research and AI Development North Reading, Massachusetts, USA Amazon is looking for talented Postdoctoral Scientists to join the Research and AI Development team at Amazon Robotics for a one-year, full-time research position with an optional extension for a second year. This Postdoctoral Scientist will innovate in the areas of multi-agent path planning, dynamic optimal transport (OT), and explainable AI (X-AI) for DeepFleet Foundation Models. They will have the opportunity to develop optimal and scalable solutions for the world's largest fleet of mobile robots, in addition to developing interpretability techniques for the DeepFleet FMs. At Amazon, we experiment and innovate relentlessly. Science is core in our offering to shoppers, advertisers and customers. Our scientists apply machine learning, optimization, and probabilistic modeling at scale to enhance customer experience, help advertisers reach relevant audiences, and support brand building. We are seeking talented scientists to invent new techniques in a variety of areas and innovate on behalf of shoppers, advertisers, and customers. Key job responsibilities In this role you will: -Work closely with a senior science advisor, collaborate with other scientists and engineers, and be part of Amazon's diverse global science community. -Publish your innovation in top-tier academic venues and hone your presentation skills. -Be inspired by challenges and opportunities to invent new techniques in your area(s) of expertise. A day in the life On a typical day in this role, you will work to progress your research projects, meet with engineering, systems, and solutions stakeholders, brainstorm with other scientists --- TITLE: Senior Site Reliability Engineer EMPLOYER: Ditto LOCATION: Remote (Atlanta, Austin, San Francisco, Seattle) (remote) SALARY: Not disclosed POSTED: 2026-06-07 APPLY_URL: https://jobs.ashbyhq.com/ditto/89254e25-82f7-4634-960a-a6f1f35bbae0 EXCERPT: Senior Site Reliability Engineer Remote (Atlanta, Austin, San Francisco, Seattle) About Ditto: Ditto is redefining how data moves at the edge. Our mission is to make it seamless for developers to build resilient, real-time applications, regardless of network conditions. Whether you're in a stadium, airplane, or remote military base, Ditto's peer-to-peer sync engine ensures devices stay connected and data stays consistent, even without internet. With more than $145 million in funding and trusted by organizations like Chick-fil-A, Delta Airlines, and the U.S. military, Ditto powers mission-critical experiences across aviation, retail, travel, hospitality, defense, and more. As a globally distributed, fast-growing startup, we're committed to building a diverse and inclusive team that reflects the wide range of perspectives needed to solve the world's hardest connectivity problems. About the position Ditto is at an inflection point. As we scale to meet the demands of our enterprise customers, we need experienced Site Reliability Engineers to ensure our infrastructure delivers enterprise-grade reliability. This is a unique opportunity to join a specialized team focused on observability, system reliability and operational excellence for our cutting-edge, edge-to-cloud, database technology. As a Site Reliability Engineer, you will play a crucial role in ensuring the reliability, performance, and scalability of Ditto's cloud infrastructure. You'll collaborate with product engineering teams to improve system resilience, lead and develop incident management processes and build observability solutions for our unique distributed architecture. As a Site Reliability Engineer, you will: - Develop and maintain observability solutions using platforms like Datadog, Prometheus and Grafana - --- TITLE: Senior Avionics Manufacturing Engineer, Electromechanical EMPLOYER: Relativity LOCATION: Long Beach, California (unspecified) SALARY: $130K-$166K POSTED: 2025-12-11 APPLY_URL: https://boards.greenhouse.io/relativity/jobs/8334809002?gh_jid=8334809002 EXCERPT: Senior Avionics Manufacturing Engineer, Electromechanical Long Beach, California At Relativity Space, we're building rockets to serve today's needs and tomorrow's breakthroughs. Our Terran R vehicle will deliver customer payloads to orbit, meeting the growing demand for launch capacity. But that's just the start. Achieving commercial success with Terran R will unlock new opportunities to advance science, exploration, and innovation, pioneering progress that reaches beyond the known. Joining Relativity means becoming part of something where autonomy, ownership, and impact exist at every level. Here, you're not just executing tasks; you're solving problems that haven't been solved before, helping develop a rocket, a factory, and a business from the ground up. Whether you're in propulsion, manufacturing, software, avionics, or a corporate function, you'll collaborate across teams, shape decisions, and see your work come to life in record time. Relativity is a place where creativity and technical rigor go hand in hand, and your voice will help define the stories we're writing together. Now is a unique moment in time where it's early enough to leave your mark on the product, the process, and the culture, but far enough along that Terran R is tangible and picking up momentum. The most meaningful work of your career is waiting. Join us. About the Team: The Avionics team is responsible for the full lifecycle of Terran R's nervous system, designing, building, testing, installing, and operating the hardware that connects and controls every major electrical system on the vehicle and ground. The team's structure intentionally combines --- TITLE: Backend Engineer EMPLOYER: Datacurve LOCATION: San Francisco | OnSite (onsite) SALARY: Not disclosed POSTED: 2026-06-10 APPLY_URL: https://jobs.ashbyhq.com/datacurve/a9c95033-12f6-4c5f-8d64-9c29aa80fc25 EXCERPT: Backend Engineer San Francisco | OnSite We're building a gamified developer platform empowering tens of thousands of coders to compete in exciting software engineering challenges - all while pushing the frontier of LLMs! As a Backend Engineer, you'll become part of a passionate, fast moving team dedicated to solving challenging technical problems, designing engaging gamification experiences, and shaping the foundations to support the collective effort of developers worldwide. WHAT YOU'LL DO - Lead technical discussions, guide architectural decisions, and identify opportunities for improvement across backend systems - Architect robust infrastructure to efficiently handle high volumes of user interactions, data processing, and real-time competition data - Collaborate closely with frontend engineers and product managers to deliver seamless user experiences - Optimize backend performance, reliability, and scalability to support rapid growth and evolving product requirements - Establish backend engineering best practices, including code quality, testing, observability, and documentation WHAT YOU HAVE - 3+ years of experience designing, building, and maintaining scalable backend systems and APIs - Solid understanding of distributed systems, asynchronous processing, and event-driven architectures - Experience designing APIs and backend services using Go - Familiarity with cloud infrastructure and services on AWS - Proficiency with IaaC tools (e.g., Terraform) and CI/CD pipelines - Expertise designing relational database schemas, optimizing SQL queries, and managing database performance and integrity - Excellent collaboration and proactive communication skills NICE TO HAVES - Experience with Kubernetes and orchestration of containerized applications - Knowledge of observability and monitoring tools (e.g., Prometheus, --- TITLE: Mechanical Engineer EMPLOYER: Cerebras Systems LOCATION: Sunnyvale, CA (unspecified) SALARY: $180K-$200K POSTED: 2026-05-29 K401_MATCH: yes (not source-backed) APPLY_URL: https://job-boards.greenhouse.io/cerebrassystems/jobs/7752149003 EXCERPT: Mechanical Engineer Sunnyvale, CA Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. Our novel wafer-scale architecture provides the AI compute power of dozens of GPUs on a single chip, with the programming simplicity of a single device. This approach allows Cerebras to deliver industry-leading training and inference speeds and empowers machine learning users to effortlessly run large-scale ML applications, without the hassle of managing hundreds of GPUs or TPUs. Cerebras' current customers include top model labs, global enterprises, and cutting-edge AI-native startups. OpenAI recently announced a multi-year partnership with Cerebras , to deploy 750 megawatts of scale, transforming key workloads with ultra high-speed inference. Thanks to the groundbreaking wafer-scale architecture, Cerebras Inference offers the fastest Generative AI inference solution in the world, over 10 times faster than GPU-based hyperscale cloud inference services. This order of magnitude increase in speed is transforming the user experience of AI applications, unlocking real-time iteration and increasing intelligence via additional agentic computation. The Role: As a Mechanical Engineer at Cerebras, you will lead the design of mechanical systems for our next-generation wafer-scale engine. Your responsibilities will include ensuring compliance with specifications, validating manufacturability, and delivering a high-quality product in a fast-paced environment-tackling some of the most challenging problems in the rapidly evolving AI space. In this role, you will develop mechanical infrastructure for Cerebras' custom hardware system. - Rapidly iterate on designs and analysis to inform high level systems trades and steer overall product direction. - Provide comprehensive support for --- TITLE: Openstack Engineer EMPLOYER: Infosys Consulting LOCATION: Location not specified (unspecified) SALARY: Not disclosed POSTED: 2026-06-10 K401_MATCH: yes (not source-backed) APPLY_URL: https://sjobs.brassring.com/TGnewUI/Search/home/HomeWithPreLoad?partnerid=25633&siteid=5439&PageType=JobDetails&jobid=2246228 EXCERPT: Openstack Engineer Role Overview Supports OpenStack‑based telco cloud infrastructure. Expanded Key Responsabilities • Operates and maintains OpenStack compute, storage, and network services • Diagnoses and resolves infrastructure incidents and performance issues • Performs capacity planning, scaling, and lifecycle management • Supports integrations with orchestration, monitoring, and automation layers • Ensures availability and resilience of NFV infrastructure • Operates and maintains OpenStack compute, storage, and network services Skillset - Strong Linux administration (RHEL / CentOS / Rocky) - Process, memory, disk, CPU troubleshooting - Systemd, journald, log analysis - Networking basics: ports, firewalls, DNS, routing - SSH, SCP, file permissions, SELinux (basic) For OpenStack‑based Ops roles (NFVI / Infra Ops). Tools: - openstack CLI - Service logs - Prometheus (metrics basics) - Grafana dashboards - Alerts & thresholds - Log analysis (Splunk / ELK / journald) Basic understanding of: - MTTR - Alert noise reduction - Basic Ansible playbooks (run & troubleshoot) - Shell scripting (log cleanup, checks) - Terraform awareness (read & validate, not design) - CI/CD awareness (pipelines, failures) About Infosys Infosys is a global leader in next-generation digital services and consulting. We enable clients in 46 countries to navigate their digital transformation. With over three decades of experience in managing the systems and workings of global enterprises, we expertly steer our clients through the many next of their digital journey. We do it by enabling the enterprise with an AI-powered core that helps prioritize the execution of change. We also empower the business with agile digital at scale --- TITLE: Applied Scientist II, Amazon AWS Agentic AI, AWS AI Fundamental Research EMPLOYER: Amazon LOCATION: Santa Clara, California, USA (unspecified) SALARY: $172K-$222K POSTED: 2026-05-27 PARENTAL_LEAVE_WEEKS: 6 (not source-backed) NON_BIRTH_PARENT_LEAVE_WEEKS: 6 (not source-backed) K401_MATCH: yes (not source-backed) FERTILITY_FAMILY_BUILDING_BENEFITS: yes (not source-backed) ADOPTION_ASSISTANCE_OFFERED: yes (not source-backed) SURROGACY_ASSISTANCE_OFFERED: yes (not source-backed) MENTAL_HEALTH_SUPPORT: yes (not source-backed) CHILDCARE_SUBSIDY: yes (not source-backed) APPLY_URL: https://www.amazon.jobs/en/jobs/10431050/applied-scientist-ii-amazon-aws-agentic-ai-aws-ai-fundamental-research EXCERPT: Applied Scientist II, Amazon AWS Agentic AI, AWS AI Fundamental Research Santa Clara, California, USA Amazon is looking for a passionate, talented, and inventive Applied Scientist with a strong machine learning background to help build industry-leading technology in generative AI and foundational models. As part of our AI team in Amazon AWS, you will work alongside internationally recognized experts to develop novel algorithms and modeling techniques to advance the state-of-the-art in generative AI. Your work will directly impact millions of our customers in the form of products and services that make use of speech, vision and language technology. You will gain hands on experience with Amazon's heterogeneous speech, text, image and structured data sources, and large-scale computing resources to accelerate advances in machine learning and foundation models. More specifically, you will have the opportunity to impact millions of our customers by researching and building innovative solutions using Agentic AI. Agentic AI drives innovation at the forefront of artificial intelligence, enabling customers to transform their businesses through cutting-edge generative AI solutions. We build and deliver the foundational AI services that power the future of cloud computing, helping organizations harness the potential of AI to solve their most complex challenges. Join our dynamic team of AI/ML practitioners, applied scientists, software engineers, and solution architects who work backwards from customer needs to create groundbreaking technologies. If you're passionate about shaping the future of AI while making a meaningful impact for customers worldwide, we want to hear from you. A day in the life --- TITLE: Senior Backend Engineer (Golang) EMPLOYER: SumUp LOCATION: Vilnius, Lithuania (unspecified) SALARY: EUR 4K-7K/mo POSTED: 2024-06-05 K401_MATCH: yes (not source-backed) LEARNING_BUDGET_OFFERED: yes (not source-backed) APPLY_URL: https://sumup.com/careers/positions/7480776002?gh_jid=7480776002 EXCERPT: Senior Backend Engineer (Golang) Vilnius, Lithuania In the Global Bank tribe , we are building the infrastructure and core services needed to provide our merchants with a digital bank account that empowers them to be successful at doing what they love. Our goal is to become the most popular banking partner for small merchants globally by offering a high-quality banking experience tailored to their needs-effortless, simple, and affordable. Joining the Global Bank tribe means playing a key role in helping us build SumUp's digital bank. You'll collaborate with a global, autonomous, cross-functional team that refines an aspect of our product from concept to execution. Work alongside colleagues from 32 nationalities across Cologne, Berlin, São Paulo, Sofia, London and Vilnius, united by a shared commitment to taking ownership, working with purpose, and helping small businesses thrive. As a Senior Backend Engineer , you will play a key role in helping us transition from a localized setup with regional banks to a unified Global Bank. A significant challenge will involve supporting the move to a fully distributed, event-driven architecture. Our tech stack includes Go, AWS, Kafka, PostgreSQL, and Kubernetes, supported by a strong observability toolchain with Prometheus, Grafana, and Honeycomb. We also actively use AI‑assisted development tools such as Cursor, GitHub Copilot, and others. SumUp's engineering culture is built around ownership. We follow a “build it, run it, improve it” model, with strong observability and a focus on learning. Engineers have clear scope and autonomy, with decisions documented, reviewed, and made collaboratively. --- TITLE: Sr. Member of Technical Staff EMPLOYER: Cerebras Systems LOCATION: Sunnyvale, CA (unspecified) SALARY: $230K-$250K POSTED: 2026-05-08 K401_MATCH: yes (not source-backed) APPLY_URL: https://job-boards.greenhouse.io/cerebrassystems/jobs/7728796003 EXCERPT: Sr. Member of Technical Staff Sunnyvale, CA Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. Our novel wafer-scale architecture provides the AI compute power of dozens of GPUs on a single chip, with the programming simplicity of a single device. This approach allows Cerebras to deliver industry-leading training and inference speeds and empowers machine learning users to effortlessly run large-scale ML applications, without the hassle of managing hundreds of GPUs or TPUs. Cerebras' current customers include top model labs, global enterprises, and cutting-edge AI-native startups. OpenAI recently announced a multi-year partnership with Cerebras , to deploy 750 megawatts of scale, transforming key workloads with ultra high-speed inference. Thanks to the groundbreaking wafer-scale architecture, Cerebras Inference offers the fastest Generative AI inference solution in the world, over 10 times faster than GPU-based hyperscale cloud inference services. This order of magnitude increase in speed is transforming the user experience of AI applications, unlocking real-time iteration and increasing intelligence via additional agentic computation. Cerebras Systems Inc. has multiple openings for Sr. Member of Technical Staff Title: Sr. Member of Technical Staff Job Duties: - Design and develop software features that support system resiliency and high availability, including automated recovery mechanisms and fault-tolerant architecture across distributed environments. - Develop and maintain cloud-based deployment workflows for AI inference software using AWS tools and services to support low-latency and scalable system performance. - Develop Python-based scripts and APIs to streamline data preprocessing, inference execution, and post-processing for real-time inference tasks. - --- TITLE: Backend Engineer EMPLOYER: LiteLLM LOCATION: San Francisco | OnSite (onsite) SALARY: Not disclosed POSTED: 2026-06-10 APPLY_URL: https://jobs.ashbyhq.com/litellm/759314e2-8415-4946-8701-88b1a7fad1b5 EXCERPT: Backend Engineer San Francisco | OnSite Backend Engineer LiteLLM is the world's most popular AI Gateway, trusted by top companies like Adobe, Netflix, and NASA. Our platform empowers developers by providing secure, reliable access to LLMs and adjacent services, and we're looking for a Senior Backend Engineer to help us build rock-solid guardrails and observability tooling at scale. About The Role You'll focus on owning our guardrails and logging world-class. You will be in charge of the backend code that ensures all guardrail calls are consistently logged, errors are surfaced to users (not silently swallowed), and our observability instrumentation works for real-world, high-volume traffic. Your attention to detail in areas like latency metrics, logging traceability, and backend guardrail registration will directly impact user trust in our security and compliance features. Responsibilities - Build and scale our product, ensuring performance, reliability, and continuous improvement. - Ensure all guardrail and policy enforcement calls (e.g., applyguardrail) are properly logged and traceable through our SpendLogs and relevant database tables - Build and design CPU-level guardrails to cover common attacks on LLM API's / MCP servers / Agents - Identify and fix areas where silent failures occur in guardrail creation, registration, and policy application-ensuring robust error handling and transparency to end users - Work with observability integrations, including Datadog, Splunk, Prometheus, and OpenTelemetry, to maintain accurate, configurable, and usable monitoring and logging for backend systems - Enhance observability integrations to work for 1B+ requests/mo., with minimal latency overhead and no memory leaks (e.g. due to --- TITLE: Applied Scientist, Amazon Robotics R&D EMPLOYER: Amazon LOCATION: Berlin, Berlin, DEU (unspecified) SALARY: Not disclosed POSTED: 2026-03-18 PARENTAL_LEAVE_WEEKS: 6 (not source-backed) NON_BIRTH_PARENT_LEAVE_WEEKS: 6 (not source-backed) K401_MATCH: yes (not source-backed) FERTILITY_FAMILY_BUILDING_BENEFITS: yes (not source-backed) ADOPTION_ASSISTANCE_OFFERED: yes (not source-backed) MENTAL_HEALTH_SUPPORT: yes (not source-backed) CHILDCARE_SUBSIDY: yes (not source-backed) APPLY_URL: https://www.amazon.jobs/en/jobs/3207770/applied-scientist-amazon-robotics-r-d EXCERPT: Applied Scientist, Amazon Robotics R&D Berlin, Berlin, DEU The Amazon Robotics team is seeking an experienced Applied Scientist to join our team. In this role you will apply the latest trends in research to solve real-world problems in robotics and AI. You will collaborate with a team of scientists and engineers building these applications. We holistically design, build, and deliver end-to-end robotic systems. Our team is also responsible for core infrastructure and tools that serve as the backbone of our robotic applications, enabling roboticists, machine learning scientists, software engineers, and hardware engineers to collaborate and deploy systems in the field. Key job responsibilities • Research, design, implement and evaluate complex computer vision and decision making algorithms integrating across multiple disciplines and leveraging machine learning. • Create experiments and prototype implementations of new learning algorithms and prediction techniques. • Work closely with software engineering team members to drive scalable, real-time implementations. • Collaborate with machine learning and robotic controls experts to implement and deploy algorithms, such as machine learning models. • Collaborate closely with hardware engineering team members on developing systems from prototyping to production level. • Represent Amazon in academia community through publications and scientific presentations. • Work with stakeholders across hardware, science, and operations teams to iterate on systems design and implementation. Basic Qualifications: - PhD in engineering, technology, computer science, machine learning, robotics, operations research, statistics, mathematics or equivalent quantitative field - Experience in publishing at major robotics and related conferences (e.g. RSS, NIPS, ICRA, CVPR, CORL, ICCV) --- TITLE: Senior Site Reliability Engineer, APAC EMPLOYER: Ditto LOCATION: APAC | Remote (remote) SALARY: Not disclosed POSTED: 2026-06-07 APPLY_URL: https://jobs.ashbyhq.com/ditto/c37d394c-3ab2-4ba3-b5a3-0e8513a4c758 EXCERPT: Senior Site Reliability Engineer, APAC APAC | Remote About Ditto: Ditto is redefining how data moves at the edge. Our mission is to make it seamless for developers to build resilient, real-time applications, regardless of network conditions. Whether you're in a stadium, airplane, or remote military base, Ditto's peer-to-peer sync engine ensures devices stay connected and data stays consistent, even without internet. With more than $145 million in funding and trusted by organizations like Chick-fil-A, Delta Airlines, and the U.S. military, Ditto powers mission-critical experiences across aviation, retail, travel, hospitality, defense, and more. As a globally distributed, fast-growing startup, we're committed to building a diverse and inclusive team that reflects the wide range of perspectives needed to solve the world's hardest connectivity problems. About the position Ditto is at an inflection point. As we scale to meet the demands of our enterprise customers, we need experienced Site Reliability Engineers to ensure our infrastructure delivers enterprise-grade reliability. This is a unique opportunity to join a specialized team focused on observability, system reliability and operational excellence for our cutting-edge, edge-to-cloud, database technology. As a Site Reliability Engineer, you will play a crucial role in ensuring the reliability, performance, and scalability of Ditto's cloud infrastructure. You'll collaborate with product engineering teams to improve system resilience, lead and develop incident management processes and build observability solutions for our unique distributed architecture. As a Site Reliability Engineer, you will: - Develop and maintain observability solutions using platforms like Datadog, Prometheus and Grafana - Take a --- TITLE: Site Reliability Engineer (SRE) EMPLOYER: Apple LOCATION: San Diego, United States of America (unspecified) SALARY: $140K-$258K POSTED: 2025-12-11 K401_MATCH: yes (not source-backed) APPLY_URL: https://jobs.apple.com/en-us/details/200636310/site-reliability-engineer-sre?team=SFTWR EXCERPT: Site Reliability Engineer (SRE) San Diego, United States of America The Video Computer Vision organization is working on exciting technologies for future Apple products. Our focus is on ML based solution around real time image and video. We have contributed to the FaceID and FaceKit project in the past and more recently the new LIDAR iPad sensor. We are looking for the right Site Reliability Engineer to help us take our efforts to the next level. In this role, you will help lead our cloud based infrastructure team for Apple's Video Computer Vision Organization. As a main contributor to our SRE team you will develop and maintain infrastructure, tooling, and engineering services for cloud based applications. You will be responsible for system bringup, deployment, reliability, security and service scalability. This role is highly cross-functional and you will work very closely with various highly skilled software development / ML teams developing cutting edge algorithms. Your core responsibility is to provide operational support of multiple cloud based applications with an emphasis on deployment, security, scalability and reliability running on AWS and Apple infrastructure. Our technologies include Terraform, Argo, Docker, Python, Postgres, Prometheus, in combination with custom Apple software and tooling. Common technologies you'll manage include: Kubernetes (eks), Elasticsearch, Redis, RDS, ELB, and other AWS based services. This role will also help drive solutions for hybrid infrastructure (on and off prem) and drive infrastructure architecture for our AWS based cloud platform. Minimum Qualifications: Experience building systems both on-premise (data center) and on public --- TITLE: Security & IT General Opportunities EMPLOYER: Cerebras Systems LOCATION: Sunnyvale CA or Toronto Canada (unspecified) SALARY: Not disclosed POSTED: 2026-05-28 K401_MATCH: yes (not source-backed) APPLY_URL: https://job-boards.greenhouse.io/cerebrassystems/jobs/7751790003 EXCERPT: Security & IT General Opportunities Sunnyvale CA or Toronto Canada Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. Our novel wafer-scale architecture provides the AI compute power of dozens of GPUs on a single chip, with the programming simplicity of a single device. This approach allows Cerebras to deliver industry-leading training and inference speeds and empowers machine learning users to effortlessly run large-scale ML applications, without the hassle of managing hundreds of GPUs or TPUs. Cerebras' current customers include top model labs, global enterprises, and cutting-edge AI-native startups. OpenAI recently announced a multi-year partnership with Cerebras , to deploy 750 megawatts of scale, transforming key workloads with ultra high-speed inference. Thanks to the groundbreaking wafer-scale architecture, Cerebras Inference offers the fastest Generative AI inference solution in the world, over 10 times faster than GPU-based hyperscale cloud inference services. This order of magnitude increase in speed is transforming the user experience of AI applications, unlocking real-time iteration and increasing intelligence via additional agentic computation. About the Team Our IT & Security team sits at the intersection of security, infrastructure, and cutting-edge AI systems. The team plays a critical role in ensuring Cerebras' environments are secure, reliable, scalable, and ready to support customers operating at the frontier of AI. We are looking for people who care deeply about operational excellence, security best practices, automation, and building systems that can support a rapidly growing organization. What You May Work On Depending on the role and team needs, you --- TITLE: Automation Engineer (Req#1278) EMPLOYER: Eplus LOCATION: India - Remote (remote) SALARY: Not disclosed POSTED: 2026-05-28 K401_MATCH: yes (not source-backed) APPLY_URL: https://job-boards.greenhouse.io/eplusinc/jobs/4253028009 EXCERPT: Automation Engineer (Req#1278) India - Remote Overview We are looking for a passionate and self-motivated Automation Developer with hands-on experience with OpenShift and Kubernetes container platforms to Infrastructure Automation team. You will play a key role in building and maintaining automation frameworks, tools, and integrations using modern infrastructure-as-code practices. This is a hands-on opportunity for a driven engineer who thrives in a collaborative, fast-paced environment and is eager to grow their technical depth in DevOps and open-source technologies. Your Impact - Design, develop, and maintain automation scripts and playbooks using Ansible and Terraform - Develop and maintain integration tools and scripts using Python - Create and consume RESTful APIs to integrate systems and services - Contribute to open-source projects and maintain internal GitHub repositories - Collaborate with DevOps, platform, and development teams to implement CI/CD pipelines and automate operational processes - Troubleshoot automation failures and improve reliability, scalability, and performance of automation systems Qualifications Basic Qualifications: - Proficient in writing Ansible playbooks and using Terraform for infrastructure provisioning - Strong Python scripting skills, with experience in writing reusable and modular code - Experience working with RESTful API integrations - 3+ years of hands-on experience with Kubernetes and OpenShift in production environments - Familiarity with GitHub workflows, version control, and collaborative open-source development - Working knowledge of CI/CD and DevOps practices - Comfortable working in Linux-based environments Preferred Qualifications: - Experience with cloud platforms such as AWS, Azure, or GCP - Knowledge of monitoring, logging, and alerting tools (e.g., Prometheus, ELK, --- TITLE: Senior Avionics Hardware Engineer, Interplanetary Sciences Program EMPLOYER: Relativity LOCATION: Long Beach, California (unspecified) SALARY: $148K-$204K POSTED: 2026-06-02 APPLY_URL: https://boards.greenhouse.io/relativity/jobs/8575422002?gh_jid=8575422002 EXCERPT: Senior Avionics Hardware Engineer, Interplanetary Sciences Program Long Beach, California At Relativity Space, we're building rockets to serve today's needs and tomorrow's breakthroughs. Our Terran R vehicle will deliver customer payloads to orbit, meeting the growing demand for launch capacity. But that's just the start. Achieving commercial success with Terran R will unlock new opportunities to advance science, exploration, and innovation, pioneering progress that reaches beyond the known. Joining Relativity means becoming part of something where autonomy, ownership, and impact exist at every level. Here, you're not just executing tasks; you're solving problems that haven't been solved before, helping develop a rocket, a factory, and a business from the ground up. Whether you're in propulsion, manufacturing, software, avionics, or a corporate function, you'll collaborate across teams, shape decisions, and see your work come to life in record time. Relativity is a place where creativity and technical rigor go hand in hand, and your voice will help define the stories we're writing together. Now is a unique moment in time where it's early enough to leave your mark on the product, the process, and the culture, but far enough along that Terran R is tangible and picking up momentum. The most meaningful work of your career is waiting. Join us. About the Team: The Interplanetary Sciences Program was established to expand access to scientific exploration across our solar system. Its mission is to make planetary research faster, more affordable, and more capable than ever before by rethinking how science missions are --- TITLE: Senior Runtime Engineer EMPLOYER: Cerebras Systems LOCATION: Sunnyvale CA or Toronto Canada (unspecified) SALARY: Not disclosed POSTED: 2025-10-28 K401_MATCH: yes (not source-backed) APPLY_URL: https://job-boards.greenhouse.io/cerebrassystems/jobs/7510635003 EXCERPT: Senior Runtime Engineer Sunnyvale CA or Toronto Canada Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. Our novel wafer-scale architecture provides the AI compute power of dozens of GPUs on a single chip, with the programming simplicity of a single device. This approach allows Cerebras to deliver industry-leading training and inference speeds and empowers machine learning users to effortlessly run large-scale ML applications, without the hassle of managing hundreds of GPUs or TPUs. Cerebras' current customers include top model labs, global enterprises, and cutting-edge AI-native startups. OpenAI recently announced a multi-year partnership with Cerebras , to deploy 750 megawatts of scale, transforming key workloads with ultra high-speed inference. Thanks to the groundbreaking wafer-scale architecture, Cerebras Inference offers the fastest Generative AI inference solution in the world, over 10 times faster than GPU-based hyperscale cloud inference services. This order of magnitude increase in speed is transforming the user experience of AI applications, unlocking real-time iteration and increasing intelligence via additional agentic computation. About The Role We are building the next generation of large-scale AI systems that power training and inference workloads at unprecedented scale and efficiency. You will design and develop high-performance distributed software that orchestrates massive compute and data pipelines across heterogeneous clusters. Your work will push the limits of concurrency, throughput, and scalability-enabling efficient execution of models at massive scale. This role sits at the intersection of systems engineering and machine learning performance, demanding both architectural depth and low-level implementation skills. You will help --- TITLE: Full Stack Software Engineer EMPLOYER: Zelos Cloud LOCATION: San Francisco, United States (unspecified) SALARY: $115K-$180K POSTED: 2025-07-28 APPLY_URL: https://jobs.gem.com/zeloscloud/am9icG9zdDqJhHJu1Y4lzIsCvOIOAbSw EXCERPT: Full Stack Software Engineer San Francisco, United States Overview Zelos is building the data platform for mission-critical systems, helping teams observe, control, and test their systems on a collaborative platform - from prototype to production. Founded by Tesla and Neuralink engineers, backed by YC and Human Capital. As a Full Stack Software Engineer, you'll be responsible for building up the cloud feature set. This is an opportunity to create scalable systems that bridge the gap between embedded devices, test systems, and collaborative engineering workflows. You'll enable teams to debug, develop, and manage physical systems at scale through innovative cloud solutions. What we're looking for - Strong proficiency in JavaScript/TypeScript, with frameworks like React - Proficiency with backend languages (Golang, Rust) for building scalable services - Proven track record working with cloud-native solutions on AWS, Azure, or GCP - Experience with modern build and deployment tooling (NixOS, Docker, Kubernetes) - Experience with time-series databases (InfluxDB, TimescaleDB) and visualization platforms (Grafana, Prometheus) What to expect You'll join a team of seasoned engineers building cutting-edge solutions for the next generation of connected systems. We value first-principles thinking, hands-on problem solving, and collaborative engineering. As a startup, you'll have significant influence over our technical direction and the opportunity to build systems that directly impact how teams develop and deploy mission-critical hardware. We understand that no one individual knows everything. We will all learn together and from each other. We strive to build a collaborative, enriching environment conducive to personal, technical, and career growth. Compensation --- TITLE: Staff Machine Learning Engineer – (ADAS/Autonomous Driving) EMPLOYER: Lucid Motors LOCATION: Newark, CA (unspecified) SALARY: $181K-$265K POSTED: 2025-01-30 K401_MATCH: yes (not source-backed) APPLY_URL: https://job-boards.greenhouse.io/lucidmotors/jobs/4620390007 EXCERPT: Staff Machine Learning Engineer – (ADAS/Autonomous Driving) Newark, CA Leading the future in luxury electric and mobility At Lucid, we set out to introduce the most captivating, luxury electric vehicles that elevate the human experience and transcend the perceived limitations of space, performance, and intelligence. Vehicles that are intuitive, liberating, and designed for the future of mobility. We plan to lead in this new era of luxury electric by returning to the fundamentals of great design - where every decision we make is in service of the individual and environment. Because when you are no longer bound by convention, you are free to define your own experience. Come work alongside some of the most accomplished minds in the industry. Beyond providing competitive salaries, we're providing a community for innovators who want to make an immediate and significant impact. If you are driven to create a better, more sustainable future, then this is the right place for you. About Lucid At Lucid, we are redefining the luxury electric vehicle experience, combining cutting-edge technology with exceptional design to deliver intuitive, safe, and exhilarating mobility. Join our team and help shape the future of autonomous driving. The Role We are seeking a Staff Software Engineer to lead the integration and deployment of advanced perception models into Lucid's production ADAS and autonomous driving systems. This role focuses on productizing ML models , ensuring robust performance on automotive-grade hardware, and building scalable pipelines for deployment and validation. You will collaborate with ML researchers, perception engineers, --- TITLE: Senior Software Engineer, ML Research EMPLOYER: Lila Sciences LOCATION: Cambridge, MA USA (unspecified) SALARY: $148K-$210K POSTED: 2025-10-06 LEARNING_BUDGET_OFFERED: yes (not source-backed) APPLY_URL: https://job-boards.greenhouse.io/lilasciences/jobs/4031328009 EXCERPT: Senior Software Engineer, ML Research Cambridge, MA USA Your Impact at LILA We're hiring a Senior Software Engineer to team with our Machine Learning Engineers and Researchers to build software that powers Lila's ML workflows and research tooling. You will be on a team of software engineers that work alongside machine learning experts to develop, support, and maintain Lila's ML libraries and tools. What You'll Be Building - Design and build high-performance, secure, and well-documented Machine Learning libraries that implement the algorithms designed by machine learning researchers and engineers. - CI/CD pipelines and integration tests for ML workflows. - Architect repository designs to follow consistent standards. - Support debugging, logging, and maintenance for Ray-based compute environments. - Data ingestion pipelines linking lab data to ML teams. What You'll Need to Succeed - Minimum of 8 years of experience writing software in a commercial setting in Go or Python. - Experience implementing scalable software solutions. - Experience with MLOps systems, GitOps (ArgoCD, Github Actions) - Familiarity with orchestration frameworks (Ray, Argo, Airflow). - Proficiency in Containerization, Kubernetes and infrastructure-as-code tools. - Acute listening skills and patience to deeply understand complex problems and algorithms. - Excellent problem-solving skills and team-first mentality. - Energetic self-starter and independent thinker with strong attention to detail. - Eager to work with highly skilled and dynamic teams in a fast-paced, entrepreneurial, and technical setting. Bonus Points For - Proficiency with monitoring and logging tools (e.g., Prometheus, Grafana) - Research engineering or scientific software experience. Compensation We offer --- TITLE: ALTERNANT - Software IOT - F/H EMPLOYER: STMicroelectronics LOCATION: Rennes, France (unspecified) SALARY: Not disclosed POSTED: 2026-06-12 K401_MATCH: yes (not source-backed) APPLY_URL: https://stmicroelectronics.eightfold.ai/careers/job/563637172464459 EXCERPT: ALTERNANT - Software IOT - F/H Rennes, France --- TITLE: Senior Backend Engineer - Accounts EMPLOYER: SumUp LOCATION: Sofia, Bulgaria (unspecified) SALARY: Not disclosed POSTED: 2026-02-11 K401_MATCH: yes (not source-backed) LEARNING_BUDGET_OFFERED: yes (not source-backed) APPLY_URL: https://sumup.com/careers/positions/8418287002?gh_jid=8418287002 EXCERPT: Senior Backend Engineer - Accounts Sofia, Bulgaria In the Global Bank tribe , we're building the infrastructure to provide merchants with a digital business account that empowers them to manage their banking needs. Our goal is to become the most popular banking partner for small merchants globally with an effortless, simple, and affordable experience. You'll help us transition from fragmented regional setups to a unified global infrastructure, directly enabling millions of merchants worldwide to access seamless banking tailored to their needs. As a Senior Backend Engineer on the Global Accounts team, you'll own critical pieces of our bank account platform. You'll help us design and build a fully distributed, event-driven system designed to scale across regions with resilience and compliance built in. You'll work primarily in Kotlin, with opportunities in Elixir and Golang. We practice Extreme Programming: small iterations, daily deliveries, and a focus on technical design quality and deep problem understanding. Our tech stack includes Kotlin, Golang, Elixir, Java, AWS, Kafka, PostgreSQL, and Kubernetes, supported by observability tools like Prometheus, Grafana, and Honeycomb. We also use AI-assisted development tools including Cursor and GitHub Copilot. What You'll Do - Build critical infrastructure from scratch: Contribute to the design and implementation of a newly architected global accounts platform. You'll help migrate existing systems to a modern, event-driven, decoupled architecture that enables scalability and resilience across regions. - Master event-driven architecture: You'll be working extensively with Kafka to build a truly decoupled, resilient system. Event-driven architecture is essential to our goal of --- TITLE: OpenStack Engineer EMPLOYER: Avepoint Inc LOCATION: Singapore (unspecified) SALARY: Not disclosed POSTED: 2025-04-09 K401_MATCH: yes (not source-backed) APPLY_URL: https://job-boards.greenhouse.io/avepoint/jobs/6787743 EXCERPT: OpenStack Engineer Singapore Key Responsibilities: • Design and implement OpenStack-based cloud solutions tailored to client requirements • Install, configure, and maintain OpenStack components (e.g., Nova, Neutron, Cinder, Swift) • Troubleshoot and resolve issues related to OpenStack deployments • Optimise OpenStack performance, scalability, and reliability • Implement security best practices and ensure compliance with industry standards • Automate OpenStack deployment and management processes • Collaborate with networking and storage teams to integrate OpenStack with existing infrastructure • Provide technical guidance and support to clients and internal teams • Stay current with OpenStack developments and contribute to the open-source community • Participate in on-call rotations for critical support Qualifications: • Bachelor's degree in computer science, Information Technology, or related field • 2+ years of experience working with OpenStack in production environments • Experience with OpenStack deployment tools (e.g., Ansible) • Proficiency in Linux system administration • Proficiency in scripting languages (e.g., Python, Bash) • Strong understanding of networking concepts (VLANs, SDN, load balancing) • Familiarity with containerisation technologies (Docker, Kubernetes) • Knowledge of storage systems and protocols (e.g., Ceph, iSCSI, NFS) • Preferred Skills: o OpenStack certification (e.g., COA - Certified OpenStack Administrator) o Experience with other cloud platforms (e.g., AWS, Azure, GCP) o Knowledge of CI/CD pipelines and DevOps practices o Familiarity with monitoring tools (e.g., Prometheus, Grafana) o Experience with database administration (e.g., MySQL, PostgreSQL) Any personal data you share with us during the application process will be processed strictly in compliance with applicable data protection laws and our Privacy --- TITLE: Forward Deployed Scientist EMPLOYER: Adaptyv LOCATION: Lausanne / Hybrid / Remote | Hybrid | Remote (hybrid) SALARY: Not disclosed POSTED: 2026-06-10 APPLY_URL: https://jobs.ashbyhq.com/adaptyv/6f9f629e-ca05-4cdf-8967-667c88eb19e2 EXCERPT: Forward Deployed Scientist Lausanne / Hybrid / Remote | Hybrid | Remote Adaptyv is building an automated lab thats let AI agents run biology experiments. We're entering the era of agentic science where AI models can now design novel proteins, propose hypotheses, and iterate on experimental results. But they can't run the experiments themselves - that's still a manual, months-long process. We're building the infrastructure that gives AI agents access to the physical world. We are one of the fastest growing biotech companies, trusted by leading biopharmas, frontier AI labs, and the techbio companies pushing the field forward. This is a rare chance to help advance some of the most important work happening in biotech today. Our automated lab is powered by a deep software + hardware stack: lab instruments worth millions of USD reverse-engineered into API-controllable hardware, dozens of devices orchestrated through complex workflows, full observability on everything that happens in the lab, processing pipelines for messy physical-world data, and AI systems that troubleshoot production results and accelerate assay development. We're growing rapidly and are hiring for talented people to scale and support the massive demand for AI-driven wet lab experimentation. ABOUT THE ROLE You're a scientist on staff who works shoulder-to-shoulder with customers to help them design better proteins and get more out of every experiment. You sit between their design models and our lab: helping them scope a campaign, pick the right design approach for the problem, make sense of the data that comes back, and decide --- TITLE: DevOps Engineer EMPLOYER: Cisco LOCATION: Caesarea, Israel (unspecified) SALARY: Not disclosed POSTED: 2026-06-11 K401_MATCH: yes (not source-backed) APPLY_URL: https://cisco.wd5.myworkdayjobs.com/Cisco_Careers/job/Caesarea-Israel/DevOps-Engineer_2013881 EXCERPT: DevOps Engineer Caesarea, Israel Cisco Silicon One is seeking DevOps Engineer Meet the Team Join the Cisco Silicon One team, where we are building a unified silicon architecture for web-scale and service provider networks. You will work in a multi-geography silicon organization that combines the energy and ownership of a startup with the scale and reach of Cisco. Our engineers collaborate across ASIC, hardware board, software infrastructure, IT, and lab teams to help deliver high-performance networking technology. This is a growth-driven environment for engineers who enjoy solving infrastructure problems that directly affect development velocity and operational excellence. Your Impact As a DevOps Infrastructure Engineer, you will help accelerate development by building, maintaining, and improving robust, scalable, and highly stable infrastructure. Your work across CI/CD automation, infrastructure as code, observability, databases, and system optimization will directly improve the speed and reliability of software delivery. You will make it easier for engineering teams to deploy, monitor, troubleshoot, and iterate while maintaining strong operational standards. Maintain and improve CI/CD pipelines using tools such as Jenkins and GitHub Actions, and support infrastructure as code workflows with Puppet and Ansible. Deploy, operate, and improve containerized services using Docker, Kubernetes, and related technologies. Develop and improve on-prem Linux images and infrastructure automation for data center and hybrid environments. Build and operate observability systems using Prometheus, Grafana, Zabbix, ELK/OpenSearch, and related tools. Design dashboards, alerts, and SLO/SLI frameworks that provide real-time visibility into pipelines, labs, and infrastructure. Develop data collection pipelines and exporters to ingest metrics, logs, --- TITLE: Security SWE EMPLOYER: Cerebras Systems LOCATION: Sunnyvale CA or Toronto Canada (unspecified) SALARY: Not disclosed POSTED: 2026-03-11 K401_MATCH: yes (not source-backed) APPLY_URL: https://job-boards.greenhouse.io/cerebrassystems/jobs/7647319003 EXCERPT: Security SWE Sunnyvale CA or Toronto Canada Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. Our novel wafer-scale architecture provides the AI compute power of dozens of GPUs on a single chip, with the programming simplicity of a single device. This approach allows Cerebras to deliver industry-leading training and inference speeds and empowers machine learning users to effortlessly run large-scale ML applications, without the hassle of managing hundreds of GPUs or TPUs. Cerebras' current customers include top model labs, global enterprises, and cutting-edge AI-native startups. OpenAI recently announced a multi-year partnership with Cerebras , to deploy 750 megawatts of scale, transforming key workloads with ultra high-speed inference. Thanks to the groundbreaking wafer-scale architecture, Cerebras Inference offers the fastest Generative AI inference solution in the world, over 10 times faster than GPU-based hyperscale cloud inference services. This order of magnitude increase in speed is transforming the user experience of AI applications, unlocking real-time iteration and increasing intelligence via additional agentic computation. About The Role As a frontend engineer on our AI cloud, you will work on our customer-facing inference, training, and admin consoles and API experiences. In this role, you will be responsible for designing and building responsive, user-friendly frontend interfaces that provide an optimal experience for our developers, handling high traffic and throughput efficiently. Your familiarity with the latest web development frameworks and best practices, coupled with a keen eye for design and user experience, will drive team success. We're looking for talented software engineers --- TITLE: Robotics Engineer EMPLOYER: STMicroelectronics LOCATION: Hal Kirkop, Malta (unspecified) SALARY: Not disclosed POSTED: 2026-06-12 K401_MATCH: yes (not source-backed) APPLY_URL: https://stmicroelectronics.eightfold.ai/careers/job/563637171789166 EXCERPT: Robotics Engineer Hal Kirkop, Malta --- TITLE: Infrastructure Site Reliability Engineer EMPLOYER: Radiant LOCATION: Gloucestershire, United Kingdom (unspecified) SALARY: Not disclosed POSTED: 2026-04-07 APPLY_URL: https://jobs.ashbyhq.com/radiant/4fd777cd-11f9-4fe8-92e1-4d054cb579fd EXCERPT: Infrastructure Site Reliability Engineer Gloucestershire, United Kingdom About Radiant Radiant is redefining how AI infrastructure is built. We design and operate AI-native cloud platforms engineered for sovereignty, performance, and scale. Our infrastructure powers GPU-native workloads, multi-tenant control planes, and high-performance AI systems designed for the most demanding environments. We are not building a generic cloud. We are building purpose-built AI infrastructure - from powered land, to compute, to software . As we scale our platform and expand our engineering organisation, we are looking for leaders who can build strong teams, uphold high standards, and deliver reliably at pace. Job Summary: We're looking for an experienced Infrastructure Site Reliability Engineer to run and evolve our infrastructure stack. You'll contribute across bare-metal, virtualization, and orchestration layers, keeping things stable and secure 24/7 x 365 - all while mentoring teammates, improving process and automation as well as helping translate deep technical concepts for a wide range of collaborators and customers. What You'll Do : - Deploy and operate resilient, scalable infrastructure supporting AI/HPC workloads - Optimize Linux system configuration, BIOS/firmware, kernel, and disk subsystem for performance - Configure, monitor and manage bare-metal infrastructure using IPMI, Redfish, etc - Build and maintain automation scripts and infrastructure as code to support platform lifecycle, as well as simplifying troubleshooting for Incident resolution and provision of tooling for our support organisation - Apply ITSM frameworks: Incident, Major Incident, Change Management, and service improvement. - Maintain and enhance 's observability stack: Prometheus, Grafana, and custom monitoring integrations - --- TITLE: Senior Systems Engineer, OS Automation EMPLOYER: CoreWeave LOCATION: Livingston, NJ / New York City, NY/ Sunnyvale, CA/ Bellevue, WA (unspecified) SALARY: $153K-$242K POSTED: 2024-08-21 K401_MATCH: yes (not source-backed) MENTAL_HEALTH_SUPPORT: yes (not source-backed) CHILDCARE_SUBSIDY: yes (not source-backed) LEARNING_BUDGET_OFFERED: yes (not source-backed) APPLY_URL: https://coreweave.com/careers/job?4396057006&board=coreweave&gh_jid=4396057006 EXCERPT: Senior Systems Engineer, OS Automation Livingston, NJ / New York City, NY/ Sunnyvale, CA/ Bellevue, WA CoreWeave is The Essential Cloud for AI™. Built for pioneers by pioneers, CoreWeave delivers a platform of technology, tools, and teams that enables innovators to build and scale AI with confidence. Trusted by leading AI labs, startups, and global enterprises, CoreWeave combines superior infrastructure performance with deep technical expertise to accelerate breakthroughs and turn compute into capability. Founded in 2017, CoreWeave became a publicly traded company (Nasdaq: CRWV) in March 2025. Learn more at www.coreweave.com . About the Role: SysEng HAVOCK ( H ardware - A cceleration - V irtualization - O perating Systems - C ontainerization - K ernel) CoreWeave is looking for a Senior Systems Engineer who is ready to evolve beyond traditional DevOps. You will start by stabilizing and scaling our Linux OS and Kernel build pipelines. Once the foundation is set, you will lead the transition to AI-native infrastructure , building "smart" workflows that don't just report errors, but understand and fix them. You are a Systems Engineer at heart, but you are ready to apply LLMs, RAG, and predictive modeling to solve infrastructure challenges at scale. Our Team's Stack: - Languages: Python, Go, bash/sh - Observability: Prometheus, Victoria Metrics, Grafana - OS & Kernel: Linux Kernel (custom build), Ubuntu - Hardware: Intel/AMD/ARM CPUs, Nvidia GPUs, DPUs, Infiniband and Ethernet NICs - Containerization: Docker, Kubernetes (k8s), KubeVirt, containerd, kubelet Responsibilities: - Pipeline Architecture: Design, maintain, and automate reproducible OS image --- TITLE: Staff Mechanical Engineer (Mechanisms) EMPLOYER: Relativity LOCATION: Long Beach, California (unspecified) SALARY: $184K-$226K POSTED: 2026-04-23 APPLY_URL: https://boards.greenhouse.io/relativity/jobs/8521183002?gh_jid=8521183002 EXCERPT: Staff Mechanical Engineer (Mechanisms) Long Beach, California At Relativity Space, we're building rockets to serve today's needs and tomorrow's breakthroughs. Our Terran R vehicle will deliver customer payloads to orbit, meeting the growing demand for launch capacity. But that's just the start. Achieving commercial success with Terran R will unlock new opportunities to advance science, exploration, and innovation, pioneering progress that reaches beyond the known. Joining Relativity means becoming part of something where autonomy, ownership, and impact exist at every level. Here, you're not just executing tasks; you're solving problems that haven't been solved before, helping develop a rocket, a factory, and a business from the ground up. Whether you're in propulsion, manufacturing, software, avionics, or a corporate function, you'll collaborate across teams, shape decisions, and see your work come to life in record time. Relativity is a place where creativity and technical rigor go hand in hand, and your voice will help define the stories we're writing together. Now is a unique moment in time where it's early enough to leave your mark on the product, the process, and the culture, but far enough along that Terran R is tangible and picking up momentum. The most meaningful work of your career is waiting. Join us. About the Team: The Structures and Mechanisms team is responsible for the full design, build, and test lifecycle of Terran R's integrated structures and flight mechanisms. Whether you're working on highly iterative mechanisms that move or tight-margin structures that shouldn't, you'll get hands-on hardware --- TITLE: Senior Mechanical Engineer (Mechanisms) EMPLOYER: Relativity LOCATION: Long Beach, California (unspecified) SALARY: $132K-$182K POSTED: 2026-04-23 APPLY_URL: https://boards.greenhouse.io/relativity/jobs/8521182002?gh_jid=8521182002 EXCERPT: Senior Mechanical Engineer (Mechanisms) Long Beach, California At Relativity Space, we're building rockets to serve today's needs and tomorrow's breakthroughs. Our Terran R vehicle will deliver customer payloads to orbit, meeting the growing demand for launch capacity. But that's just the start. Achieving commercial success with Terran R will unlock new opportunities to advance science, exploration, and innovation, pioneering progress that reaches beyond the known. Joining Relativity means becoming part of something where autonomy, ownership, and impact exist at every level. Here, you're not just executing tasks; you're solving problems that haven't been solved before, helping develop a rocket, a factory, and a business from the ground up. Whether you're in propulsion, manufacturing, software, avionics, or a corporate function, you'll collaborate across teams, shape decisions, and see your work come to life in record time. Relativity is a place where creativity and technical rigor go hand in hand, and your voice will help define the stories we're writing together. Now is a unique moment in time where it's early enough to leave your mark on the product, the process, and the culture, but far enough along that Terran R is tangible and picking up momentum. The most meaningful work of your career is waiting. Join us. About the Team: The Structures and Mechanisms team is responsible for the full design, build, and test lifecycle of Terran R's integrated structures and flight mechanisms. Whether you're working on highly iterative mechanisms that move or tight-margin structures that shouldn't, you'll get hands-on hardware --- TITLE: Senior Platform Engineer, Operator EMPLOYER: Ditto LOCATION: Remote (Atlanta, Austin, San Francisco, Seattle) (remote) SALARY: Not disclosed POSTED: 2026-06-07 APPLY_URL: https://jobs.ashbyhq.com/ditto/d94a99db-1acf-408e-b07b-b76942bd57c7 EXCERPT: Senior Platform Engineer, Operator Remote (Atlanta, Austin, San Francisco, Seattle) About Ditto: Ditto is redefining how data moves at the edge. Our mission is to make it seamless for developers to build resilient, real-time applications, regardless of network conditions. Whether you're in a stadium, airplane, or remote military base, Ditto's peer-to-peer sync engine ensures devices stay connected and data stays consistent, even without internet. With more than $145 million in funding and trusted by organizations like Chick-fil-A, Delta Airlines, and the U.S. military, Ditto powers mission-critical experiences across aviation, retail, travel, hospitality, defense, and more. As a globally distributed, fast-growing startup, we're committed to building a diverse and inclusive team that reflects the wide range of perspectives needed to solve the world's hardest connectivity problems. WHAT YOU'LL BE UP TO… - Own the architecture and evolution of the SMBP Operator's CRD API surface, designing extensions that meet enterprise expectations - consistent status conditions, configurable security contexts, fine-grained network configuration, and admission validation that surfaces bad configs at apply time rather than deep in the reconciliation loop. - Design and implement flexible network configuration patterns that give customers choice of ingress controller, load balancer, and traffic management approach - without prescribing specific Kubernetes ecosystem components or hardcoding controller-specific behaviour. - Build the observability integration enterprise customers expect: Prometheus-native metrics, ServiceMonitor CRDs for automatic scraping, Kubernetes Events emission throughout the reconciliation lifecycle, and pre-built dashboards that surface operator and cluster health without manual instrumentation. - Harden the operator's security posture across pod --- TITLE: Site Reliability Engineer (Edge Services), Infrastructure Services EMPLOYER: Apple LOCATION: Austin, United States of America (unspecified) SALARY: $132K-$245K POSTED: 2026-05-18 K401_MATCH: yes (not source-backed) APPLY_URL: https://jobs.apple.com/en-us/details/200663929/site-reliability-engineer-edge-services-infrastructure-services?team=SFTWR EXCERPT: Site Reliability Engineer (Edge Services), Infrastructure Services Austin, United States of America We are seeking a proactive Site Reliability Engineer to champion the evolution of our production ecosystems. In this role, you will help drive the vision for our visibility, moving beyond simple uptime metrics to build a sophisticated, data-driven reliability framework. You will play a pivotal role in ensuring our services are resilient, scalable, and observable, bridging the gap between complex distributed systems and seamless user experiences. As a key member of the SRE team, your mission is to treat operations as a software problem. You will focus on designing and implementing a next-generation observability and alerting strategy that prioritizes high-cardinality data and meaningful signals over noise. You will spend your time building "self-healing" systems, reducing toil through aggressive automation, and partnering with development teams to bake reliability into the CI/CD pipeline. Your goal is to move us toward a proactive stance where performance bottlenecks are identified and mitigated before they impact the customer. Minimum Qualifications: Understanding of Linux internals and deep networking expertise, including HTTP/2, HTTP/3 (QUIC), and HTTPS/TLS. You should be comfortable debugging protocol-level issues and optimizing traffic flow. Proven ability to automate repetitive tasks and complex workflows using Python or Go Experience configuring and managing modern monitoring suites (e.g., Prometheus, Grafana, ClickHouse) with a focus on creating actionable, high-signal quality alerting. Grasp of Data Structures and Algorithms (DSA) to write efficient, performant code and troubleshoot complex system bottlenecks. Practical knowledge of SLIs, SLOs, Error Budgets, Release --- TITLE: Staff Inference ML Runtime Engineer EMPLOYER: Cerebras Systems LOCATION: Sunnyvale CA or Toronto Canada (unspecified) SALARY: Not disclosed POSTED: 2025-11-25 K401_MATCH: yes (not source-backed) APPLY_URL: https://job-boards.greenhouse.io/cerebrassystems/jobs/7523546003 EXCERPT: Staff Inference ML Runtime Engineer Sunnyvale CA or Toronto Canada Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. Our novel wafer-scale architecture provides the AI compute power of dozens of GPUs on a single chip, with the programming simplicity of a single device. This approach allows Cerebras to deliver industry-leading training and inference speeds and empowers machine learning users to effortlessly run large-scale ML applications, without the hassle of managing hundreds of GPUs or TPUs. Cerebras' current customers include top model labs, global enterprises, and cutting-edge AI-native startups. OpenAI recently announced a multi-year partnership with Cerebras , to deploy 750 megawatts of scale, transforming key workloads with ultra high-speed inference. Thanks to the groundbreaking wafer-scale architecture, Cerebras Inference offers the fastest Generative AI inference solution in the world, over 10 times faster than GPU-based hyperscale cloud inference services. This order of magnitude increase in speed is transforming the user experience of AI applications, unlocking real-time iteration and increasing intelligence via additional agentic computation. About The Role The Inference ML Engineering team at Cerebras Systems is dedicated to enabling our fast generative inference solution through simple APIs powered by a distributed runtime that runs on large clusters of our own hardware. Our mission is to empower enterprises, developers, and researchers to unlock the full potential of our platform, leveraging its performance, scalability, and flexibility. The team works closely with cross-functional groups, including compiler developers, cluster orchestrators, ML scientists, cloud architects, and product teams, to deliver high-impact --- TITLE: Sr Data Engineer / Apache Cassandra Administrator EMPLOYER: PDF Solutions INC LOCATION: Dallas, Texas (unspecified) SALARY: $115K-$129K POSTED: 2026-04-20 K401_MATCH: yes (not source-backed) APPLY_URL: https://careers-pdf.icims.com/jobs/1486/login EXCERPT: Sr Data Engineer / Apache Cassandra Administrator Dallas, Texas Overview PDF Solutions is redefining the way the semiconductor industry approaches data, analytics, and experience design. As part of our journey, we're building Next gen platform - a modern, human-centered analytics platform. We believe design systems aren't just about consistency - they're about scalability , collaboration , and performance . As a Data Engineer at PDF, you will be part of a global team dedicated to leveraging innovative approaches and public cloud infrastructure to refine and enhance the design and architecture of our Big Data Analytics software. In this role, you will: Architect robust databases : Develop and maintain scalable NoSQL Apache Cassandra to support our advanced analytical solutions. Maintain scalable Database : Play a critical role in ensuring the high availability, performance, and scalability of our Cassandra clusters while supporting development teams and troubleshooting complex database issues. Contribute to scalable solutions : Design, implement, and maintain the scalable data processing components of PDF's Analytics software. Bring DevOps expertise : Automate infrastructure using modern tools like Terraform and Ansible, ensuring efficiency and reliability. This position is ideal for engineers passionate about working with cutting-edge technology to drive impactful results. Responsibilities Participate in the continuous efforts to improve the design and architecture of our application Install, configure, upgrade, and maintain Apache Cassandra clusters (on-premises and/or cloud-based) Monitor database health and performance using tools such as nodetool , JMX , Prometheus , Grafana , Medusa Perform regular backup and restore operations using native --- TITLE: AI Hardware Systems Engineer, Annapurna Labs, Trainium Machine Learning Fleet Operations EMPLOYER: Amazon LOCATION: Austin, Texas, USA (unspecified) SALARY: $136K-$184K POSTED: 2026-03-19 PARENTAL_LEAVE_WEEKS: 6 (not source-backed) NON_BIRTH_PARENT_LEAVE_WEEKS: 6 (not source-backed) K401_MATCH: yes (not source-backed) FERTILITY_FAMILY_BUILDING_BENEFITS: yes (not source-backed) ADOPTION_ASSISTANCE_OFFERED: yes (not source-backed) SURROGACY_ASSISTANCE_OFFERED: yes (not source-backed) MENTAL_HEALTH_SUPPORT: yes (not source-backed) CHILDCARE_SUBSIDY: yes (not source-backed) APPLY_URL: https://www.amazon.jobs/en/jobs/10370980/ai-hardware-systems-engineer-annapurna-labs-trainium-machine-learning-fleet-operations EXCERPT: AI Hardware Systems Engineer, Annapurna Labs, Trainium Machine Learning Fleet Operations Austin, Texas, USA Annapurna Labs designs silicon and software that accelerates innovation. Customers choose us to create cloud solutions that solve challenges that were unimaginable a short time ago-even yesterday. Our custom chips, accelerators, and software stacks enable us to take on technical challenges that have never been seen before, and deliver results that help our customers change the world. In Annapurna Labs we are at the forefront of hardware/software co-design not just in Amazon Web Services (AWS) but across the industry. The Machine Learning Acceleration Fleet Operations Team is looking for candidates interested in diving deep into our fleet of ML servers deployed around the world. We are seeking an engineer who is comfortable debugging emergent problems in GPU and server hardware, writing scripts in languages such as Python or Bash, running large scale experiments on a fleet of complex hardware, developing data infrastructure and analyzing trends, and developing automation software to scale operations. Our team has end to end ownership of some of the most advanced server hardware in the world. We drive technical debug efforts and write truly massive scale autonomous software to monitor, optimize, and remediate machine learning hardware. Come join us! Key job responsibilities - Member of a team responsible for system remediation, operational excellence, and customer experience on bleeding edge ML products - Utilize data to root cause hardware failures and identify live trends on the most complex systems in AWS - Implement --- TITLE: Platform Site Reliability Engineer EMPLOYER: Radiant LOCATION: Gloucestershire, United Kingdom (unspecified) SALARY: Not disclosed POSTED: 2026-05-06 APPLY_URL: https://jobs.ashbyhq.com/radiant/084e5679-435e-4bd4-9a36-a17f5d47d574 EXCERPT: Platform Site Reliability Engineer Gloucestershire, United Kingdom About Radiant Radiant is redefining how AI infrastructure is built. We design and operate AI-native cloud platforms engineered for sovereignty, performance, and scale. Our infrastructure powers GPU-native workloads, multi-tenant control planes, and high-performance AI systems designed for the most demanding environments. We are not building a generic cloud. We are building purpose-built AI infrastructure - from powered land, to compute, to software . As we scale our platform and expand our engineering organisation, we are looking for leaders who can build strong teams, uphold high standards, and deliver reliably at pace. Role Responsibilities - Deploy and Manage Kubernetes Clusters, deployed at scale to support AI centric workloads, across both our bare metal clusters and via trusted partner infrastructure - Develop Kubernetes Manifests and Operators: Facilitate application deployments and maintain Kubernetes-native services for networking, storage, security, identity and infrastructure management - Optimize Linux system configuration including kernel, driver, filesystem and services to support workloads running via our orchestration layer - Build and maintain automation scripts and infrastructure as code to support platform lifecycle, as well as simplifying troubleshooting for Incident resolution and provision of tooling for our support organisation - Apply ITSM frameworks: Incident, Major Incident, Change Management, and service improvement. - Maintain and enhance Radiant's observability stack: Prometheus, Grafana, and custom monitoring integrations - Operate and support services in 24x7 production environments, including on-call rotation - Contribute to Incident postmortem analyses, root cause analysis, document learnings, and automate remediations - Mentor junior engineers --- TITLE: Software Engineer - ML Infrastructure EMPLOYER: Genesis Molecular AI LOCATION: San Mateo, CA, Remote, New York, NY (remote) SALARY: Not disclosed POSTED: 2025-11-24 APPLY_URL: https://jobs.ashbyhq.com/genesis-molecular-ai/388e7964-e9a3-4620-89ae-ff21fe6444cc EXCERPT: Software Engineer - ML Infrastructure San Mateo, CA, Remote, New York, NY About the Team We're a tight-knit team of proven drug hunters, deep learning researchers, and software engineers united by a common mission - drive AI innovation in biochemistry, discovering and developing groundbreaking therapies for patients suffering from severe disorders. Genesis AI team is focused on developing foundation models for small molecule drug discovery by conducting fundamental research at the intersection of machine learning, physics, and computational chemistry, as well as engineering robust software systems that enable running large scale simulations and training generative and predictive AI models designed to learn from all kinds of molecular data, leveraging our cluster with 1000s of GPUs and 10,000s of CPUs. About the Role We're seeking experienced ML infrastructure engineers to join the team and lead engineering efforts focused on driving forward our ML research agenda for generative modeling of molecular systems, which is instrumental to our mission. As an engineer at Genesis, you will lead rapid iteration on our AI platform and infrastructure, unlocking the next level of performance, efficiency, and scale that was not previously possible. You will build massively distributed training and inference pipelines, core MLOps tools and frameworks, and optimize GPU operations to speed up ML models. Genesis is a highly-collaborative and cross functional environment, and you will work in close partnership with our exceptional engineers, researchers, and scientists. You Will - Lead engineering efforts focused on continuous improvement of the AI platform, focused on rapid build out --- TITLE: CI/CD Automation Software Engineer EMPLOYER: Viasat INC LOCATION: Duluth, Georgia (unspecified) SALARY: $117K-$185K POSTED: 2026-05-01 K401_MATCH: yes (not source-backed) APPLY_URL: https://careers-viasat.icims.com/jobs/6122/login EXCERPT: CI/CD Automation Software Engineer Duluth, Georgia About us One team. Global challenges. Infinite opportunities. At Viasat, we're on a mission to deliver connections with the capacity to change the world. For more than 35 years, Viasat has helped shape how consumers, businesses, governments and militaries around the globe communicate. We're looking for people who think big, act fearlessly, and create an inclusive environment that drives positive impact to join our team. What you'll do We are seeking a lead CI/CD Automation Software Engineer to join our team in delivering a high quality and robust Monitor and Control system for Antenna Ground Stations! In this role, you will be responsible for architecting, implementing, and maintaining CI/CD pipelines using tools such as Jenkins and GitHub Actions for automated deployment, testing, and system provisioning. You will develop and coordinate infrastructure-as-code through Ansible, Terraform, and related tools, and build automated test environments that simulate production-like systems for performance and scale testing. You will collaborate with cross-functional teams to diagnose system issues, work to improve system reliability, and deploy updates. Your role will also involve developing and maintaining tooling and scripts in Python and Bash, and supporting observability with tools like Prometheus, Grafana, and ELK. The day-to-day Architect, implement, and maintain CI/CD pipelines (Jenkins, GitHub Actions) for automated deployment, testing, and system provisioning Capable of leading designing and developing complex CI test automation pipelines, end-to-end. Build and manage infrastructure-as-code using Ansible, Terraform, or similar tools Create and maintain automated Linux test environments, simulating production-like systems --- TITLE: Senior ML Software Engineer - Integration & Quality EMPLOYER: Cerebras Systems LOCATION: Sunnyvale CA or Toronto Canada (unspecified) SALARY: Not disclosed POSTED: 2026-02-05 K401_MATCH: yes (not source-backed) APPLY_URL: https://job-boards.greenhouse.io/cerebrassystems/jobs/7620384003 EXCERPT: Senior ML Software Engineer - Integration & Quality Sunnyvale CA or Toronto Canada Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. Our novel wafer-scale architecture provides the AI compute power of dozens of GPUs on a single chip, with the programming simplicity of a single device. This approach allows Cerebras to deliver industry-leading training and inference speeds and empowers machine learning users to effortlessly run large-scale ML applications, without the hassle of managing hundreds of GPUs or TPUs. Cerebras' current customers include top model labs, global enterprises, and cutting-edge AI-native startups. OpenAI recently announced a multi-year partnership with Cerebras , to deploy 750 megawatts of scale, transforming key workloads with ultra high-speed inference. Thanks to the groundbreaking wafer-scale architecture, Cerebras Inference offers the fastest Generative AI inference solution in the world, over 10 times faster than GPU-based hyperscale cloud inference services. This order of magnitude increase in speed is transforming the user experience of AI applications, unlocking real-time iteration and increasing intelligence via additional agentic computation. About the Role We are looking for a Software Engineer to join the ML Integration and Quality team at Cerebras. This team sits at the intersection of machine learning infrastructure, distributed systems, and hardware/software co-design. In this role, you will help integrate and validate the software stack that powers the Cerebras AI platform, ensuring large-scale ML workloads run reliably and efficiently across our systems. You will work closely with engineers across runtime, compiler, kernel, and hardware teams to debug --- TITLE: Senior Backend Engineer - Transfers Gateway EMPLOYER: SumUp LOCATION: Vilnius, Lithuania (unspecified) SALARY: EUR 4K-7K/mo POSTED: 2026-01-23 K401_MATCH: yes (not source-backed) LEARNING_BUDGET_OFFERED: yes (not source-backed) APPLY_URL: https://sumup.com/careers/positions/8383464002?gh_jid=8383464002 EXCERPT: Senior Backend Engineer - Transfers Gateway Vilnius, Lithuania At SumUp, the Global Bank tribe builds the core infrastructure and services that give merchants a digital business account, helping millions of small businesses manage their banking needs easily and reliably. The Transfers Gateway squad connects SumUp's internal systems to external financial networks across the EU and the UK. We build and run the services that move money through payment schemes every day. The team owns end‑to‑end scheme integrations and operates production services that process real money at scale. You'll help modernise existing systems while balancing reliability, correctness, and change, making architectural decisions that directly affect uptime and merchant experience across multiple markets. Our tech stack includes Go, AWS, Kafka, PostgreSQL, and Kubernetes, supported by a strong observability toolchain with Prometheus, Grafana, and Honeycomb. We also actively use AI‑assisted development tools such as Cursor, GitHub Copilot, and others. SumUp's engineering culture is built around ownership. We follow a “build it, run it, improve it” model, with strong observability and a focus on learning. Engineers have clear scope and autonomy, with decisions documented, reviewed, and made collaboratively. The Vilnius team is a core part of the Global Bank domain. We own production systems, shape technical direction, and work closely with colleagues across SumUp's global offices on critical merchant banking capabilities. Join our Vilnius Team today ( Vilnius office ). What You'll Do - Own high availability money movement production services end‑to‑end, including delivery, observability, maintenance, and operational support. - Lead small‑to‑mid sized initiatives --- TITLE: Senior Scientist - Lentiviral Vectors EMPLOYER: Revvity LOCATION: Cambridge (UK) (unspecified) SALARY: Not disclosed POSTED: 2026-06-03 APPLY_URL: https://revvity.wd103.myworkdayjobs.com/External/job/Cambridge-UK/Senior-Scientist_JR-044575 EXCERPT: Senior Scientist - Lentiviral Vectors Cambridge (UK) posted: Posted 9 Days Ago --- TITLE: Sr. Embedded Engineer - End Devices EMPLOYER: Hubble Network LOCATION: Seattle, Washington, United States (unspecified) SALARY: $126K-$249K POSTED: 2026-06-07 APPLY_URL: https://hubble.com/careers?gh_jid=4738770008 EXCERPT: Sr. Embedded Engineer - End Devices Seattle, Washington, United States Hubble Network was founded with the intention of delivering on the promise of what Internet-of-Things (IoT) was supposed to be. We're building a global Bluetooth® network dedicated to machine-to-machine connectivity. We differentiate ourselves as the first modem-less and gateway-less, direct-to-satellite network from off-the-shelf Bluetooth® Low Energy chips. Hubble is ideal for applications in logistics, AgTech, and maritime where economies of scale for volume consumer and enterprise asset tracking is a priority. Our goal is to be the first billion-endpoint-connected network in the world. Hubble is an early-stage, venture-backed startup supported by some of the best investors in the world. In their previous lives, the founding team has been successful in raising $100s of millions in venture funding, developing the Amazon Sidewalk network, launching billions of dollars of space assets, and leading their teams to successful exits, both through acquisition and IPO. We are now looking to bring on talented team members who are the best at what they do to help us make Hubble a reality for the world. Hubble is building a global Bluetooth® Low Energy + satellite network that connects devices from anywhere on Earth. As we scale, we're seeking a mission-driven Embedded Engineer - End Devices to lead integration efforts with enterprise customers. You'll serve as the technical bridge between Hubble's SDK and our partners' diverse hardware platforms-enabling rapid onboarding, resolving device-side integration issues, and accelerating time-to-value. This is a hands-on, customer-facing engineering role with high visibility --- TITLE: Full stack Lead EMPLOYER: Flextock LOCATION: Almazah, Cairo Governorate, Egypt (unspecified) SALARY: Not disclosed POSTED: 2026-06-10 APPLY_URL: https://jobs.smartrecruiters.com/Flextock/743999846441451-full-stack-lead EXCERPT: Full stack Lead Almazah, Cairo Governorate, Egypt Company Description: Flextock is a YC-backed company focused to power the next generation of commerce in the region by offering e-commerce merchant the ability to scale their online businesses on demand. For local online merchants, we leverage technology to efficiently fulfill and deliver their orders through our network of last mile providers. Flextock is a purpose driven company on a mission to enable more than 1 million merchants in Africa and the Middle East to sell online without carrying out the hassle of running their own operations. Job Description: Technical Requirements: Full stack Tech Lead with skills and experience with Python, React, JavaScript, TypeScript, Perl, Oracle, SQL, MySQL, Apache Tomcat, Maven, XML, XSLT, JSON, RESTful APIs, etc. Analyzing customer requirements Ability to understand client requirements as well as underlying infrastructure applications, systems and processes. Ability to oversee development efforts. Strong capability in juggling priorities so that deadlines are met while retaining consistently high quality outcomes. Software architecture design, together with architecture team Technical knowledge of MS Project Server, Report Builder and SharePoint Optimize applications for maximum speed and scalability Creating technical specifications, writing program code and documenting Testing the products in controlled situations before going live Maintaining the systems once they are up and running Preparation of training manuals (user guides) for users Experience with systems management tools as Nagios, Grafana, Prometheus, Rundeck are a plus. Experience of working in infrastructure is a plus Experience in Automation, and Orchestration to drive efficiencies within --- TITLE: Senior DevSecOps Engineer EMPLOYER: Goodyear TIRE & Rubber LOCATION: IN Hyderabad (unspecified) SALARY: Not disclosed POSTED: 2026-05-12 K401_MATCH: yes (not source-backed) APPLY_URL: https://goodyear.wd1.myworkdayjobs.com/GoodyearCareers/job/IN-Hyderabad/Senior-DevSecOps-Engineer_JR-40107354 EXCERPT: Senior DevSecOps Engineer IN Hyderabad Job Description Job Description We are looking for an experienced AWS DevOps / DevSecOps Engineer to join our cloud engineering team in India. In this role, you will be responsible for operating and continuously improving our AWS‑based cloud platform Fleethub, with a strong focus on automation, CI/CD, reliability, and security. You will work closely with AWS Cloud Architects, Software Engineers, and Security teams to ensure scalable, secure, and highly available cloud environments that support modern application development and rapid delivery. Key Responsibilities: In close collaboration with Cloud Architects and AWS Developers, you will: Design, build, and maintain CI/CD pipelines to automate build, test, and deployment processes Manage and evolve Infrastructure as Code (IaC) using tools such as Terraform, CloudFormation, or AWS CDK Operate and optimize AWS cloud infrastructure (compute, networking, storage, containers) Monitor system health, performance, and reliability using tools like Prometheus, Grafana, and CloudWatch Implement DevSecOps practices, including security scanning, secrets management, and compliance controls Troubleshoot production issues and participate in incident response and root‑cause analysis Improve system availability, scalability, and cost efficiency Maintain architecture diagrams, technical documentation, and operational runbooks Collaborate with global teams across time zones and contribute to continuous improvement initiatives Required Skills & Experience: Strong hands‑on experience with AWS services (e.g., VPC, EC2, AM, RDS, SQS, S3, SES, Route53) Proven expertise in CI/CD tools such as Jenkins, SonarQube, GitLab CI/CD, GitHub Actions, or similar (Jenkins is a must) Solid experience with container orchestration : AWS EKS and Kubernetes Hands‑on experience --- TITLE: Senior Software Engineer, Robotics EMPLOYER: Divergent 3D LOCATION: Torrance, California, United States (unspecified) SALARY: $141K-$224K POSTED: 2026-04-01 APPLY_URL: https://job-boards.greenhouse.io/divergent/jobs/5173190008 EXCERPT: Senior Software Engineer, Robotics Torrance, California, United States Divergent is a technology company that has architected, invented, built, and commercialized an end-to-end factory system called the Divergent Adaptive Production System (DAPS) that comprehensively uses machine learning to optimally engineer, additively manufacture, and flexibly assemble complex integrated vehicle structures and subsystems. Products created using DAPS are superior in performance, lower in cost, rapidly customizable to meet mission and customer-specific requirements, faster to market, and scalable on demand to high volume production. Divergent is a qualified Tier 1 supplier to global automotive OEMs, and Divergent is now expanding to support mission critical needs in the Aerospace and Defense sector. Join us to be a part of this transformative journey, where your impact will shape the future of technology and production. Purpose We are seeking a talented and motivated Software Developer to join our Robotics & AI Software team. In this role, you will have the opportunity to work closely with experienced automation and software engineers to develop and implement innovative algorithms for task and motion planning plus computer vision in robotic systems. As a key member of our team, you will contribute to the design, development, and testing of software solutions aimed at enhancing the efficiency and reliability of robotic software algorithms. This involves symbolic task planning, robot kinematics, collision detection, optimization, and computer vision. The Role - Develop and maintain software development best practices to ensure high-quality code, efficient testing, and timely releases. - Collaborate with other teams to ensure that --- TITLE: Member of Technical Staff, Multimodal Reasoning - Applied Science , AGI Autonomy EMPLOYER: Amazon LOCATION: San Francisco, California, USA (unspecified) SALARY: $255K-$345K POSTED: 2026-02-02 PARENTAL_LEAVE_WEEKS: 6 (not source-backed) NON_BIRTH_PARENT_LEAVE_WEEKS: 6 (not source-backed) K401_MATCH: yes (not source-backed) FERTILITY_FAMILY_BUILDING_BENEFITS: yes (not source-backed) ADOPTION_ASSISTANCE_OFFERED: yes (not source-backed) MENTAL_HEALTH_SUPPORT: yes (not source-backed) CHILDCARE_SUBSIDY: yes (not source-backed) APPLY_URL: https://www.amazon.jobs/en/jobs/3172291/member-of-technical-staff-multimodal-reasoning-applied-science-agi-autonomy EXCERPT: Member of Technical Staff, Multimodal Reasoning - Applied Science , AGI Autonomy San Francisco, California, USA Amazon has launched a new research lab in San Francisco to develop foundational capabilities for useful AI agents. We're enabling practical AI to make our customers more productive, empowered, and fulfilled. Our work leverages large vision language models (VLMs) with reinforcement learning (RL) and world modeling to solve perception, reasoning, and planning to build useful enterprise agents. Our lab is a small, talent-dense team with the resources and scale of Amazon. Each team in the lab has the autonomy to move fast and the long-term commitment to pursue high-risk, high-payoff research. We're entering an exciting new era where agents can redefine what AI makes possible. Key job responsibilities You will lead our efforts to improve the multi-model perception and reasoning abilities of our AI agent in an applied research role. Responsibilities including model training, dataset design, and pre- and post-training optimization. You will be hired as a Member of Technical Staff. Basic Qualifications: - 5+ years' experience building machine learning models - PhD or Master's degree in computer science or related field - Proficiency in Python, Java, C++, or related language - Experience with deep learning methods and tools, e.g., PyTorch, JAX Preferred Qualifications: - Background in scientific research with a proven ability to generate and implement new ideas in machine learning - Experience with post-training of large Vision Language Models (VLMs). - Willingness to step outside typical role boundaries to get things done --- TITLE: Developer Experience Engineer EMPLOYER: Etched LOCATION: San Jose, CA, United States (unspecified) SALARY: Not disclosed POSTED: 2026-02-19 APPLY_URL: https://jobs.ashbyhq.com/etched/3f83fd5f-5e50-403a-9a03-7ddb56502f49 EXCERPT: Developer Experience Engineer San Jose, CA, United States About Etched Etched is building the world's first AI inference system purpose-built for transformers - delivering over 10x higher performance and dramatically lower cost and latency than a B200. With Etched ASICs, you can build products that would be impossible with GPUs, like real-time video generation models and extremely deep & parallel chain-of-thought reasoning agents. Backed by hundreds of millions from top-tier investors and staffed by leading engineers, Etched is redefining the infrastructure layer for the fastest growing industry in history. Job Summary We are looking for a Developer Experience Enginee r to enhance developer productivity, automation, and infrastructure across our hardware and software teams. You will work at the intersection of DevOps, software engineering, and high-performance computing (HPC), building systems that accelerate chip design, simulation, and AI model deployment in a cloud and on-prem hybrid environment. Key responsibilities - Develop and maintain automation tools to streamline development, testing, and deployment workflows. - Optimize and manage Slurm-based job scheduling for AI workloads, simulation, and chip design workflows. - Build observability solutions using Grafana, Prometheus, and OpenTelemetry for monitoring pipelines, infrastructure, and compute clusters. - Manage and optimize containerized environments using Docker and Kubernetes to enhance scalability and reproducibility. - Enhance build, test, and deployment pipelines with CI/CD tools like GitHub Actions, Jenkins, Buildkite, or Bazel. - Develop caching and artifact management systems to reduce build times and improve dependency resolution. - Integrate and manage cloud resources (AWS, GCP) for scaling compute, storage, --- TITLE: Research Scientist – LTX Model Evaluation EMPLOYER: Lightricks LOCATION: Jerusalem (unspecified) SALARY: Not disclosed POSTED: 2026-01-27 APPLY_URL: https://careers.lightricks.com/position?gh_jid=8392230002&gh_jid=8392230002 EXCERPT: Research Scientist – LTX Model Evaluation Jerusalem Who we are Lightricks is an AI-first company creating next-generation content creation technology for businesses, enterprises, and studios with a mission to bridge the gap between imagination and creation. At our core is LTX-2, an open-source generative video model, built to deliver expressive, high-fidelity video at unmatched speed. It powers both our own products and a growing ecosystem of partners through API integration. The company is also known globally for pioneering consumer creativity through products like Facetune, one of the world's most recognized creative brands, which helped introduce AI-powered visual expression to hundreds of millions of users worldwide. We combine deep research, user-first design, and end-to-end execution from concept to final render to bring the future of expression to all. The Team Lightricks is an AI-first company creating next-generation content-creation technology for businesses, enterprises, and studios, with a mission to bridge the gap between imagination and creation. At our core is LTX-2, an open-source generative video model, built to deliver expressive, high-fidelity video at unmatched speed. It powers both our own products and a growing ecosystem of partners through API integration. Following the success of LTX-2, our widely adopted open-source text-to-audio+video model, we are expanding our efforts to develop cutting-edge audio+video generation models and are hiring Research Scientists to join our Model Evaluation team - part of LTX Foundational Model group. The Model Evaluation team is the central nervous system of the LTX Foundation Model group. We don't just measure performance; we define --- TITLE: Senior Linux Platform Engineer - Trading Infrastructure EMPLOYER: Qube Research & Technologies LOCATION: London (unspecified) SALARY: Not disclosed POSTED: 2025-09-25 APPLY_URL: https://job-boards.greenhouse.io/quberesearchandtechnologies/jobs/8185063002 EXCERPT: Senior Linux Platform Engineer - Trading Infrastructure London Qube Research & Technologies (QRT) is a global quantitative and systematic investment manager, operating in all liquid asset classes across the world. We are a technology- and data-driven group implementing a scientific approach to investing. Combining data, research, technology, and trading expertise has shaped our collaborative mindset, which enables us to solve the most complex challenges. QRT's culture of innovation continuously drives our ambition to deliver high-quality returns for our investors. Your Future Role within QRT You will join the global infrastructure team, responsible for QRT's Linux estate, including trading infrastructure and low-latency platforms. The team works closely with application and support functions to deliver secure, high-performance, and scalable solutions for the business. - Manage and evolve QRT's Linux platform to support global trading environments - Work with application and support teams to maintain a secure and scalable infrastructure - Lead the design and implementation of automation frameworks, tools, and processes - Contribute to the development of monitoring and observability capabilities - Analyse and tune performance across systems, networks, and applications - Support build/release management processes - Collaborate within a version-controlled development environment Your Present Skillset - Strong expertise in Linux (preferably Red Hat-based distributions) - Experience building and administering large-scale or latency-sensitive platforms - Extensive experience with architecture and long-term evolution of automation and configuration management systems (e.g. Ansible, Puppet, Chef) - Experience with monitoring and observability tools (e.g. Splunk, Elastic, Prometheus, Grafana) - Experience with build/release management systems. - Solid --- TITLE: Senior DevSecOps Engineer EMPLOYER: Goodyear TIRE & Rubber LOCATION: IN Hyderabad (unspecified) SALARY: Not disclosed POSTED: 2026-05-12 K401_MATCH: yes (not source-backed) APPLY_URL: https://goodyear.wd1.myworkdayjobs.com/GoodyearCareers/job/IN-Hyderabad/AWS-DevSecOps-Engineer_JR-40107353 EXCERPT: Senior DevSecOps Engineer IN Hyderabad Job Description We are looking for an experienced AWS DevOps / DevSecOps Engineer to join our cloud engineering team in India. In this role, you will be responsible for operating and continuously improving our AWS‑based cloud platform Fleethub, with a strong focus on automation, CI/CD, reliability, and security. You will work closely with AWS Cloud Architects, Software Engineers, and Security teams to ensure scalable, secure, and highly available cloud environments that support modern application development and rapid delivery. Key Responsibilities: In close collaboration with Cloud Architects and AWS Developers, you will: Design, build, and maintain CI/CD pipelines to automate build, test, and deployment processes Manage and evolve Infrastructure as Code (IaC) using tools such as Terraform, CloudFormation, or AWS CDK Operate and optimize AWS cloud infrastructure (compute, networking, storage, containers) Monitor system health, performance, and reliability using tools like Prometheus, Grafana, and CloudWatch Implement DevSecOps practices, including security scanning, secrets management, and compliance controls Troubleshoot production issues and participate in incident response and root‑cause analysis Improve system availability, scalability, and cost efficiency Maintain architecture diagrams, technical documentation, and operational runbooks Collaborate with global teams across time zones and contribute to continuous improvement initiatives Required Skills & Experience: Strong hands‑on experience with AWS services (e.g., VPC, EC2, AM, RDS, SQS, S3, SES, Route53) Proven expertise in CI/CD tools such as Jenkins, SonarQube, GitLab CI/CD, GitHub Actions, or similar (Jenkins is a must) Solid experience with container orchestration : AWS EKS and Kubernetes Hands‑on experience with configuration --- TITLE: Senior Hardware Engineer, EEE Components EMPLOYER: Relativity LOCATION: Long Beach, California (unspecified) SALARY: $148K-$204K POSTED: 2026-05-01 APPLY_URL: https://boards.greenhouse.io/relativity/jobs/8533753002?gh_jid=8533753002 EXCERPT: Senior Hardware Engineer, EEE Components Long Beach, California At Relativity Space, we're building rockets to serve today's needs and tomorrow's breakthroughs. Our Terran R vehicle will deliver customer payloads to orbit, meeting the growing demand for launch capacity. But that's just the start. Achieving commercial success with Terran R will unlock new opportunities to advance science, exploration, and innovation, pioneering progress that reaches beyond the known. Joining Relativity means becoming part of something where autonomy, ownership, and impact exist at every level. Here, you're not just executing tasks; you're solving problems that haven't been solved before, helping develop a rocket, a factory, and a business from the ground up. Whether you're in propulsion, manufacturing, software, avionics, or a corporate function, you'll collaborate across teams, shape decisions, and see your work come to life in record time. Relativity is a place where creativity and technical rigor go hand in hand, and your voice will help define the stories we're writing together. Now is a unique moment in time where it's early enough to leave your mark on the product, the process, and the culture, but far enough along that Terran R is tangible and picking up momentum. The most meaningful work of your career is waiting. Join us. About the Team: The Avionics team is responsible for all aspects of the electrical design on the rocket, from electronics, interconnects, associated mechanical design, testing and initial production. This team's scope includes the flight computers, power systems, RF systems, PCB design, autonomous flight --- TITLE: Prognostics & Health Monitoring Engineer EMPLOYER: Cerebras Systems LOCATION: Sunnyvale, CA (unspecified) SALARY: $150K-$250K POSTED: 2026-04-30 K401_MATCH: yes (not source-backed) APPLY_URL: https://job-boards.greenhouse.io/cerebrassystems/jobs/7720309003 EXCERPT: Prognostics & Health Monitoring Engineer Sunnyvale, CA Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. Our novel wafer-scale architecture provides the AI compute power of dozens of GPUs on a single chip, with the programming simplicity of a single device. This approach allows Cerebras to deliver industry-leading training and inference speeds and empowers machine learning users to effortlessly run large-scale ML applications, without the hassle of managing hundreds of GPUs or TPUs. Cerebras' current customers include top model labs, global enterprises, and cutting-edge AI-native startups. OpenAI recently announced a multi-year partnership with Cerebras , to deploy 750 megawatts of scale, transforming key workloads with ultra high-speed inference. Thanks to the groundbreaking wafer-scale architecture, Cerebras Inference offers the fastest Generative AI inference solution in the world, over 10 times faster than GPU-based hyperscale cloud inference services. This order of magnitude increase in speed is transforming the user experience of AI applications, unlocking real-time iteration and increasing intelligence via additional agentic computation. Role Summary Quality, reliability, and uptime are foundational to scaling Cerebras systems. We are seeking an engineer to define and build our prognostics and health monitoring (PHM) capability-developing frameworks to monitor, assess, and predict hardware health across our fleet. In this role, you will transform telemetry and operational data into actionable insights and automated responses, enabling early detection of degradation, accurate failure prediction, and proactive actions to keep systems highly available, performant, and resilient. This is a highly cross-functional role spanning reliability engineering, data science, --- TITLE: Platform Database Engineer (MONGO DB) EMPLOYER: Valtech LOCATION: Colombia - Remote (remote) SALARY: Not disclosed POSTED: 2026-04-02 APPLY_URL: https://job-boards.eu.greenhouse.io/valtech/jobs/4831257101 EXCERPT: Platform Database Engineer (MONGO DB) Colombia - Remote Platform Database Engineer Location: Colombia - Remote About the Role The Platform Database Engineer is responsible for designing, deploying, administering, and optimizing MongoDB (Atlas and on‑premise) within a large-scale, cloud-based enterprise ecosystem. This role blends deep database engineering expertise with SRE principles to ensure performance, reliability, scalability, and automation across mission‑critical data platforms. What You Will Do - Architect, implement, and maintain MongoDB Atlas and non‑Atlas environments, ensuring high availability, scalability, and security. - Design and enforce resiliency and disaster recovery strategies, including backup, restore, and multi‑region failover. - Optimize database performance and query execution to support application development teams throughout all SDLC phases. - Develop and manage infrastructure‑as‑code solutions (Terraform, GitOps) for database provisioning and automation. - Implement and refine observability and monitoring solutions using Dynatrace, CloudWatch, and MongoDB native telemetry. - Manage access control, auditing, and encryption to meet enterprise security and compliance requirements. - Collaborate closely with application, DevOps, and platform teams to continuously improve reliability, performance, and operational excellence. - Lead incident response for database‑related issues, drive root cause analysis, and implement corrective and preventative measures. - Mentor and guide technical teams, contributing to documentation, standards, and design reviews. Minimum Qualifications - 4+ years of experience as a MongoDB DBA, including 4+ years administering MongoDB Atlas. - Strong experience with the AWS ecosystem (EC2, EKS, CloudWatch, IAM, KMS, VPC, Lambda). - Proficiency in Terraform, GitOps, CI/CD automation, and containerized workloads (EKS/Kubernetes). - Experience with Dynatrace, Prometheus, or equivalent --- TITLE: NOC Tech I (NOC & Cloud Operations) EMPLOYER: Five9 Inc LOCATION: Manila, Manila, Philippines (Hybrid) (hybrid) SALARY: Not disclosed POSTED: 2026-04-28 NON_BIRTH_PARENT_LEAVE_WEEKS: 12 (not source-backed) K401_MATCH: yes (not source-backed) MENTAL_HEALTH_SUPPORT: yes (not source-backed) APPLY_URL: https://www.five9.com/about/careers/job-detail?gh_jid=5983971004 EXCERPT: NOC Tech I (NOC & Cloud Operations) Manila, Manila, Philippines (Hybrid) Join us in bringing joy to customer experience. Five9 is a leading provider of cloud contact center software, bringing the power of cloud innovation to customers worldwide. Living our values everyday results in our team-first culture and enables us to innovate, grow, and thrive while enjoying the journey together. We celebrate diversity and foster an inclusive environment, empowering our employees to be their authentic selves. Who we are: Acqueon's conversational engagement software lets customer-centric brands orchestrate campaigns and proactively engage with consumers using voice, messaging, and email channels. Acqueon leverages a rich data platform, statistical and predictive models, and intelligent workflows to let enterprises maximize the potential of every customer conversation. Acqueon is trusted by 200 clients across industries to increase sales, drive proactive service, improve collections, and develop loyalty. At our core, Acqueon is a customer-centric company with a burning desire (backed by asuite of awesome, AI-powered technology) to help businesses provide friction-free, delightful, and referral-worthy customer experiences. For NOC Engineers, Acqueon offers a dynamic environment focused on 24/7 infrastructure and application monitoring, incident response, service reliability, and continuous improvement. Engineers work with enterprise-scale systems, ensuring uptime, optimizing performance, and maintaining strict SLA and compliance standards in a fast-paced SaaS ecosystem. Key Responsibilities - Ensure high availability and maximum uptime in SaaS environments by proactively monitoring and managing infrastructure, applications, and network systems using tools like Zabbix, Grafana, Prometheus, Sumo Logic, and Amazon CloudWatch. - Proactively identify anomalies, --- TITLE: Director, Avionics Hardware Development EMPLOYER: Relativity LOCATION: Long Beach, California (unspecified) SALARY: $245K-$336K POSTED: 2026-02-23 APPLY_URL: https://boards.greenhouse.io/relativity/jobs/8432770002?gh_jid=8432770002 EXCERPT: Director, Avionics Hardware Development Long Beach, California At Relativity Space, we're building rockets to serve today's needs and tomorrow's breakthroughs. Our Terran R vehicle will deliver customer payloads to orbit, meeting the growing demand for launch capacity. But that's just the start. Achieving commercial success with Terran R will unlock new opportunities to advance science, exploration, and innovation, pioneering progress that reaches beyond the known. Joining Relativity means becoming part of something where autonomy, ownership, and impact exist at every level. Here, you're not just executing tasks; you're solving problems that haven't been solved before, helping develop a rocket, a factory, and a business from the ground up. Whether you're in propulsion, manufacturing, software, avionics, or a corporate function, you'll collaborate across teams, shape decisions, and see your work come to life in record time. Relativity is a place where creativity and technical rigor go hand in hand, and your voice will help define the stories we're writing together. Now is a unique moment in time where it's early enough to leave your mark on the product, the process, and the culture, but far enough along that Terran R is tangible and picking up momentum. The most meaningful work of your career is waiting. Join us. About the Team: The Avionics team is responsible for the full lifecycle of Terran R's nervous system, designing, building, testing, installing, and operating the hardware that connects and controls every major electrical system on the vehicle and ground. The team leverages close partnership between --- TITLE: Senior Technical Program Manager, Enterprise Projects Office EMPLOYER: Voyager Technologies INC LOCATION: Remote (remote) SALARY: $140K-$175K POSTED: 2026-06-07 APPLY_URL: https://job-boards.greenhouse.io/voyagertechnologiesinc/jobs/4254574009 EXCERPT: Senior Technical Program Manager, Enterprise Projects Office Remote Voyager is an innovative defense, national security and space technology company committed to advancing and delivering transformative, mission-critical solutions. We tackle the most complex challenges to unlock new frontiers for human progress, fortify national security, and protect critical assets to lead in the race for technological and operational superiority from ground to space. Forge the Future: Join Voyager Technologies The future belongs to those who build it. At Voyager Technologies, we're building technologies that protect lives, expand frontiers and prepare us for what's next. And we're doing that with people who are wired to solve, build, adapt and lead. These roles are not for the faint of heart. You'll help lay the foundation for humanity's future. Join a culture where innovation thrives, curiosity is rewarded, and impact is real. We're a company of doers, thinkers and builders, united by purpose and grounded in reality. If you want to put your skills to work where the stakes are real and the mission is bigger than any one person, forge the future with Voyager. The Senior Program Manager, IT Program Management Office is responsible for leading Voyager's enterprise IT program portfolio, driving project management, designing the IT roadmaps and resourcing plans, shaping data strategy and architecture, and driving governance, compliance, and cross-company technology delivery at scale. This position reports to the the Senior Director, Enterprise Projects Office and is part of the administrative organization and this role, the essential functions are: • Lead the --- TITLE: Platform Database Engineer EMPLOYER: Valtech LOCATION: Argentina - Remote (remote) SALARY: Not disclosed POSTED: 2026-04-02 APPLY_URL: https://job-boards.eu.greenhouse.io/valtech/jobs/4831263101 EXCERPT: Platform Database Engineer Argentina - Remote Platform Database Engineer Location: Colombia - Remote About the Role The Platform Database Engineer is responsible for designing, deploying, administering, and optimizing MongoDB (Atlas and on‑premise) within a large-scale, cloud-based enterprise ecosystem. This role blends deep database engineering expertise with SRE principles to ensure performance, reliability, scalability, and automation across mission‑critical data platforms. What You Will Do - Architect, implement, and maintain MongoDB Atlas and non‑Atlas environments, ensuring high availability, scalability, and security. - Design and enforce resiliency and disaster recovery strategies, including backup, restore, and multi‑region failover. - Optimize database performance and query execution to support application development teams throughout all SDLC phases. - Develop and manage infrastructure‑as‑code solutions (Terraform, GitOps) for database provisioning and automation. - Implement and refine observability and monitoring solutions using Dynatrace, CloudWatch, and MongoDB native telemetry. - Manage access control, auditing, and encryption to meet enterprise security and compliance requirements. - Collaborate closely with application, DevOps, and platform teams to continuously improve reliability, performance, and operational excellence. - Lead incident response for database‑related issues, drive root cause analysis, and implement corrective and preventative measures. - Mentor and guide technical teams, contributing to documentation, standards, and design reviews. Minimum Qualifications - 4+ years of experience as a MongoDB DBA, including 4+ years administering MongoDB Atlas. - Strong experience with the AWS ecosystem (EC2, EKS, CloudWatch, IAM, KMS, VPC, Lambda). - Proficiency in Terraform, GitOps, CI/CD automation, and containerized workloads (EKS/Kubernetes). - Experience with Dynatrace, Prometheus, or equivalent observability platforms. --- TITLE: HPC Performance Engineer EMPLOYER: CoreWeave LOCATION: New York, NY / Sunnyvale, CA / Bellevue, WA (unspecified) SALARY: $165K-$242K POSTED: 2025-09-15 K401_MATCH: yes (not source-backed) MENTAL_HEALTH_SUPPORT: yes (not source-backed) CHILDCARE_SUBSIDY: yes (not source-backed) LEARNING_BUDGET_OFFERED: yes (not source-backed) APPLY_URL: https://coreweave.com/careers/job?4601657006&board=coreweave&gh_jid=4601657006 EXCERPT: HPC Performance Engineer New York, NY / Sunnyvale, CA / Bellevue, WA CoreWeave is The Essential Cloud for AI™. Built for pioneers by pioneers, CoreWeave delivers a platform of technology, tools, and teams that enables innovators to build and scale AI with confidence. Trusted by leading AI labs, startups, and global enterprises, CoreWeave combines superior infrastructure performance with deep technical expertise to accelerate breakthroughs and turn compute into capability. Founded in 2017, CoreWeave became a publicly traded company (Nasdaq: CRWV) in March 2025. Learn more at www.coreweave.com . What you'll do: CoreWeave is seeking a highly skilled and motivated HPC Performance Engineer to join our HAVOCK Team, reporting into the Manager of Systems Engineering. In this role, you will play a crucial part in the design, development, and optimization of our bare-metal systems from POST through joining a Kubernetes cluster. The team's primary responsibilities include maintaining a custom Linux kernel, various OS images (Ubuntu-based), the virtualization stack (kubevirt/qemu/vfio), and the container/pod runtime stack (containerd/nydus/kubelet). You will collaborate closely with cross-functional teams, up stack engineering teams, and stakeholders to ensure our low-level software stack is performant in the context of hardware updates; and providing data, metrics, dashboards, and analysis to substantiate performance assertions. Kernel H ardware - A cceleration - V irtualization - O perating Systems - C ontainerization - K ubelet Our Team's Stack: - Python, Go, bash/sh, C - Prometheus, Victoria Metrics, Grafana - Linux Kernel (custom build), Ubuntu - Intel/AMD/ARM CPUs, Nvidia GPUs, DPUs, Infiniband and Ethernet NICs --- [PASTE YOUR RESUME OR SKILLS HERE]