Remora Jobs
Find JobsCompanies

Filters

Remote

Pay

Job Type

Date Posted

Sort By

Infrastructure Engineer Jobs

  1. Home
  2. /
  3. Infrastructure Engineer Jobs

446 jobs found

Hadrian

Infrastructure Engineer

Hadrian

United States

Are you obsessed withbuilding cloud-native platformsthat enable engineers to operate at internet scale? Then join us as our newInfrastructure Engineerand be part of defining the next generation of the Infrastructure platform atHadrian! Who are we: We are Hadrian - an offensive cybersecurity startup. We are reshaping the future of cybersecurity through the power of agentic AI injected with the hacker’s perspective. Founded in August 2021, we've secured funding and the backing of ABN AMRO Ventures and Cherry Ventures, propelling us into a new era. Our goal is to provide companies with an autonomous Threat Exposure Management platform. At Hadrian, we view security through a hacker's eyes because hackers understand hackers best. We continuously map the digital footprint of organisations, discover risks, and prioritise remediation for security teams to harden their external attack surfaces. Located in Amsterdam's buzzing Leidseplein, by London's famous Paddington Station, and our most recent US HQ in the heart of New York City, our diverse team from 25+ countries is on a mission to shake things up in offensive security. Join us in making waves and securing over 700 businesses, and be at the forefront of automated offensive security. Your role: As an Infrastructure Engineer at Hadrian, you will design, build, and evolve the cloud-native platform that powers our offensive security operations. You will enable engineering teams to move faster, deploy safely, and operate reliably by creating scalable self-service infrastructure and resilient platform capabilities. This role is ideal for engineers who enjoy combining deep infrastructure expertise with a strong software engineering mindset. We're looking for someone who thrives on automation, reliability, and platform design, and who sees infrastructure as a product that empowers others to do their best work. If you enjoy solving complex distributed systems challenges, building developer-friendly platforms, and creating the foundations that allow autonomous security systems to operate at scale, we would love to hear from you What you will do as an Infrastructure Engineer at Hadrian: Help build and maintain the cloud-native platform that powers Hadrian's offensive security systems. Design and automate infrastructure delivery using Infrastructure-as-Code, enabling scalable and self-service platform capabilities. Manage and improve our Kubernetes environment, ensuring reliability, performance, and security. Collaborate closely with software engineers to improve deployment workflows, developer experience, and platform adoption. Operate and optimize our core cloud infrastructure on AWS, while supporting our distributed scanning environments. Improve observability, monitoring, incident response, and platform reliability across critical systems. Help implement platform standards, security controls, and engineering best practices that enable teams to move quickly and safely. Take on a unique challenge. We’re building autonomous security systems that continuously analyze internet-scale attack surfaces. You'll help create the platform that makes this possible. You are fit for this job because you: Are fluent in English. Have 3-5 years of experience as an Infrastructure Engineer, Platform Engineer, DevOps Engineer, SRE, or similar role. Have hands-on experience with Kubernetes, AWS, and Terraform (/OpenTofu), and are comfortable operating cloud-native infrastructure in production environments. Are comfortable managing core environments on AWS while extending workloads into secondary cloud environments used for scanning. Have a solid foundation in systems engineering, containerization (Docker), and distributed systems. Are experienced with automation and CI/CD tooling, such as Terraform, Ansible, Jenkins, or GitLab CI. Enjoy building platforms, tooling, and automation that help engineering teams move faster and more safely. Possess strong troubleshooting and problem-solving skills, and enjoy solving reliability, scalability, and operational challenges. Use AI pragmatically to accelerate workflows while maintaining high standards of reliability and security. Communicate clearly and collaborate effectively across technical teams. Bonus points if you have: Experience with KNative, event-driven architectures, or service mesh technologies. Experience with Kafka or other distributed messaging systems. Hands-on experience with observability and monitoring tooling. Familiarity with Site Reliability Engineering (SRE) practices. Certifications related to Kubernetes such as CKA, CKAD, or AWS certifications. Contributions to open-source infrastructure or cloud-native projects. Our Stack: GoLang / Python KNative Serving & Eventing / Kafka / Redis Docker / Kubernetes / Helm Charts Terraform / OpenTofu PostgreSQL / BigQuery AWS / GCP / SCW (plus) Benefits at Hadrian: Permanent contract Unlimited paid holiday Stock options package Mobile phone stipend Work-from-home budget Business travel card Swapfiets Referral bonuses Why you want to be part of Hadrian: Make the internet safer: We enable companies to protect their customers, employees, and other stakeholders from malicious hackers by providing real-time Exposure Management from the Hacker’s perspective Board a primed rocket ship: Hadrian is growing fast; the team is approaching 90 wonderful people, notable customers are signed, and significant funding is raised. Now is the time to join our high-growth journey Be part of a strong and dynamic culture: Our values define who we are as a group and serve as the cornerstones to guide us in building a product that our customers love. At Hadrian, we achieve this by championing a culture that takes ownership, succeeds together, increases velocity, and pursues growth Sounds like the role for you? Apply now! Our clients are seeking a world-class and reliable digital security service. That is exactly what we are offering. Therefore, a background check will be part of the process. Originally posted on Himalayas

On-siteFull-time2w agoApply
Apex Systems

Infrastructure Engineer

Apex Systems

Costa Mesa, CA

Job#: 3047225 Job Description: Infrastructure Engineer (Systems & Networks) Location: Costa Mesa, California (Onsite) Role Overview: We are seeking a Systems and Network Engineer to join a dynamic IT team. This role is responsible for high-level, end-to-end IT support across the organization, from desktop support to enterprise VoIP, LAN/WAN/Wi-Fi networking, hybrid cloud infrastructure, and security operations. This is an in-office position that requires on-call availability to support global

On-siteFull-time$1.0L – $1.2L3w agoApply
Simplex Trading

Infrastructure Engineer

Simplex Trading

Chicago, IL

As an Infrastructure Engineer, you will develop and maintain systems, triage issues, and manage Linux environments, working both independently and as part of a team. Top Skills: Ansible, Bash, Ipmi, Linux, Nagios, Perl, Python, Snmp Industry: Fintech , Software

On-siteContract$1.3L – $2.0LNaNy agoApply
People Culture Talent

Infrastructure Engineer

People Culture Talent

San Francisco, CAMontreal, QC, Seattle, WA, Chicago, IL, Atlanta, GA, Los Angeles, CA, New York, NY

PCT partners with venture-backed technology companies looking for Infrastructure Engineers to build the systems, platforms, and tooling that enable engineering teams and products to operate reliably at scale. Infrastructure opportunities across our portfolio range from early-stage companies building their foundational systems to later-stage organizations tackling significant challenges around scale, reliability, performance, developer productivity, and cloud infrastructure. What You Might Work On Design, build, and operate scalable cloud and production infrastructure. Improve the reliability, availability, performance, and efficiency of production systems. Build internal platforms and tooling that improve developer productivity and deployment velocity. Develop and maintain CI/CD, observability, monitoring, and incident-response systems. Automate infrastructure provisioning, deployment, and operational workflows. Partner with application, security, data, and product engineering teams on infrastructure requirements. Identify and eliminate bottlenecks as products and engineering organizations scale. Help establish infrastructure architecture, standards, and operational best practices. What We`re Looking For While requirements vary across our client companies, strong candidates often bring: Experience building and operating production infrastructure at meaningful scale. Strong knowledge of modern cloud environments and infrastructure technologies. Experience with infrastructure-as-code, containers, orchestration, CI/CD, observability, or related tooling. Strong software engineering and automation skills. An understanding of distributed systems, networking, reliability, and production operations. A pragmatic approach to balancing reliability, scalability, developer experience, and cost. Strong ownership and an interest in solving foundational technical problems. When you apply, tell us about the infrastructure challenges you`ve worked on, the technologies and environments you`re strongest in, and the types of companies or technical problems you`re interested in next. We`ll use that information to identify opportunities across our client portfolio that may be a strong fit. About PCT People Culture Talent helps VCs and venture-backed startups build thriving teams through exceptional Human Capital Management. We create people-first talent strategies for high-growth companies, partnering with organizations from pre-seed to IPO to transform their growth into a masterclass in scaling with purpose. About Our Client Companies PCT partners with exceptional venture-backed companies to build resilient, people-centric, diverse teams. Our clients span industries and stages, from pre-seed startups to companies preparing for IPO, including leading tech organizations like Notion, Lyft, Modern Health, Instacart, Uber, Google, Render, and GitHub . Our clients operate across a wide range of technical environments and stages of scale, creating opportunities to build foundational infrastructure as well as solve complex systems challenges at growing organizations. These companies share a commitment to building strong culture, high-performing teams, and innovative people programs. They`re growing fast, scaling thoughtfully, and seeking exceptional talent to join their journey.

On-siteFull-time1w agoApply

Infrastructure Engineer

Elicit

USA

About Elicit Elicit is an AI research assistant that uses language models to help researchers figure out what’s true and make better decisions, starting with common research tasks like literature review. What we're aiming for: Elicit radically increases the amount of good reasoning in the world. For experts, Elicit pushes the frontier forward. For non-experts, Elicit makes good reasoning more affordable. People who don't have the tools, expertise, time, or mental energy to make well-reasoned decisions on their own can do so with Elicit. Elicit is a scalable ML system based on human-understandable task decompositions, with supervision of process, not outcomes. This expands our collective understanding of safe AGI architectures. Visit our Twitter to learn more about how Elicit is helping researchers and making progress on our mission. Why we're hiring for this role Elicit is an AI research platform used by scientists, pharma companies, and decision-makers for high-stakes evidence synthesis. A single session could trigger hundreds of thousands of language model invocations across multiple providers, which means our infrastructure decisions directly impact cost, reliability, and the quality of research outcomes for users making decisions worth millions of dollars. Our infra is well-architected using best practices — Terraform, Kubernetes, Argo CD, GitHub Actions. But we're at an inflection point: enterprise contracts are getting larger, single-tenant deployments are multiplying, and the surface area that needs dedicated attention has outgrown what our current team can cover part-time. This is the first dedicated infrastructure hire and you'll define how this function works at Elicit. James (Head of Engineering, ex-Square) set up the original infrastructure and will be your close partner. This role will own and evolve the infrastructure platform that underpins Elicit's product. You will ensuring it is reliable, secure, cost-efficient, and ready for the demands of a growing enterprise customer base. Under your ownership, our systems will scale gracefully across single-tenant deployments, our SLAs will be backed by real engineering rigor rather than best intentions, and our compliance posture will be a selling point rather than an afterthought. What you'll own Own our cloud infrastructure across AWS and GCP — Kubernetes clusters, networking, databases (Aurora PostgreSQL, Redis, MongoDB Atlas), Cloudflare, and our CI/CD pipeline. Scale single-tenant deployments from a handful to many — each with distinct data retention, geographic, monitoring, and compliance requirements. Make a private cloud deployment a repeatable, low-overhead operation. Build our observability and incident response practice — proactive monitoring, alerting, SLA tracking, and structured post-mortems that make the whole team better at diagnosing and resolving issues. Drive compliance and security operations — ensure we follow through on the policies we've written (SOC 2, NIST AI framework, EU Cyber Resilience). Own disaster recovery exercises, database restoration drills, and security event monitoring (SIEM). Manage infrastructure cost and capacity — make smart decisions about where we run workloads (AWS, CoreWeave, Parasail), optimize spend, and plan capacity as usage grows. Improve developer experience — CI/CD pipeline performance, preview environments, local development tooling, and deployment confidence. Contribute to backend systems where infrastructure and application intersect — circuit breakers, inference routing, data connector infrastructure for enterprise customers bringing their own data. What success will look like (6-12 months) A private cloud deployment is a ~1-day turnkey operation. Playbooks and templated Terraform make standing up Elicit in a customer's cloud routine, which opens up 8-figure enterprise deals. Our observability signal:noise ratio improves 10-fold. Health monitors cover every endpoint and job, and an alert firing means something needs attention. Disaster recovery is practiced. We run database restoration drills and provider-outage dry runs on a schedule, with post-mortems that make the whole team better at diagnosis. Our SLAs are backed by engineering rigor. We follow through on SOC 2, NIST AI framework, and EU Cyber Resilience commitments, and enterprise security reviews go faster because of it. Inference is faster and cheaper. You've found and executed opportunities like shifting load between providers to cut p95 latency and cost at the same time. What we're looking for 5+ years of hands-on infrastructure/SRE/platform engineering experience. An AI-native way of working. Agentic coding tools (Claude Code, Cursor, Devin, etc.) are how we build at Elicit, and infrastructure is no exception: agents help us write IaC and investigate incidents. You should be an enthusiastic practitioner who uses AI to multiply your impact, and excited to find new places agents can safely take on infrastructure work. Bonus: you've written about, spoken about, or built projects demonstrating this. Solid Terraform experience. this is our primary infrastructure-as-code layer and the most important technical requirement. Strong Kubernetes expertise. You've operated production clusters, not just deployed to them. Comfortable with EKS, networking, autoscaling (Karpenter), and debugging cluster-level issues. AWS experience (primary), with GCP familiarity a plus. GitOps and CI/CD fluency. Argo CD, GitHub Actions, or equivalent. You understand deployment automation, rollback strategies, and change management. SRE mindset. You've built or significantly improved observability stacks (DataDog or equivalent), incident response processes, and on-call practices. Security and compliance awareness. Experience with SOC 2 or similar frameworks, SIEM tooling, and translating compliance requirements into engineering practice. Ability to write software. You can contribute to our backend codebases where infrastructure meets application logic. Am I a good fit? Strong applicants will find it easy to answer these questions: Can you describe a time you designed and executed a multi-tenant or single-tenant deployment architecture for enterprise customers? How have you approached disaster recovery planning and testing at a previous company? Walk me through how you'd evaluate whether to build vs. buy for a new infrastructure component at a ~30-person startup. Have you owned compliance follow-through (not just policy writing) for a framework like SOC 2? Who will I work with? James (Head of Engineering): set up the original infrastructure and will be your closest partner. You'll own execution, with James as sounding board and advocate. Panda: the engineer who has been covering infrastructure part-time, with deep context on our cluster bootstrapping, Cloudflare setup, and inference providers. Product: PMs covering the core product, ML, and evals. They carry the customer side of enterprise deployments, so you'll work together to turn requirements like data residency, compliance commitments, and SLAs into architecture, and to weigh the cost and latency tradeoffs behind product decisions. Eval infrastructure is a shared surface with Ben, from CI integration to inference capacity. The whole engineering team: we're ~30 people company-wide, so you'll work directly with the engineers whose developer experience you're improving, and pair with them where infrastructure meets application code. Andreas and Jungwon (cofounders): you'll meet both during the interview process, and infrastructure decisions with strategic weight (enterprise deployments, compliance posture) get their direct attention. On-call & incident expectations We don't have a formal on-call rotation yet. Incidents today are handled by the engineers closest to the affected system. Part of this role is building the incident response practice we should have: sensible alerting, SLA tracking, structured post-mortems, and eventually a rotation designed so it doesn't burn anyone out. You'd design the on-call setup you'll then live with. Why join us? The work matters. Our mission is to radically improve reasoning for high-stakes decisions. Over 2 million people use Elicit, including pharma teams making R&D decisions worth tens of millions of dollars, and the reliability of our platform is part of what makes those decisions sound. Infrastructure work is close to the business. Single-tenant deployments open enterprise deals, and inference routing choices show up directly in our costs and latency. You'll see the effect of your work in the company's trajectory. Built for the long term. We're a Public Benefit Corporation that spun out of a non-profit AI research lab. Long-term impact and AI safety are part of the corporate charter. High agency, low bureaucracy. ~30 curious, slightly weird people who write things down and trust each other to run with a vague brief. Serious investment in your growth. $1,000 per quarter for every person to explore AI tools, courses, and events, plus quarterly in-person team retreats. Location and travel We have a great office in Oakland, CA, and we'd love to see you there if you're local. That said, we're just as happy for you to work remotely. We do get the whole team together for a quarterly retreat somewhere fun, because in-person time matters to us. Benefits and perks In addition to working on important problems as part of a productive and positive team, we also offer great benefits (with some variation based on location): Flexible work environment: work from our office in Oakland or remotely with time zone overlap (between GMT and GMT-8), as long as you’re comfortable traveling for quarterly in-person offsites. Fully covered health, dental, vision, and life insurance for you, generous coverage for the rest of your family (FSA/HSA, too). Flexible vacation policy, with a minimum recommendation of 20 days / year and plenty of company holidays. Every Elician receives a $200 monthly wellbeing stipend to spend on whatever supports your health and wellbeing. 401K with a 6% employer match. A new Mac + $1,000 budget to set up your workstation or home office in your first year, then $500 every year thereafter. $1,000 quarterly AI Experimentation & Learning budget, so you can freely experiment with new AI tools to incorporate into your workflow, take courses, purchase educational resources, or attend AI-focused conferences and events. A team administrative assistant who can help you with personal and work tasks. You can find more reasons to work with us in this thread! Compensation For all roles at Elicit, we use a data-backed compensation framework to keep salaries market-competitive, equitable, and simple to understand. For this role, we target starting ranges of: Senior (L4): $185-260k + equity Expert (L5): $250-280k + equity Principal (L6): >$260 + significant equity We're optimizing for a hire who can contribute at a L4/senior-level or above. We offer above-market equity for all roles at Elicit, as well as employee-friendly equity terms.

On-siteFull-time$1.9L – $2.6L3w agoApply
Foundry Robotics

Infrastructure Engineer

Foundry Robotics

San Francisco

About Us Foundry Robotics is building an AI-native robotics manufacturing company focused on deploying advanced assembly and production capability for leading robotics companies and national-security-critical hardware. Basically, we’re building robots that build robots. Factory OS is building software that runs factories. Our platform spans cloud backends, on-prem services, and edge devices deployed on the factory floor. The operators, supervisors, and managers who use our tools need interfaces that are fast, reliable, and intuitive—even when network conditions aren’t perfect and the environment is demanding. The Role You will work on the infrastructure that our services, our factory-floor software, and our ML training run on: AWS, Kubernetes, the network boundary between them, the deployment pipeline, GPU workload scheduling, and the monitoring that tells us when something is wrong. This is a unique position with huge scope, spanning traditional containerized software platforms, as well as robotics, networking and AI infrastructure. Unlike an internet company, you will work on capabilities that support real hardware, and unlike traditional robotics, this platform also orchestrates the manufacturing process and the continued operation of the robots themselves. It is a small team with a lot of surface area. Expect to move between infrastructure as code, cluster operations, GPU scheduling, and edge-device support in the same week. What We`re Looking For 3+ years running production infrastructure with hands-on Kubernetes. Fluent in Terraform and GitOps. You have worked in infrastructure CI/CD where merge means apply and you know how to make that safe. Solid networking fundamentals: VPCs, routing, VPN, DNS, TLS, overlay networks, and the ability to debug a routing problem with a packet capture. Comfortable in Go or Python for tooling and glue, and able to read production service code well enough to debug it. Experience operating an observability stack and a track record of making alerts trustworthy. Security-minded by default: least-privilege access, no public endpoints, secrets from config, and the reflexes to catch a committed credential. Strong plus Experience running Kubernetes on-prem. Edge or fleet experience: arm64 Linux devices, over-the-air rollout, offline-tolerant design. WireGuard-based overlay networks at organization scale. Time spent near robots or industrial hardware. Experience in an ITAR or CMMC environment. Why Join Us? This is one of the only places where world-class manufacturing operators, mechanical engineers, robotics researchers, and software engineers sit in the same room — building production systems together. We are committed to being deeply embedded in the U.S. industrial base. Our focus is simple: build adaptive robotic assembly systems that make American manufacturing scalable, resilient, and competitive again. If you want to run a mature, well-defined commercial org, this may not be the role. If you want to build the commercial engine that brings AI-driven manufacturing to every industrial and energy customer in America — this is it. The base salary range for this full-time position in the location of San Francisco is: $150,000—$250,000 USD Compensation packages at Foundry Robotics for eligible roles include base salary, equity, and benefits. The range displayed on each job posting reflects the minimum and maximum target for new hire salaries for the position, determined by work location and additional factors, including job-related skills, experience, interview performance, and relevant education or training. Foundry Robotics employees in eligible roles are also granted equity based compensation, subject to Board of Director approval. You`ll also receive benefits including, but not limited to: Comprehensive health, dental and vision coverage, and generous PTO.

On-siteFull-time$1.5L – $2.5L1w agoApply

Infrastructure Engineer

Roboflow

NY, SF or Remote

Who We Are Our mission is to make the world programmable. Sight is one of the key ways we understand the world, and soon this will be true for the software we use, too. We’re building the tools, community, and resources needed to make the world programmable with artificial intelligence. Roboflow simplifies building and using computer vision models. Today, over 1M+ developers, including those from half the Fortune 100, use Roboflow’s machine learning open source and hosted tools. That includes counting cells to accelerate cancer research, improving construction site safety , digitizing floor plans , preserving coral reef populations , guiding drone flight , and much more . Roboflow is supported by great customers and investors, having raised over 63 million from Y Combinator, Google Ventures, Craft Ventures, Sam Altman, Lachy Groom, amongst other leading software investors. Roboflowers love building great things with passionate teammates. We value ownership, accountability, and a bias toward action—whether it`s a big initiative or a small fix. You’re naturally curious, hands-on with new tech (maybe even played with ChatGPT or AI products early on), and prefer to show your work over talking about it. Many of us have founder mindsets and thrive in Roboflow’s high-autonomy environment—some even started as side hustlers in school. What We`re Looking For Primarily, you like to make great things with passionate colleagues. You are someone that likes to own outcomes, not only inputs. You’re motivated by having responsibility and accountability. You’re eager to ‘do the work,’ big and small. You’re curious and learning about new technologies, perhaps an early tinkerer with MLOps products. You show more than you tell. You’re motivated by the question, “How can I improve this?” and have a track record of doing so, even in ways adjacent to your role. Much of our current team is made up of former founders and thrive in the level of autonomy at Roboflow. Maybe you had a side hustle in high school or college. Many Roboflowers have used our tools before joining. One of the best ways to stand out amongst other applicants is to write about something you have built with Roboflow or contribute to one of our open source projects . Likewise we highly value users with meaningful contributions to successful open source devtool and security projects. What You`ll Do As a member of our infrastructure team, you`ll be at the heart of a fast-paced startup environment. Your primary focus will be on striking the right balance between rapid delivery, high reliability, and robust security . This isn`t a traditional, siloed role; you`ll need to wear many hats—acting as an infrastructure engineer one moment, and a developer, or even a security analyst. You will be securing, scaling, and maintaining the core infrastructure that powers our product. This includes our cloud architecture, databases, file storage, search clusters, microservices, and machine learning pipelines. You`ll work closely with our product team and collaborate across the company on product, operations, and customer-facing projects, constantly context-switching to solve the next critical challenge. Skillset We`re looking for a versatile engineer excited by high-impact challenges. At Roboflow, we are AI-native: we expect our team to use AI to accelerate everything from writing code and fixing bugs to analyzing security, cost, and performance. Experience in some or all of the following areas will be crucial: Production experience with Kubernetes : Building and managing containerized applications at scale. Infrastructure-as-Code (IaC) : Using Terraform, Helm charts, bash scripting, and Python to automate everything. Scale & Site Reliability : Operating, monitoring, and scaling large-scale applications (especially in ML/AI) in AWS and/or GCP. Development Skills : Proficiency in Node.js and Python, with the ability to collaborate with full-stack developers on designing and operating SaaS applications. ML/Big Data Ops : Hands-on experience with the infrastructure required for machine learning at scale (GPUs, Docker, Kubernetes) and familiarity with libraries like PyTorch or Tensorflow. CI/CD Automation : Experience with tools like GitHub Actions or Spacelift to build and deploy code efficiently. Pragmatic Security : Awareness of security best practices for cloud operations and how they can be applied to startup environments. AI-Native Engineering: Leveraging LLMs and AI tools to accelerate the development lifecycle—from writing and refactoring code to identifying security vulnerabilities and optimizing infrastructure costs. A Glimpse of Your Work No two days will be the same. Your tasks will be a blend of strategic projects and hands-on implementation. Examples include: Running and optimizing a high-availability machine learning inference service. Collaborating with customer security teams to ensure secure integration. Developing creative IaC solutions to scale our platform cost-effectively. Working with the engineering team to define SLOs/SLAs and participating in incident response. Improving the Observability and Alerting stack and the processes built around it. Diving deep into our stack to identify and act on cost-optimization opportunities. Contributing code (Python, JavaScript, etc.) as part of a team designing and deploying new product features. Fixing security vulnerabilities and bugs Hardening our systems and processes to meet SOC 2, HIPAA, and GDPR requirements, making us audit-ready. Participating in an on-call rotation to ensure platform reliability. ???? Within one week, you will… Learn all about computer vision, our product, company, customers, and vision. Ship something substantial to an end user Start learning our infrastructure and security practices. ???? Within one month, you will… Onboard in person with your manager Build your first computer vision project with Roboflow (if you haven`t already) Start contributing to infra-as-code Start working with customers to help with their security questions and onboarding Understand the architecture of Roboflow ???? Within six months, you will… Attend your first all company onsite Be ramped up on other relevant parts of the Roboflow product. Who You`ll Be Working With Our team of ~100 attracts talent like executives that wanted to return to building, founders with a 100M+ exit, Roboflow users turned team members, open source contributors, a cyclist who biked across the United States, prolific high school hackers, a CTO from 100+ engineering organization, amongst many exceptional others. You will directly be working with our Engineering Lead and a team of product, infrastructure and security engineers. Where You`ll Work Roboflow is distributed across the US and Europe. We currently have Hubs in New York City and San Francisco (and plan to open more as we grow density in new cities). We provide opportunities (like team on-sites in different cities) and resources (like a $4000/yr travel stipend) to work in person with other team members as much as you`d like, while also supporting remote team members. You can work from one of our Hubs (we offer a relocation bonus), work from home, work at co-working spaces, etc. We want you to work where you work best! When You`ll Work Roboflow primarily operates during the daytime hours in the US and there are some synchronous meetings you’ll be expected to attend each week. Apart from that, we have a flexible schedule that allows you to work collaboratively with other team members and asynchronously when needed. What You`ll Receive To determine your salary, we use a number of market and data-driven salary sources. We review all salaries every six months to ensure we stay in line with the market. ???? The target compensation for this role is USD $165,000 base - $200,000 base. ???? In addition to our cash compensation, we offer generous perks and benefits. Below are some of the highlights: $4000/yr Travel Stipend to travel anywhere anytime to work alongside other Roboflowers $350/mo Productivity stipend to spend on things that make your work environment more productive, like high-speed internet at home or a co-working space Cover up to 100% of your health insurance costs for you and your partner or family Equity in the company so we are all invested in the future of computer vision Interview Process (~5 hours) Below is the interview process you can expect for this role. You will be speaking directly with our team about what it`s like to work and thrive at Roboflow. Before the Interview: We’ll review your application, LinkedIn, Github, etc. The best way to stand out is to write about something you’ve built with Roboflow or contribute to one of our open source projects , or highlight your contributions to devtools/infrastructure/security engineering open-source projects. We may send you a technical screen if applicable. Introduction Phase: [45m] Meet with hiring manager for introduction, Sachin Agarwal, to assess overall mindset and skillset. This first interview is a time to get to know more about the role, allow us to get to know you better, and ensure it`s a good fit for both parties to continue moving forward in the process Team Interview Phase: [45m] Meet with our CTO, Brad Dwyer [90m] Meet with hiring manager and team for a technical infrastructure hands-on interview Ask questions! Final Interview Stage: [45m] Meet with Kate Wagner, Head of Operations for a culture discussion [60m] Meet with Joseph Nelson, CEO We check references and conduct a background check Note: you are welcome to request additional conversations with anyone you would like to meet and we will accommodate as best we can. Learn More About Us We are building a diverse Distributed team that is distributed across the globe. Roboflow is an equal opportunity workplace; we welcome people from all backgrounds, communities, and experiences. We provide competitive compensation and stellar benefits to accelerate your personal and work life. Learn more about what it is like to work at Roboflow by reading these blog posts. See our careers page for all open listings.

RemoteFull-time2w agoApply

Infrastructure Engineer

Bishopfox

U.S. Remote

At Bishop Fox, security isn't just a job—it's our passion. As leaders in continuous offensive security and penetration testing, we deliver world-class customer experiences. Trusted by over a quarter of the Fortune 100, half of the Fortune 10, and top global media companies, we help safeguard digital landscapes. Our Cosmos platform, honored as Best Emerging Technology by SC Media, exemplifies our commitment to innovation. Joining Bishop Fox means collaborating with a curious and dedicated team. You'll tackle complex challenges for some of the world's most recognized organizations, securing their networks against real-world threats. With nearly 20 years of industry contributions—including 16 open-source tools and 50 security advisories published in the past five years—we're committed to making the digital world safer. We’re looking for talented, experienced professional hackers to help us secure some of the world’s most complex software and sophisticated technologies. You’ll be working alongside our US and internationally-based teams supporting clients across multiple industries. Who You Are and What You’ll Do As an Infrastructure Engineer, you are a service owner who designs, builds, deploys, and operates the infrastructure Bishop Fox runs on across a multi-cloud AWS and Azure estate and the Microsoft 365/Entra ID platform. Infrastructure is defined as code in Terraform and deployed through GitHub Actions, with services built primarily in Go. As a knowledgeable member of the IT team, you will monitor service health, respond to incidents, use AI tooling day-to-day, and mentor more junior members while working to streamline and improve IT tools, processes, and procedures. Write, review, and maintain Terraform modules and infrastructure as code across our AWS and Azure environments Build and maintain production Go and Python services and CI/CD pipelines in GitHub Actions Deploy and own new services, monitor their health, and respond to incidents Maintain the Azure environment, including compute, storage, networking, and secrets management Support the Microsoft 365 and Entra ID platform, including conditional access and endpoint management through MDM platforms Build and maintain integrations between business systems Update and create automations for tasks and workflows General network troubleshooting for small business locations Perform infrastructure/maintenance changes while adhering to change management processes Use AI-assisted development tooling and support the internal AI services the IT team operates Provide assistance, guidance, and mentorship to analyst roles Engage in practice development activities by developing tools, improving processes, and developing training material Serve as a configuration manager to create, apply, and enforce processes for promoting all infrastructure components from the development environment to the testing, demonstrating, and production environments Various day-to-day tasks or projects as assigned by leadership Your Experience Associates or Bachelors Degree in IT, Computer Science, or a related field OR 6+ years relevant professional experience Production experience with infrastructure as code (Terraform) and application/service development in Go or Python Experience building or maintaining CI/CD pipelines, GitHub Actions preferred Hands-on experience administering Azure Experience integrating systems Working knowledge of identity and access management; Entra ID experience preferred Prior experience with Salesforce preferred but not required Knowledge of scripting (PowerShell, Bash) and Linux/Windows server administration Comfort using AI coding assistants as part of an engineering workflow Demonstrates solid judgement in decision making that positively impacts users, projects, and Bishop Fox Excels at process building with a demonstrated ability to create templates, documentation for new processes Provides excellent customer service to internal customers and vendors Navigates unusual hours and last-minute requests with flexibility and calm Good written and verbal communication skills, including the ability to explain technical trade-offs to non-technical stakeholders Excels at building professional trust and maintaining confidentiality Team player mentality with the ability to work as a self-motivated individual contributor Detail-oriented with excellent time management, prioritization, and multi-tasking skills Why Bishop Fox At Bishop Fox, we're driven by a simple mission: deliver exceptional quality to our clients, foster a vibrant and fulfilling environment for our team, and champion excellence within our industry. Our core values, which we live by every day, are: Be Excellent to Each Other Do the Right Thing Do What You’ll Say You’ll Do Get Better Together Give a Sh*t Get Better Together At Bishop Fox, we're committed to providing benefits that support your well-being and professional growth. Here's a glimpse of what we offer: Generous Time Off and Company-Wide Holidays Team Events and International Travel Opportunities Work From Home Support Monthly Allowance for Cell Phone and Internet Training Budget Retirement; 401k Matching for Traditional and Roth Accounts in the US Health Insurance Options Including Medical, Dental, Vision Paid Parental Leave Bishop Fox is an Equal Opportunity employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex including sexual orientation and gender identity, national origin, disability, protected veteran status, or any other characteristic protected by applicable federal, state, or local law. All new hires must pass a background check as a condition of employment. Interested? Apply today!

RemoteFull-time3w agoApply

Infrastructure Engineer

Bishop Fox

USA

At Bishop Fox, security isn't just a job—it's our passion. As leaders in continuous offensive security and penetration testing, we deliver world-class customer experiences. Trusted by over a quarter of the Fortune 100, half of the Fortune 10, and top global media companies, we help safeguard digital landscapes. Our Cosmos platform, honored as Best Emerging Technology by SC Media, exemplifies our commitment to innovation. Joining Bishop Fox means collaborating with a curious and dedicated team. You'll tackle complex challenges for some of the world's most recognized organizations, securing their networks against real-world threats. With nearly 20 years of industry contributions—including 16 open-source tools and 50 security advisories published in the past five years—we're committed to making the digital world safer. We’re looking for talented, experienced professional hackers to help us secure some of the world’s most complex software and sophisticated technologies. You’ll be working alongside our US and internationally-based teams supporting clients across multiple industries. Who You Are and What You’ll Do As an Infrastructure Engineer, you are a service owner who designs, builds, deploys, and operates the infrastructure Bishop Fox runs on across a multi-cloud AWS and Azure estate and the Microsoft 365/Entra ID platform. Infrastructure is defined as code in Terraform and deployed through GitHub Actions, with services built primarily in Go. As a knowledgeable member of the IT team, you will monitor service health, respond to incidents, use AI tooling day-to-day, and mentor more junior members while working to streamline and improve IT tools, processes, and procedures. Write, review, and maintain Terraform modules and infrastructure as code across our AWS and Azure environments Build and maintain production Go and Python services and CI/CD pipelines in GitHub Actions Deploy and own new services, monitor their health, and respond to incidents Maintain the Azure environment, including compute, storage, networking, and secrets management Support the Microsoft 365 and Entra ID platform, including conditional access and endpoint management through MDM platforms Build and maintain integrations between business systems Update and create automations for tasks and workflows General network troubleshooting for small business locations Perform infrastructure/maintenance changes while adhering to change management processes Use AI-assisted development tooling and support the internal AI services the IT team operates Provide assistance, guidance, and mentorship to analyst roles Engage in practice development activities by developing tools, improving processes, and developing training material Serve as a configuration manager to create, apply, and enforce processes for promoting all infrastructure components from the development environment to the testing, demonstrating, and production environments Various day-to-day tasks or projects as assigned by leadership Your Experience Associates or Bachelors Degree in IT, Computer Science, or a related field OR 6+ years relevant professional experience Production experience with infrastructure as code (Terraform) and application/service development in Go or Python Experience building or maintaining CI/CD pipelines, GitHub Actions preferred Hands-on experience administering Azure Experience integrating systems Working knowledge of identity and access management; Entra ID experience preferred Prior experience with Salesforce preferred but not required Knowledge of scripting (PowerShell, Bash) and Linux/Windows server administration Comfort using AI coding assistants as part of an engineering workflow Demonstrates solid judgement in decision making that positively impacts users, projects, and Bishop Fox Excels at process building with a demonstrated ability to create templates, documentation for new processes Provides excellent customer service to internal customers and vendors Navigates unusual hours and last-minute requests with flexibility and calm Good written and verbal communication skills, including the ability to explain technical trade-offs to non-technical stakeholders Excels at building professional trust and maintaining confidentiality Team player mentality with the ability to work as a self-motivated individual contributor Detail-oriented with excellent time management, prioritization, and multi-tasking skills Why Bishop Fox At Bishop Fox, we're driven by a simple mission: deliver exceptional quality to our clients, foster a vibrant and fulfilling environment for our team, and champion excellence within our industry. Our core values, which we live by every day, are: Be Excellent to Each Other Do the Right Thing Do What You’ll Say You’ll Do Get Better Together Give a Sh*t Get Better Together At Bishop Fox, we're committed to providing benefits that support your well-being and professional growth. Here's a glimpse of what we offer: Generous Time Off and Company-Wide Holidays Team Events and International Travel Opportunities Work From Home Support Monthly Allowance for Cell Phone and Internet Training Budget Retirement; 401k Matching for Traditional and Roth Accounts in the US Health Insurance Options Including Medical, Dental, Vision Paid Parental Leave Bishop Fox is an Equal Opportunity employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex including sexual orientation and gender identity, national origin, disability, protected veteran status, or any other characteristic protected by applicable federal, state, or local law. All new hires must pass a background check as a condition of employment. Interested? Apply today!

On-siteFull-time3w agoApply
SimSpace Corporation

Infrastructure Engineer

SimSpace Corporation

Boston Office

SimSpace serves as an AI Proving Ground where organizations can confidently train, test, and outmaneuver adversaries in any environment. Trusted by allied governments, militaries, enterprises, and research institutions worldwide, SimSpace enables adaptive, AI-ready defenses that stay ahead of evolving threats. Founded in 2015 by experts from U.S. Cyber Command and MIT Lincoln Laboratory, the platform unifies training, testing, and validation in a realistic, live-fire simulation—helping teams evaluate security investments, optimize performance, and compress cyber readiness cycles from months to days. Why join SimSpace? We are an organization that is focused on building our culture and mindfully enhancing our atmosphere every day which is why we have collaborated on an integral value system. Our governing philosophy of being Human Centered is deeply embedded within our value system. We apply this philosophy to every one of our internal team members, external clients, and their customers. How Do We Work? We believe that people are at the center of everything we do. SimSpace fosters a culture of continuous learning, curiosity, and professional growth. That belief shows up in action: in-house training, internal and external learning platforms, cyber conferences, industry events, and dedicated time for skill development. Our people are empowered to shape their careers - and it shows. Year over year, SimSpace consistently outperforms industry benchmarks in internal mobility, promotions, and total rewards growth. Who Thrives Here? We are a team of innovators, protectors, and problem-solvers. We believe diversity of thought and experience fuels better solutions, and we’re committed to building teams that reflect the communities we serve. Whether you’re remote or office-based, you’ll collaborate with talented colleagues across departments and time zones, united by the mission to create a safer digital world. We invite you to apply today! The Infrastructure / Data Center Operations Engineer is responsible for the reliability, operation, and continuous improvement of critical data center and virtualization infrastructure. This is a hands-on role spanning VMware platforms, physical data center operations, monitoring, technical documentation, basic Linux administration, and network infrastructure. The ideal candidate is comfortable operating in a high-availability environment where disciplined change management, accurate documentation, physical-layer execution, and rapid troubleshooting are essential. This role is based in the Greater Boston area and participates in an on-call rotation to support business-critical infrastructure. What You Will Do • Virtualization Operations: Deploy, operate, troubleshoot, and maintain VMware-based infrastructure, including day-to-day administration of vCenter-managed environments and associated compute resources. • Data Center Operations: Perform Layer 1 data center work including rack-and-stack, structured cable management, hardware installation and replacement, server repair, and physical troubleshooting. • Availability & Incident Response: Operate infrastructure in a 99.999% uptime environment, respond to monitoring alerts and incidents, participate in root-cause analysis, and support disciplined recovery and remediation activities. • Monitoring & Operational Health: Use infrastructure and monitoring tools to assess system health, identify emerging issues, validate service availability, and support proactive capacity and reliability management. • Technical Documentation: Create and maintain accurate infrastructure documentation, including Visio diagrams, IP-address management records, rack layouts, cable documentation, and network visualizations. • Linux & Remote Systems: Perform basic Linux deployment and administration tasks, use common troubleshooting tools, and connect securely to remote systems for operational support. • Network Operations: Perform basic configuration and troubleshooting of Cisco and Dell switching environments and support connectivity changes associated with server and data center operations. • On-Call Support: Participate in an on-call rotation and respond to infrastructure issues that require after-hours investigation, escalation, or hands-on remediation. What You Will Bring • VMware Expertise: VMware Certified Professional (VCP) or equivalent demonstrated expertise, with 5+ years of relevant experience operating VMware-based infrastructure. • High-Availability Operations: Experience supporting infrastructure in a 99.999% uptime environment with strong operational discipline and attention to risk. • Data Center Experience: Hands-on Layer 1 experience with rack-and-stack, cable management, server hardware, component replacement, and physical troubleshooting. • Documentation: Strong technical-documentation skills using tools and practices such as Microsoft Visio, IP management, rack layouts, and network visualization. • Monitoring: Experience using infrastructure-monitoring tools to identify, troubleshoot, and respond to operational issues. • Linux: Working knowledge of Linux sufficient to deploy systems, use common operational tools, and connect to remote environments. • Networking: Working knowledge of basic networking concepts and hands-on experience with Cisco and/or Dell switch configuration. • Location & Availability: Located in the Greater Boston area and able to participate in an on-call rotation, including occasional after-hours response as required. Preferred Qualifications • Storage Systems: Experience with enterprise storage platforms such as Pure Storage, NetApp, or comparable technologies. • Software-Defined Networking: Experience with technologies such as VMware NSX, vCenter integrations, GENEVE, VXLAN, or related software-defined networking concepts. • Public Cloud: Operational or design experience with Azure, Google Cloud Platform, AWS, or similar public-cloud environments. • Firewalls: Experience administering or supporting firewall platforms such as Palo Alto Networks, Fortinet, or comparable products. • Project & Technical Design Skills: Ability to plan and execute infrastructure projects, develop technical designs, coordinate dependencies, and communicate implementation requirements. • Technical Awareness: Habit of monitoring relevant CVEs, security breaches, vendor advisories, geopolitical events, and other developments that may affect infrastructure risk or operations. • Customer / VAR Experience: Experience working in a VAR, customer-facing, or end-user-facing environment, including the ability to join customer calls and clearly explain technical requirements, configurations, and operational tradeoffs. We’re proud to offer a competitive and comprehensive package designed to support your well-being, growth, and success: Compensation. Base salary range: $95,000 - $145,000, reflecting our confidence in your expertise and impact, with the opportunity for annual bonuses tied to company performance and individual contributions. Health & Wellness. Comprehensive medical, dental, and vision benefits, plus savings plans—coverage starts on day one! Mental Health Support. Access to company-paid counseling, coaching, and resources for you and your family through Spring Health. Financial Well-Being. Plan for your future with a 401(k)-retirement savings plan featuring a company match. Flexible Time Off. Take the time you need with flex vacation, company holidays, and dedicated health & wellness days. SimSpace provides flexible solutions to meet the diverse work-life needs of team members. Parental Leave. Paid leave plans to support you and your loved ones during life’s most important moments. Ownership Opportunities: Equity stock options at hire, with annual performance-based grants—become an invested stakeholder in our shared success. Peloton Interactive Wellness Program: Full- and partial- subsidized membership plans and equipment discounts to help you reach your personalized fitness goals. Continuous Learning: Access a LinkedIn Learning membership to prioritize your personal and professional development. Social Connections: Monthly reimbursements for meaningful connections with teammates through our SocialSpace Community. Extra Perks: Legal plan coverage, pet insurance, wellness reimbursements, and more to simplify life’s details. SimSpace is an Equal Opportunity Employer: In compliance with federal law, all persons hired will be required to verify identity and eligibility to work in the United States and to complete the required employment eligibility verification document form upon hire. SimSpace is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, age, pregnancy, genetic information, disability, status as a protected veteran, or any other protected category under applicable federal, state, and local laws. We are committed to providing an inclusive and welcoming environment for all members of our staff, clients, volunteers, subcontractors, vendors, and clients. Research shows that women and people from underrepresented groups only apply to jobs if they meet all of the qualifications. However, no one ever meets 100% of the qualifications. SimSpace encourages you to break that statistic and to apply. We look forward to your application! We also consider qualified applicants regardless of criminal histories, in accordance with applicable law. We are committed to providing reasonable accommodations for qualified individuals with disabilities in our job application procedures. If you need assistance or accommodation due to a disability, please contact careers@simspace.com . SimSpace does not accept unsolicited resumes from employment agencies. Actual compensation for the position is based on a variety of factors, including, but not limited to affordability, skills, qualifications and experience, and may vary from the range.

On-siteFull-time$95K – $1.4L1w agoApply
Writer

Infrastructure engineer

Writer

New York City, NYSeattle, WA, San Francisco, CA

???? About WRITER WRITER is where the world`s leading enterprises orchestrate AI-powered work. Our vision is to expand human capacity through superintelligence. And we`re proving it`s possible – through powerful, trustworthy AI that unites IT and business teams together to unlock enterprise-wide transformation. With WRITER`s end-to-end platform, hundreds of companies like Mars, Marriott, Uber, and Vanguard are building and deploying AI agents that are grounded in their company`s data and fueled by WRITER`s enterprise-grade LLMs. Valued at $1.9B and backed by industry-leading investors including Premji Invest, Radical Ventures, and ICONIQ Growth, WRITER is rapidly cementing its position as the leader in enterprise generative AI. Founded in 2020 with office hubs in San Francisco, New York City, Seattle, Austin, Chicago, and London, our team thinks big and moves fast, and we`re looking for smart, hardworking builders and scalers to join us on our journey to create a better future of work with AI. ???? About the role At WRITER, our mission to expand human capacity with superintelligence relies on a foundational truth: our platform must be available, performant, and reliable, 24/7. As an Infrastructure engineer, you`ll be at the heart of making this a reality, impacting every enterprise customer who trusts us with their AI-powered workflows. This isn`t just about keeping the lights on; it`s about pushing the boundaries of what`s possible, proactively identifying and solving complex systemic challenges, and laying the groundwork for our rapid growth and the evolving demands of enterprise generative AI. You`ll build resilient systems, automate across the stack, and champion reliability best practices, directly enabling our ambitious product roadmap and ensuring our customers always have access to the powerful tools they need. This is a hybrid position, based out of our New York City, San Francisco, Seattle, or London hubs. You`ll report to our director of engineering. ????????‍♀️ What you`ll do Technical Breadth across disciplines. Bring deep focus to one problem at a time, with the breadth to move between SRE, DevOps, Infrastructure, and Platform work over a quarter or two as the leverage shifts. This is not a thrash-every-week role — most of the time you`re heads-down on one substantial initiative (the on-call posture, the release pipeline, the multi-region Terraform layout, the internal platform surface). Cross-layer fluency is what lets you pick the right next initiative; it isn`t a weekly context-switch. Simplicity / via negativa. Challenge the status quo and remove toil before adding features — automate operational tasks and infrastructure management with Python or Go, reject tools that don`t fit the problem, and treat manual on-call work as a defect to be designed out, not a status quo to be staffed up. Breadth across the stack. Design scalable, fault-tolerant infrastructure across AWS (preferred), GCP, and Azure, working fluently across Kubernetes, Helm, Terraform, and the supporting cloud and AI tooling that backs WRITER`s high-traffic platform. AI in workflow. Run agents in your daily loop — Claude Code, Droid, Codex, internal skills — to investigate incidents, draft Terraform / Helm changes, write runbooks, scaffold tooling, and review PRs. Build the agentic setup as a collective surface: humans and digital teammates working as one team, with shared skills, shared context, and shared on-call workflows. Encode recurring infra tasks as internal skills any teammate (human or agent) can pick up and run, so the team`s throughput compounds — not just your own. Debugging fluency. Lead incident response, post-mortems, and root-cause analyses — trace failures to the underlying problem (never the symptom), apply the learning back into the architecture, and prevent the same incident from happening twice. Non-technical End-to-end ownership. Own the reliability, performance, and efficiency of WRITER`s core services end-to-end — define and uphold the SLOs and error budgets, carry the on-call pager, and stand behind the outcome metric, not just the system you shipped. Strategic vs. tactical balance. Balance this week`s critical work with the 6–12-month platform direction — ship the on-call-driving fix today while shaping the multi-year observability, cost, and reliability investments that move WRITER`s enterprise customers. Cross-functional collaboration. Operate at the seams with product, security, and engineering peers — provide expert guidance on system design for reliability, performance, and scalability from conception through launch, Connect the infra agenda to product and revenue context, and disagree with evidence, not volume. ⭐️ What you need Technical Track record. 5+ years of experience in infrastructure engineering, DevOps, or a similar role focused on building and operating large-scale, high-availability production systems at a high-growth product company. Breadth. Experience running containerisation in production (a real cluster, not a lab), with experience in Helm and Terraform or Pulumi on at least one major cloud (AWS preferred), plus good proficiency in Python or Go for automation and tooling. AI in workflow. AI is part of how you ship, not a thing you`ve read about — agentic tooling (Claude Code, Droid, Codex, internal skills) is in your daily loop, you`ve built or adopted AI-assisted workflows others now use, and you have strong opinions on where it`s unreliable. This is a hard requirement, not a bonus. Candidates whose actual daily workflow does not already include AI tooling will not be advanced. First-principles + decision-making . Demonstrated ability to Challenge the status quo, proactively identify systemic weaknesses, and propose innovative solutions to complex reliability problems — reason from constraints and failure modes (not analogy or vendor defaults), name the tradeoff in business terms (reliability vs. velocity, cost vs. blast radius, standardisation vs. one-off), and reject the "best practices" answer when it doesn`t fit the problem. Reversibility & blast-radius. Make reversible calls by default — write the rollback before you touch production, work fluently with monitoring and logging stacks (Prometheus, Grafana, ELK or equivalent), and stress the system in safe places so it comes back stronger. Non-technical Cross-functional collaboration. Excellent communication, collaboration, and problem-solving skills, with a talent for building strong relationships and Connecting with cross-functional teams — surface non-goals before anyone asks, and partner with product, security, and platform peers as one delivery surface. Autonomy & end-to-end ownership. A strong sense of ownership and accountability, eager to Own mission-critical systems and drive them toward peak performance and unparalleled reliability. At least one 0-to-1 infrastructure build you owned end-to-end, with the outcome metric attached. ???? Bonus if you have Software-engineering depth. A software-engineering background, not only config and scripting — you`ve designed, built, and shipped non-trivial production code (services, libraries, internal frameworks) in Python, Go, or a comparable language, you can read and modify the codebases your infrastructure runs, and you move between infra automation and feature engineering without changing brains. ???? Benefits & perks (US Full-time employees) Generous PTO, plus company holidays Medical, dental, and vision coverage for you and your family Paid parental leave for all parents (16 weeks) Fertility and family planning support Early-detection cancer testing through Galleri Flexible spending account and dependent FSA options Health savings account for eligible plans with company contribution Annual work-life stipends for: Wellness stipend for gym, massage/chiropractor, personal training, etc. Learning and development stipend Company-wide off-sites and team off-sites Competitive compensation, company stock options and 401k WRITER is an equal-opportunity employer and is committed to diversity. We don`t make hiring or employment decisions based on race, color, religion, creed, gender, national origin, age, disability, veteran status, marital status, pregnancy, sex, gender expression or identity, sexual orientation, citizenship, or any other basis protected by applicable local, state or federal law. Under the San Francisco Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records. By submitting your application on the application page, you acknowledge and agree to WRITER`s Global Candidate Privacy Notice .

On-siteFull-time$7K – $8K1w agoApply
Onebrief

Infrastructure Engineer

Onebrief

United States

Design, build, and operate secure cloud-native and edge infrastructure for classified and commercial environments. Own end-to-end platform outcomes, harden artifact pipelines for signed builds across cloud and air-gapped deployments, embed with product teams to advise on secure deployments, and support compliance/audit activities (STIGs, CVE remediation) to enable operation in IL5/IL6/JWICS environments. Top Skills: Air-Gapped Appliances, Amazon Aws, Application Streaming, Artifact Pipeline, Artifact Signing, Aws Govcloud, Azure Government, Cve Remediation, Fedramp, Go, Identity, Il5, Il6, Jwics, Kubernetes, Azure, Multi-Cluster Operations, Network Segmentation, Python, Rust, Secrets Management, Stigs, Typescript Industry: Software , Defense

On-siteContract$1.8L – $2.9LNaNy agoApply
Zensar Technologies Inc.

Infrastructure Engineer

Zensar Technologies Inc.

Phoenix, AZ

No Description

On-siteFull-time$1.0L – $1.2L1w agoApply
V2Soft

Infrastructure Engineer

V2Soft

Sandston, VA

Roles and Responsibilities:Coordinating and performing Physical data center operations such as rack and stack, structured copper and fiber cabling installation, hardware break fix, and asset management.Must Have Skills: Direct experience with physical data center operations such as rack and stack, structured copper and fiber cabling.Hardware break-fixAsset ManagementSolid understanding of power, cooling and physical security controls.Team members are expected to operate within formal change, inc

On-siteFull-time2w agoApply
Quantum Technologies LLC

Azure Infrastructure Engineer

Quantum Technologies LLC

Rockville, MD

Job Title : Azure Infrastructure Engineer Location : Rockville, MD Duration: 12 Months Bill Rate: $83/hr. on W2 Job type: W2-Contract Work Authorization: US-Citizen, H-1B, OPT-EAD, GC-EAD Job Description: Lead and execute end-to-end cloud migration projects (Lift-and-Shift, Re-Platforming, etc.). Collaborate with application owners, network engineers, and security teams to architect cloud-based solutions. Implement infrastructure as code (IaC) using ARM templates, Bicep, or Terraform. Eval

On-siteContract$80 – $903w agoApply
Hadrian

Infrastructure Engineer (LATAM)

Hadrian

United States

Are you obsessed withbuilding cloud-native platformsthat enable engineers to operate at internet scale? Then join us as our newInfrastructure Engineerand be part of defining the next generation of the Infrastructure platform atHadrian! Who are we: We are Hadrian - an offensive cybersecurity startup. We are reshaping the future of cybersecurity through the power of agentic AI injected with the hacker’s perspective. Founded in August 2021, we've secured funding and the backing of ABN AMRO Ventures and Cherry Ventures, propelling us into a new era. Our goal is to provide companies with an autonomous Threat Exposure Management platform. At Hadrian, we view security through a hacker's eyes because hackers understand hackers best. We continuously map the digital footprint of organisations, discover risks, and prioritise remediation for security teams to harden their external attack surfaces. Located in Amsterdam's buzzing Leidseplein, by London's famous Paddington Station, and our most recent US HQ in the heart of New York City, our diverse team from 25+ countries is on a mission to shake things up in offensive security. Join us in making waves and securing over 700 businesses, and be at the forefront of automated offensive security. Your role: As an Infrastructure Engineer at Hadrian, you will design, build, and evolve the cloud-native platform that powers our offensive security operations. You will enable engineering teams to move faster, deploy safely, and operate reliably by creating scalable self-service infrastructure and resilient platform capabilities. This role is ideal for engineers who enjoy combining deep infrastructure expertise with a strong software engineering mindset. We're looking for someone who thrives on automation, reliability, and platform design, and who sees infrastructure as a product that empowers others to do their best work. If you enjoy solving complex distributed systems challenges, building developer-friendly platforms, and creating the foundations that allow autonomous security systems to operate at scale, we would love to hear from you What you will do as an Infrastructure Engineer at Hadrian: Help build and maintain the cloud-native platform that powers Hadrian's offensive security systems. Design and automate infrastructure delivery using Infrastructure-as-Code, enabling scalable and self-service platform capabilities. Manage and improve our Kubernetes environment, ensuring reliability, performance, and security. Collaborate closely with software engineers to improve deployment workflows, developer experience, and platform adoption. Operate and optimize our core cloud infrastructure on AWS, while supporting our distributed scanning environments. Improve observability, monitoring, incident response, and platform reliability across critical systems. Help implement platform standards, security controls, and engineering best practices that enable teams to move quickly and safely. Take on a unique challenge. We’re building autonomous security systems that continuously analyze internet-scale attack surfaces. You'll help create the platform that makes this possible. You are fit for this job because you: Are fluent in English. Have 3-5 years of experience as an Infrastructure Engineer, Platform Engineer, DevOps Engineer, SRE, or similar role. Have hands-on experience with Kubernetes, AWS, and Terraform (/OpenTofu), and are comfortable operating cloud-native infrastructure in production environments. Are comfortable managing core environments on AWS while extending workloads into secondary cloud environments used for scanning. Have a solid foundation in systems engineering, containerization (Docker), and distributed systems. Are experienced with automation and CI/CD tooling, such as Terraform, Ansible, Jenkins, or GitLab CI. Enjoy building platforms, tooling, and automation that help engineering teams move faster and more safely. Possess strong troubleshooting and problem-solving skills, and enjoy solving reliability, scalability, and operational challenges. Use AI pragmatically to accelerate workflows while maintaining high standards of reliability and security. Communicate clearly and collaborate effectively across technical teams. Bonus points if you have: Experience with KNative, event-driven architectures, or service mesh technologies. Experience with Kafka or other distributed messaging systems. Hands-on experience with observability and monitoring tooling. Familiarity with Site Reliability Engineering (SRE) practices. Certifications related to Kubernetes such as CKA, CKAD, or AWS certifications. Contributions to open-source infrastructure or cloud-native projects. Our Stack: GoLang / Python KNative Serving & Eventing / Kafka / Redis Docker / Kubernetes / Helm Charts Terraform / OpenTofu PostgreSQL / BigQuery AWS / GCP / SCW (plus) Benefits at Hadrian: Permanent contract Unlimited paid holiday Stock options package Mobile phone stipend Work-from-home budget Business travel card Swapfiets Referral bonuses Why you want to be part of Hadrian: Make the internet safer: We enable companies to protect their customers, employees, and other stakeholders from malicious hackers by providing real-time Exposure Management from the Hacker’s perspective Board a primed rocket ship: Hadrian is growing fast; the team is approaching 90 wonderful people, notable customers are signed, and significant funding is raised. Now is the time to join our high-growth journey Be part of a strong and dynamic culture: Our values define who we are as a group and serve as the cornerstones to guide us in building a product that our customers love. At Hadrian, we achieve this by championing a culture that takes ownership, succeeds together, increases velocity, and pursues growth Sounds like the role for you? Apply now! Our clients are seeking a world-class and reliable digital security service. That is exactly what we are offering. Therefore, a background check will be part of the process. Originally posted on Himalayas

On-siteFull-time3d agoApply
Bright Vision Technologies

Intelligent Infrastructure Engineer

Bright Vision Technologies

Sterling, VA 20166

Intelligent Infrastructure Engineer – Remote Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States. This is a fantastic opportunity to join an established and well-respected organization offering tremendous career growth potential. Job Title: Intelligent Infrastructure Engineer Location: 100% Remote (United States) Position Type: Full-time, Direct W2 Salary Range: $100,000–$150,000 Annually Experience: 6+ years Sponsorship: U.S. Citizens, Green Card Holders, EAD Holders, and H-1B transfer candidates are encouraged to apply. We are unable to sponsor new H-1B visa petitions for this position. Job Summary We are seeking an Intelligent Infrastructure Engineer to design, build, and operate the platform layer that powers large-scale AI training and inference workloads. The role focuses on GPU clusters, distributed training frameworks, scheduling, storage performance, and developer experience for ML engineers and researchers, with strong emphasis on reliability, efficiency, and cost control. The ideal candidate has built or operated production AI infrastructure at scale, understands the interaction between hardware, kernel, scheduler, and ML framework, and brings strong software engineering discipline to platform work. Required Qualifications Bachelor’s or Master’s degree in Computer Science or a related field. Six or more years of experience in infrastructure, platform, or HPC engineering. Hands-on experience operating GPU clusters or large-scale ML training infrastructure. Strong proficiency in Python and at least one systems language such as Go or C++. Deep understanding of distributed training, accelerator architectures, and collective communication. Experience with Kubernetes, Slurm, Ray, or similar scheduling systems for ML workloads. Strong understanding of Linux internals, networking, and high-performance storage. Experience with at least one major cloud provider’s ML infrastructure offerings. Strong software engineering practices including testing, CI/CD, and code review. Excellent communication and cross-functional collaboration skills. Preferred Qualifications Experience operating InfiniBand or RDMA networking at scale. Contributions to open-source ML infrastructure projects. Familiarity with custom orchestrators or research-grade training stacks. Exposure to frontier model training operations. Experience with FinOps for AI workloads. How to Apply Would you like to know more about this opportunity? For immediate consideration, please send your resume to boon@bvteck.com or contact us at (908) 650-6699 . Learn more about Bright Vision Technologies at www.bvteck.com . Bright Vision Technologies is an Equal Opportunity Employer. Equal Employment Opportunity (EEO) Statement Bright Vision Technologies (BV Teck) is committed to equal employment opportunity (EEO) for all employees and applicants without regard to race, color, religion, sex, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, veteran status, or any other protected status as defined by applicable federal, state, or local laws. This commitment extends to all aspects of employment, including recruitment, hiring, training, compensation, promotion, transfer, leaves of absence, termination, layoffs, and recall. BV Teck expressly prohibits any form of workplace harassment or discrimination. Any improper interference with employees' ability to perform their job duties may result in disciplinary action up to and including termination of employment.

On-siteFull-time$1.0L – $1.5L1mo agoApply
Bright Vision Technologies

Intelligent Infrastructure Engineer

Bright Vision Technologies

Sterling, VA 20166

Intelligent Infrastructure Engineer – Remote Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States. This is a fantastic opportunity to join an established and well-respected organization offering tremendous career growth potential. Job Title: Intelligent Infrastructure Engineer Location: 100% Remote (United States) Position Type: Full-time, Direct W2 Salary Range: $100,000–$150,000 Annually Experience: 6+ years Sponsorship: U.S. Citizens, Green Card Holders, EAD Holders, and H-1B transfer candidates are encouraged to apply. We are unable to sponsor new H-1B visa petitions for this position. Job Summary We are seeking an Intelligent Infrastructure Engineer to design, build, and operate the platform layer that powers large-scale AI training and inference workloads. The role focuses on GPU clusters, distributed training frameworks, scheduling, storage performance, and developer experience for ML engineers and researchers, with strong emphasis on reliability, efficiency, and cost control. The ideal candidate has built or operated production AI infrastructure at scale, understands the interaction between hardware, kernel, scheduler, and ML framework, and brings strong software engineering discipline to platform work. Required Qualifications Bachelor’s or Master’s degree in Computer Science or a related field. Six or more years of experience in infrastructure, platform, or HPC engineering. Hands-on experience operating GPU clusters or large-scale ML training infrastructure. Strong proficiency in Python and at least one systems language such as Go or C++. Deep understanding of distributed training, accelerator architectures, and collective communication. Experience with Kubernetes, Slurm, Ray, or similar scheduling systems for ML workloads. Strong understanding of Linux internals, networking, and high-performance storage. Experience with at least one major cloud provider’s ML infrastructure offerings. Strong software engineering practices including testing, CI/CD, and code review. Excellent communication and cross-functional collaboration skills. Preferred Qualifications Experience operating InfiniBand or RDMA networking at scale. Contributions to open-source ML infrastructure projects. Familiarity with custom orchestrators or research-grade training stacks. Exposure to frontier model training operations. Experience with FinOps for AI workloads. How to Apply Would you like to know more about this opportunity? For immediate consideration, please send your resume to boon@bvteck.com or contact us at (908) 650-6699 . Learn more about Bright Vision Technologies at www.bvteck.com . Bright Vision Technologies is an Equal Opportunity Employer. Equal Employment Opportunity (EEO) Statement Bright Vision Technologies (BV Teck) is committed to equal employment opportunity (EEO) for all employees and applicants without regard to race, color, religion, sex, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, veteran status, or any other protected status as defined by applicable federal, state, or local laws. This commitment extends to all aspects of employment, including recruitment, hiring, training, compensation, promotion, transfer, leaves of absence, termination, layoffs, and recall. BV Teck expressly prohibits any form of workplace harassment or discrimination. Any improper interference with employees' ability to perform their job duties may result in disciplinary action up to and including termination of employment.

On-siteFull-time$1.0L – $1.5L1mo agoApply
Tekfortune Inc.

Infrastructure Engineer L3

Tekfortune Inc.

Dallas, TX

Role Infrastructure Engineer L3 Location- Dallas, TX Hybrid Role Overview An experienced Infrastructure Engineer / Technical Specialist with 8+ years of experience in supporting, implementing, and troubleshooting enterprise IT infrastructure environments. Strong hands-on expertise across compute, storage, networking, hosting services, security, and cloud platforms, with the ability to contribute to infrastructure improvements, maintain operational stability, and support scalable and secure IT

On-siteFull-time$1.0L – $1.1L3w agoApply
Apex Systems

Azure Infrastructure Engineer

Apex Systems

Cumming, GA

Job#: 3047058 Job Description: Azure Infrastructure Engineer Location: REMOTE Role Overview The Enterprise Automation Team is seeking a skilled Azure Infrastructure Engineer to design and support enterprise cloud and hybrid infrastructure solutions. This role will partner with Cloud, Application, Network, and Security teams to deliver scalable, secure, and highly available solutions across both Azure and on-premises environments. The ideal candidate will have experience with Azure Kubernetes

On-siteFull-time$1.4L – $1.5L3w agoApply
Remora Jobs

Connecting talent with opportunity. Find your dream job or the perfect candidate today.

  • Browse Jobs
  • Remote Jobs
  • Companies

Popular Job Searches

More searches
Product Engineer JobsProject Engineer JobsInfrastructure Engineer JobsQA Analyst JobsBusiness Development Manager JobsSecurity Specialist JobsBranch ManagerOffice Associate JobsJobs In TampaJobs In AnchorageOperations Manager JobsImplementation Specialist JobsTechnical JobsNode.js DeveloperJava DeveloperSolutions EngineerMore searches
Remora Jobs

Connecting talent with opportunity. Find your dream job or the perfect candidate today.

Browse JobsRemote JobsCompanies
© 2026 Remora Jobs. All rights reserved.remorajobs.com