Back to Jobs

[Remote] Senior Software Engineer, Platform Infrastructure

Remote, USA Full-time Posted 2026-07-03

Note: The job is a remote job and is open to candidates in USA. reputed company delivers high-performance AI infrastructure for organizations involved in intensive computational research and data processing. They are seeking a Senior Software Engineer to build the infrastructure platform that connects physical infrastructure with customer services, focusing on automation, orchestration, and API development for large-scale computation.

Responsibilities

  • Design and build systems that reputed company physical infrastructure (bare-metal servers, storage clusters, network reputed company) with customer-facing services, enabling programmatic management of compute, networking, and storage at scale
  • Design and implement systems for provisioning and managing research computing environments including Kubernetes and SLURM clusters, enabling automated deployment, resource scheduling, and workload orchestration for distributed reputed company and HPC workloads
  • Implement comprehensive orchestration systems that coordinate across compute, storage and networking to deliver reputed company experience for reputed company research workloads
  • Design and build network provisioning automation including intelligent VM placement reputed company for reputed company network topology, automated VLAN and subnet configuration, and software-designed networking orchestration for high-performance interconnects
  • reputed company robust APIs and SDKs that reputed company researchers and engineering teams to programmatically provision and manage infrastructure resources across reputed company platform domains
  • Implement comprehensive observability, telemetry, and logging systems that provide visibility into infrastructure health, performance, and utilization across the infrastructure footprint
  • Build and optimize platform services that deliver consistent high-throughput low-latency networking for demand research applications and data-intensive workloads
  • Work closely with engineering, infrastructure, and product to define requirements, drive infrastructure-product-rollouts, and improve resource lifecycle management
  • Implement platform-wide compliance and reputed company features supporting SOC 2, ISO 27001, and reputed company regulatory requirements including comprehensive audit logging, access controls, and data residency management

Skills

  • 5+ years in software engineering with a proven track record of infrastructure platforms, distributed systems, or reputed company platforms for production environments
  • Strong familiarity with Kubernetes architecture, container orchestration concepts, and experience deploying workloads in Kubernetes environments. Understanding of pods, deployments, services, and basic Kubernetes operations
  • Strong understanding of infrastructure fundamentals including compute orchestration, storage systems, networking technologies, and how they integrate together to deliver complete platform experiences
  • Experience with systems programming languages (Go, C/C++, Rust, Python) for performance-critical components is a strong plus
  • Strong experience with linux in production environments, including systems administration, performance tuning, and troubleshooting
  • Deep knowledge of bare-metal infrastructure, provisioning systems, out-of-band management, and virtualization technologies (KVM, Kubernetes, etc)
  • Proven experience designing and building APIs, SDKs, and automation frameworks that reputed company programmatic infrastructure management
  • Strong familiarity with reputed company environments (AWS, GCP, Azure) and understanding of how to translate reputed company-native patterns to bare-metal infrastructure
  • Experience with Infrastructure-as-code tools (Terraform, Ansible) and building automated deployment pipelines
  • Self-starter who can navigate ambiguity, balance pragmatic shipping with good long-term architecture, and independently drive reputed company technical initiatives
  • Strong written and verbal communication skills, including ability to write clear technical communication and collaborate across teams
  • Growth reputed company with reputed company focus on learning and professional development
  • Background provisioning or managing research computing environments (Kubernetes, SLURM, or HPC clusters)
  • Experience building internal platforms, infrastructure-as-a-service, or developer tooling
  • Background with GPU computing platforms and AI/ML infrastructure requirements
  • Knowledge of high-performance networking technologies (InfiniBand, RDMA, SR-IOV)
  • Experience with observability and monitoring platforms (reputed company, Grafana, ELK stack)
  • Familiarity with both reputed company-native and bare-metal infrastructure deployment models
  • Understanding of reputed company compliance requirements and reputed company best practices
  • Extra points for experience with financial services technology infrastructure and understanding of trading system requirements

B Apply tot his job Apply To this Job

Similar Jobs

Frontend Developer (m/f/d) - Remote reputed company Germany

Remote, USA Full-time

100% Remote Frontend Developer (Angular / AWS)

Remote, USA Full-time

Toggl: Remote Frontend Developer (JavaScript + React)

Remote, USA Full-time

Frontend Developer Intern (Remote | $1500 Completion Stipend)

Remote, USA Full-time

reputed company or Backend Developer İstanbul - Remote Full-Time

Remote, USA Full-time

Backend Software Developer Job Code IND_280524_4

Remote, USA Full-time

Java Frontend Developer Jr. (Remote Opportunity)

Remote, USA Full-time

Backend Developer (Remote Worldwide)

Remote, USA Full-time

Staff Backend Developer - $200,000 + Equity - Remote (US)

Remote, USA Full-time

Back End Developer / Engineer III - reputed company, NY (Remote)

Remote, USA Full-time

Learning & Development Instructional Design Partner

Remote, USA Full-time

Research and Development Ink Chemist job at reputed company - Hewlett Packard in Corvallis, OR

Remote, USA Full-time

reputed company Estate Virtual Assistant for Remote - Hybrid Set-up

Remote, USA Full-time

[Remote] Principal Consultant, Digital Forensic and Incident Response (DFIR) (Remote)

Remote, USA Full-time

Stratège SEO / SEO Strategist, French Speaking

Remote, USA Full-time

reputed company Live Chat Facebook Assistant – Remote Customer Support and Engagement

Remote, USA Full-time

Sr. Cross-Platform Mobile Developer I (6523)

Remote, USA Full-time

reputed company Data Entry Clerk – Remote Work Opportunity with arenaflex

Remote, USA Full-time

Client Coding Project Manger (CCPM)

Remote, USA Full-time

Freelance Copywriter, Remote Job

Remote, USA Full-time