Skip to main content
Optiscale
Log inContact us
Solutions
HPC Cluster Design · CPU supercomputersAI Supercomputer · GPU training & inference clustersParallel Storage · Spectrum Scale · Lustre · BeeGFS · VASTInterconnect · InfiniBand · RoCE fabricsSchedulers · Slurm · KubernetesApplication Workload · Reference diagrams by workloadSizing Calculator · 5-minute sizing + 5-year TCO
Products
Luxe Vision · Unified monitoring across heterogeneous resourcesLuxe Ray · Heterogeneous resource operations & unified controlLuxe Vantage · Usage & billing built on the AI/HPC schedulerLuxe Orbit · Heterogeneous software provisioning & parallel managementLuxe Series Overview · The unified 4-layer storyRequest a Live Demo · Vision · Ray · VantageRequest a Closed PoC · Orbit — 1:1 environment setup
Services
Architecture Consulting · Workload definition → specification designImplementation & PM · Vendor integration & project managementPerformance Tuning · Architecture & configuration optimization plus code-level performance gainsManaged Operations · Multi-year operations contracts
Resources
TCO Calculator · On-premises vs. the three major cloudsSelf-Check · 8-question workload assessmentApplication Workload Library · 6 workload diagramsWhite Papers · In-depth technical white papersBlog · Tech Notes · Engineering blogNewsletter · Biweekly infrastructure updates
00Company · Careers

People who mean it when they say they'll see it through.

One company owns all five stages, from proposal to operations. Measurement, honesty, and hands-on execution are how we work — and we look for people who work that way.

01Open Positions

Now hiring.

  • Full-timeConsulting / Deployment

    Senior AI Infrastructure Engineer

    Design, deployment, and tuning of GPU clusters. Own the full path from capacity sizing to 95%+ NCCL efficiency.

    Responsibilities

    • Analyze AI training/inference workloads and size capacity
    • Design Rail-Optimized topologies and apply adaptive routing
    • Integrate NVMe-oF with Spectrum Scale and isolate IO
    • Partner with the deployment PM to validate acceptance criteria
    • Facilitate RFP authoring workshops

    Qualifications

    • 5+ years deploying or operating GPU clusters
    • Hands-on experience with NCCL / MPI / Slurm
    • Experience troubleshooting InfiniBand or RoCEv2
    • Proficient in writing technical documentation in English

    Preferred

    • PoC experience at 64+ H100 node scale
    • Experience tuning Spectrum Scale / Lustre metadata
    • Experience operating Bright Cluster Manager / Warewulf / xCAT
  • Full-timeConsulting / Deployment

    Senior Storage Engineer

    Design, deployment, and migration of Spectrum Scale-centric parallel file systems.

    Responsibilities

    • Design metadata IO isolation
    • Design AFM-based multi-site and DR architectures
    • Spectrum Scale protection group policies and HSM tiering
    • Performance benchmarking and IOR/mdtest automation

    Qualifications

    • 3+ years operating Spectrum Scale
    • Experience analyzing parallel file system IO patterns

    Preferred

    • Comparative operations experience with Lustre, WekaIO, VAST Data
    • Experience with 12PB+ scale migrations
  • Full-timeProduct

    Product Engineer (Luxe Series)

    Full-stack development of one of four product lines — Vision / Vantage / Orbit / Ray. TypeScript plus Go or Rust.

    Responsibilities

    • Slurm and Kubernetes metrics integration (Vantage)
    • OS provisioning and failover (Orbit)
    • Heterogeneous hardware event normalization (Ray)
    • Role-based dashboards (Vision)

    Qualifications

    • 3+ years hands-on with TypeScript / React
    • Understanding of Linux and distributed systems
    • Open-source contributions or published side projects

    Preferred

    • Experience with Go / Rust
    • Experience operating Prometheus / OpenSearch
    • Experience with IPMI / Redfish
02Evergreen

Always hiring.

You'll be added to our talent pool, and we'll reach out when the timing is right.

  • Full-timeConsulting / Deployment

    Deployment Project Manager (PMP)

    Schedule, quality, and issue management for multi-vendor integration projects.

    Responsibilities

    • Weekly progress management and issue tracking
    • Vendor interface integration
    • Client reporting and risk communication
    • Preparing handover packages

    Qualifications

    • PMP or equivalent certification
    • 3+ years as an IT infrastructure deployment PM

    Preferred

    • PM experience deploying HPC/AI clusters
    • Able to run projects in English
  • ContractMarketing

    Technical Writer / Content Director

    White papers, blog posts, and case studies — fact-based technical content.

    Responsibilities

    • Turn engineer interviews into white papers and blog posts (2-3 per month)
    • Write case studies (T2 anonymization and metrics tables)
    • Maintain terminology consistency

    Qualifications

    • 3+ years writing IT/engineering content
    • Experience conducting technical interviews

    Preferred

    • Familiarity with the HPC/AI domain
    • Able to produce English content
03Benefits

Why you'll like working here.

  • Senior-led

    Work alongside engineers with 25+ years of experience. A high senior ratio means deeper decisions.

  • Real impact

    30-day RFPs, 14-week deployments, 6-month tuning cycles — results measured, not claimed.

  • We build products too

    We do more than consulting. We develop and operate all four Luxe Series product lines in-house.

  • Conferences & seminars

    Attend and present at conferences like SC, ISC HP, and KOSA.

  • Training & certification

    The company covers NVIDIA / IBM / SchedMD certification costs.

  • Remote & flexible

    Autonomous outside deployment sites. Office 2-3 days a week recommended.

How to apply

Send your résumé and portfolio (or GitHub) to [email protected]. We reply personally.

Get in touch first
Careers — Optiscale | Optiscale