
IMC · Amsterdam
At IMC, we thrive on pushing boundaries, embracing innovation, and leveraging cutting-edge technology to stay ahead in the competitive trading landscape. We are...
At IMC, we thrive on pushing boundaries, embracing innovation, and leveraging cutting-edge technology to stay ahead in the
competitive trading landscape. We are seeking an exceptional HPC Systems Architect who will lead the charge in designing,
implementing, and optimizing our global HPC infrastructure, driving innovation, and delivering scalable solutions that empower our
advanced computational needs.
System Design and Implementation
learning workloads
scalability, efficiency, and reliability
aligns with IMC’s long-term business goals
Performance Optimization
Project Management
time, within budget, and to the highest standards
Collaboration
impact on the trading industry
translate them into scalable, high-performance solutions
consistent expansion of our global research environment
environment
About Us
IMC is a global trading firm powered by a cutting-edge research environment and a world-class technology backbone. Since 1989,
we’ve been a stabilizing force in financial markets, providing essential liquidity upon which market participants depend. Across
our offices in the US, Europe, Asia Pacific, and India, our talented quant researchers, engineers, traders, and business
operations professionals are united by our uniquely collaborative, high-performance culture, and our commitment to giving back.
From entering dynamic new markets to embracing disruptive technologies, and from developing an innovative research environment to
diversifying our trading strategies, we dare to continuously innovate and collaborate to succeed.
At IMC, we thrive on pushing boundaries, embracing innovation, and leveraging cutting-edge technology to stay ahead in the competitive trading landscape. We are seeking an exceptional HPC Systems Architect who will lead the charge in designing, implementing, and optimizing our global HPC infrastructure, driving innovation, and delivering scalable solutions that empower our advanced computational needs. Your Core Responsibilities System Design and Implementation * Architect and oversee the implementation of advanced HPC systems tailored to high-impact applications, including machine learning workloads * Integrate cutting-edge technologies to enhance computational power and capabilities * Evaluate and select best-in-class hardware and software solutions, optimizing our infrastructure for peak performance, scalability, efficiency, and reliability * Partner with global teams to establish and enforce architectural standards and best practices across HPC environments * Ensure seamless interoperability between different systems and teams, creating a cohesive, high-performance environment that aligns with IMC’s long-term business goals Performance Optimization * Develop and deploy solutions to continuously optimize system performance and resource utilization across the HPC ecosystem * Implement sophisticated monitoring and analytics tools to proactively identify and resolve performance bottlenecks Project Management * Lead and manage the full lifecycle of HPC projects, from concept through implementation, ensuring projects are delivered on time, within budget, and to the highest standards * Oversee resource allocation, budgeting, and timeline management, driving the successful execution of critical projects * Develop policies for allocating compute time, storage, and networking resources to users based on priority, need, and fairness Collaboration * Act as a trusted technical advisor to senior leadership, offering strategic insights on emerging HPC trends and their potential impact on the trading industry * Collaborate closely with quant researchers, software engineers, and traders to understand computational requirements and translate them into scalable, high-performance solutions * Partner with global teams to coordinate system design, resource allocation, and parallel project delivery, ensuring the consistent expansion of our global research environment * Evaluate success based on the HPC environment’s utilization and end-user satisfaction Your Skills and Experience * Bachelor’s, Master’s, or PhD degree in Computer Science, Electrical Engineering, or a related field * Extensive experience (typically 10+ years) in HPC architecture, system design, or a similar role within a large-scale compute environment * Deep expertise in some or all of the following areas: * HPC system design and optimization * Linux systems administration * Enterprise storage solutions (e.g., Vast, DDN, Isilon) * HPC management tools (e.g., Kubernetes, Docker, Slurm) * High-performance processors and compute offload devices (e.g., GPUs) * Low-latency network architecture, including high-speed interconnects (e.g., InfiniBand, Ethernet) * Datacenter design and optimization * AI/ML frameworks and their integration into HPC systems * Programming languages such as Python, Go, C++, or similar * Exceptional communication skills, with the ability to translate complex technical concepts for non-technical audiences * Proven ability to influence decision-making and align global teams to achieve a unified vision The Base Salary range for the role is included below. Base salary is only one component of total compensation; all full-time, permanent positions are eligible for a discretionary bonus and benefits, including paid leave and insurance. Please visit Benefits - US | IMC Trading for more comprehensive information. Salary Range $200,000—$225,000 USD About Us IMC is a global trading firm powered by a cutting-edge research environment and a world-class technology backbone. Since 1989, we’ve been a stabilizing force in financial markets, providing essential liquidity upon which market participants depend. Across our offices in the US, Europe, Asia Pacific, and India, our talented quant researchers, engineers, traders, and business operations professionals are united by our uniquely collaborative, high-performance culture, and our commitment to giving back. From entering dynamic new markets to embracing disruptive technologies, and from developing an innovative research environment to diversifying our trading strategies, we dare to continuously innovate and collaborate to succeed.
ABOUT US Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation. JOB SUMMARY We are seeking a Staff Hardware Engineer to provide advanced operational, diagnostic, and engineering support for Graphcore’s Arm-based hardware platforms across lab and data center environments. This role focuses on supporting hardware bring-up, validation, and troubleshooting of complex AI compute platforms, including server blades, racks, and rack-scale infrastructure. The successful candidate will collaborate closely with engineering, platform, and data center teams to ensure the reliability and performance of next-generation AI systems. THE TEAM The Systems Engineering and Hardware Engineering teams are responsible for enabling the bring-up, validation, and operational reliability of Graphcore’s AI infrastructure platforms. The team works closely with server engineering, firmware teams, platform architects, and data center operations to support the development, testing, and deployment of next-generation AI compute systems. This collaborative environment enables rapid problem-solving and continuous improvement of Graphcore’s hardware platforms from early development through production deployment. RESPONSIBILITIES AND DUTIES * Lead advanced break-fix troubleshooting for server blades, motherboards, power systems, and rack-scale infrastructure. * Support engineering bring-up activities, including component validation and firmware interaction testing. * Diagnose system-level failures involving thermal behavior, power anomalies, network configuration, and BIOS/BMC issues. * Collaborate with server engineering teams to perform root cause analysis and propose corrective actions or design improvements. * Support deployment and rollout of next-generation hardware platforms through structured validation and qualification cycles. * Interface with facilities and infrastructure teams to understand environmental factors impacting system reliability. * Develop and maintain standard operating procedures (SOPs), troubleshooting guides, and validation documentation. * Provide guidance and mentorship to junior technicians and engineers on troubleshooting methodologies and hardware diagnostics. * Participate in on-call rotations or off-hours support during critical engineering milestones or hardware bring-up phases. CANDIDATE PROFILE ESSENTIAL * Bachelor’s degree in Electrical Engineering, Computer Engineering, Computer Science, or related discipline. * 7 years experience with server hardware architectures and board-level debugging. * Experience analyzing system logs, hardware telemetry, and power/thermal metrics to isolate hardware failures. * Hands-on experience with HPC systems, AI compute platforms, or rack-scale infrastructure. * Strong collaboration skills and ability to work effectively in fast-paced engineering environments. * Excellent written and verbal communication skills. DESIRABLE * Experience supporting prototype or pre-production hardware bring-up. * Familiarity with data center facilities, including liquid cooling and power distribution systems. * Experience using Python, Bash, or automation tools for hardware validation or troubleshooting. * Exposure to structured failure analysis and reliability engineering methodologies. USA Benefits In addition to a competitive salary, Graphcore offers flexible working and a comprehensive benefits package designed to support your health, wellbeing and financial future. Our benefits include medical, dental and vision coverage, Flexible Spending Accounts (FSAs), Health Savings Accounts (HSAs), disability and life insurance, a 401(k) retirement plan, commuter benefits, wellness services and an Employee Assistance Programme (EAP). We welcome people of different backgrounds and experiences; we're committed to building an inclusive work environment that makes Graphcore a great home for everyone. We offer an equal opportunity process and understand that there are visible and invisible differences in all of us. We can provide a flexible approach to interview and encourage you to chat to us if you require any reasonable adjustments.
ABOUT US Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation. JOB SUMMARY We are seeking a Staff Hardware Engineer to provide advanced operational, diagnostic, and engineering support for Graphcore’s Arm-based hardware platforms across lab and data center environments. This role focuses on supporting hardware bring-up, validation, and troubleshooting of complex AI compute platforms, including server blades, racks, and rack-scale infrastructure. The successful candidate will collaborate closely with engineering, platform, and data center teams to ensure the reliability and performance of next-generation AI systems. THE TEAM The Systems Engineering and Hardware Engineering teams are responsible for enabling the bring-up, validation, and operational reliability of Graphcore’s AI infrastructure platforms. The team works closely with server engineering, firmware teams, platform architects, and data center operations to support the development, testing, and deployment of next-generation AI compute systems. This collaborative environment enables rapid problem-solving and continuous improvement of Graphcore’s hardware platforms from early development through production deployment. RESPONSIBILITIES AND DUTIES * Lead advanced break-fix troubleshooting for server blades, motherboards, power systems, and rack-scale infrastructure. * Support engineering bring-up activities, including component validation and firmware interaction testing. * Diagnose system-level failures involving thermal behavior, power anomalies, network configuration, and BIOS/BMC issues. * Collaborate with server engineering teams to perform root cause analysis and propose corrective actions or design improvements. * Support deployment and rollout of next-generation hardware platforms through structured validation and qualification cycles. * Interface with facilities and infrastructure teams to understand environmental factors impacting system reliability. * Develop and maintain standard operating procedures (SOPs), troubleshooting guides, and validation documentation. * Provide guidance and mentorship to junior technicians and engineers on troubleshooting methodologies and hardware diagnostics. * Participate in on-call rotations or off-hours support during critical engineering milestones or hardware bring-up phases. CANDIDATE PROFILE ESSENTIAL * Bachelor’s degree in Electrical Engineering, Computer Engineering, Computer Science, or related discipline. * 10 years experience with server hardware architectures and board-level debugging. * Experience analyzing system logs, hardware telemetry, and power/thermal metrics to isolate hardware failures. * Hands-on experience with HPC systems, AI compute platforms, or rack-scale infrastructure. * Strong collaboration skills and ability to work effectively in fast-paced engineering environments. * Excellent written and verbal communication skills. DESIRABLE * Experience supporting prototype or pre-production hardware bring-up. * Familiarity with data center facilities, including liquid cooling and power distribution systems. * Experience using Python, Bash, or automation tools for hardware validation or troubleshooting. * Exposure to structured failure analysis and reliability engineering methodologies. USA Benefits In addition to a competitive salary, Graphcore offers flexible working and a comprehensive benefits package designed to support your health, wellbeing and financial future. Our benefits include medical, dental and vision coverage, Flexible Spending Accounts (FSAs), Health Savings Accounts (HSAs), disability and life insurance, a 401(k) retirement plan, commuter benefits, wellness services and an Employee Assistance Programme (EAP). We welcome people of different backgrounds and experiences; we're committed to building an inclusive work environment that makes Graphcore a great home for everyone. We offer an equal opportunity process and understand that there are visible and invisible differences in all of us. We can provide a flexible approach to interview and encourage you to chat to us if you require any reasonable adjustments.