Engineering Manager, Infrastructure Engineering
Company Name
**
Livingston, NJ / New York, NY / Sunnyvale, CA / Bellevue, WA
Basic
Posted 17 days ago
CoreWeave is The Essential Cloud for AI™. Built for pioneers by pioneers, CoreWeave delivers a platform of technology, tools, and teams that enables innovators to build and scale AI with confidence. Trusted by leading AI labs, startups, and global enterprises, CoreWeave combines superior infrastructure performance with deep technical expertise to accelerate breakthroughs and turn compute into capability. Founded in 2017, CoreWeave became a publicly traded company (Nasdaq: CRWV) in March 2025. Learn more at www.coreweave.com .
About the role:
CoreWeave is seeking an experienced, people-focused Engineering Manager to lead our MetalDev RAS (Reliability, Availability & Serviceability) team within Hardware Engineering Dev (Metal Dev). In this first-line management role, you will build, coach, and grow a team of infrastructure and site reliability engineers responsible for the reliability, availability, and serviceability of the services that manage CoreWeave's bare-metal infrastructure at scale.
You will own both the health of your team and the health of the systems they operate — setting technical direction, driving operational excellence, and partnering with cross-functional teams, external vendors, and other stakeholders to deliver highly performant and resilient infrastructure solutions. You will balance hands-on technical leadership with the people leadership needed to develop engineers and sustain a strong on-call and incident-management culture.
The MetalDev RAS team is responsible for the software and infrastructure that keeps CoreWeave's fleet of GPU servers reliable, available, and serviceable at massive scale. This team sits at the intersection of software engineering, infrastructure, and hardware, building the automation and operational systems that power one of the world's fastest-growing AI clouds.
What You’ll Do
People Leadership & Team Development
Hire, coach, and grow a team of infrastructure/SRE engineers, fostering a culture of ownership and reliability.
Run 1:1s, performance reviews, and career development; set clear goals and hold the team accountable.
Maintain a healthy, sustainable on-call culture and balanced team workload.
Reliability, Availability & Serviceability (RAS) Leadership
Own the reliability, availability, and serviceability of the MetalDev RedFish services in production environment, ensuring that the services are performing optimally.
Define and drive KPIs, SLAs, and SLOs for the team, ensuring alignment with organizational reliability objectives.
Champion system observability and health using tools like Prometheus and Grafana to proactively detect performance bottlenecks before they become incidents.
Set the automation strategy that streamlines incident detection and recovery and reduces manual toil across the server hardware lifecycle.
Drive the team to continuously reduce on-call queries and incidents over time through systemic, long-term improvements.
Incident Management & Operational Excellence
Establish and improve incident response processes, runbooks, and RCA/PIR practices.
Lead communication to stakeholders during major incidents and drive systemic fixes over repeat firefighting.
Oversee the full server hardware lifecycle through automation, dashboards, and CI/CD pipelines.
Cross functional collaboration
Collaborate with engineering teams across the organization on platform reliability, resilience improvements, and disaster recovery.
Represent the team in planning and prioritization, managing dependencies with partner teams and external vendors.
Engage with upstream communities (e.g., Go and Redfish-based services) and guide the team's technical standards for tooling, automation, and documentation.
Who You Are:
3+ years of engineering management experience leading and growing teams of software / infrastructure / SRE engineers (first-line management), on top of a strong individual-contributor background.
7+ years of combined experience in cloud operations, site reliability engineering (SRE), infrastructure, or related technical roles.
Strong understanding of cloud platforms and K8S, and of cloud and bare-metal infrastructure.
Solid grounding in incident management practices and frameworks (e.g., ITIL, SRE best practices), including on-call, RCA, and PIR processes.
Experience leading teams that develop software in Go or (comparable systems languages) with the ability to participate in architecture and code reviews.
Experience with observability tooling such as Prometheus and Grafana.
Track record of defining and driving reliability metrics (SLAs / SLOs / KPIs) and operational improvements at scale.
Excellent communication, documentation, and stakeholder-management skills, with strong analytical and problem-solving abilities.
Experience supporting production services on an on-call rotation, and a demonstrated ability to build a healthy on-call and operational culture within a team.
Preferred Qualifications
Experience managing bare-metal, hardware, or hardware-adjacent infrastructure teams.
Familiarity with Redfish, BMC, or server lifecycle management technologies.
Experience scaling teams, tooling, and processes in a high-growth environment.
The base salary range for this role is $182,000 to $242,000. The starting salary will be determined based on job-related knowledge, skills, experience, and market location. We strive for both market alignment and internal equity when determining compensation. In addition to base salary, our total rewards package includes a discretionary bonus, equity awards, and a comprehensive benefits program (all based on eligibility).
What We Offer
The range we’ve posted represents the typical compensation range for this role. To determine actual compensation, we review the market rate for each candidate which can include a variety of factors. These include qualifications, experience, interview performance, and location.
In addition to a competitive salary, we offer a variety of benefits to support your needs. The benefits below reflect our US-based offerings for full-time employees; for roles in other locations, benefits vary and are shared during the hiring process. These include:
Medical, dental, and vision insurance - 100% paid for by CoreWeave
Company-paid Life Insurance
Voluntary supplemental life insurance
Short and long-term disability insurance
Flexible Spending Account
Health Savings Account
Tuition Reimbursement
Ability to Participate in Employee Stock Purchase Program (ESPP)
Mental Wellness Benefits through Spring Health
Family-Forming support provided by Carrot
Paid Parental Leave
Flexible, full-service childcare support with Kinside
401(k) with a generous employer match
Flexible PTO
Catered lunch each day in our office and data center locations
A casual work environment
A work culture focused on innovative disruption
California Applicants
California Consumer Privacy Act
Equal Opportunity & Accommodations
CoreWeave is an equal opportunity employer, committed to fostering an inclusive and supportive workplace. All qualified applicants and candidates will receive consideration for employment without regard to race, color, religion, sex, disability, age, sexual orientation, gender identity, national origin, veteran status, or genetic information.
As part of this commitment and consistent with the Americans with Disabilities Act (ADA) , CoreWeave will ensure that qualified applicants and candidates with disabilities are provided reasonable accommodations for the hiring process, unless such accommodation would cause an undue hardship. If reasonable accommodation is needed, please contact: careers@coreweave.com .
Export Control Compliance
This position requires access to export controlled information. To conform to U.S. Government export regulations applicable to that information, applicant must either be (A) a U.S. person, defined as a (i) U.S. citizen or national, (ii) U.S. lawful permanent resident (green card holder), (iii) refugee under 8 U.S.C. § 1157, or (iv) asylee under 8 U.S.C. § 1158, (B) eligible to access the export controlled information without a required export authorization, or (C) eligible and reasonably likely to obtain the required export authorization from the applicable U.S. government agency. CoreWeave may, for legitimate business reasons, decline to pursue any export licensing process.
AI Engineer (US)
Company Name
**
New York, USA; Philadelphia, Pennsylvania, United States; San Francisco, California, United States
Basic
Posted 18 days ago
We are investing in agentic AI and need a Senior AI Engineer to lead the design and delivery of these systems. This is a foundational hire: you will own both the agent-facing workstreams — pipelines, orchestration, conversational interfaces — and the underlying context layer that makes them reliable, including memory management, knowledge graph integration, and retrieval infrastructure.
You will work closely with data engineers, project leads, and client stakeholders, and play a key role in shaping how Lynx builds and ships AI solutions at scale.
What This Involves:
Lead the architecture and delivery of agentic AI systems end-to-end: agents, orchestration, tool use, and multi-step reasoning workflows.
Own the context layer: design and implement memory architectures (episodic, semantic, working memory) and integrate GraphRAG and knowledge graph retrieval into agentic pipelines.
Build robust RAG systems — including vector retrieval, graph traversal, and hybrid search — and ensure retrieval quality through evaluation frameworks.
Translate client requirements into technical designs, presenting approaches and trade-offs to both technical and non-technical stakeholders.
Define standards and reusable patterns for agentic AI development that other engineers at Lynx can build on.
Set up observability, evaluation, and monitoring pipelines to ensure AI systems perform correctly in production.
Requirements:
5–8 years of software or ML engineering experience, with at least 2–3 years building LLM-based or agentic AI systems in production.
Deep hands-on experience with agentic frameworks (LangChain, LlamaIndex, AutoGen, CrewAI, or similar) and LLM APIs (OpenAI, Anthropic, etc.).
Strong understanding of agent design patterns: ReAct, planning loops, tool use, multi-agent coordination, and memory architectures.
Practical experience with GraphRAG or knowledge graph-based retrieval (e.g., Neo4j, Microsoft GraphRAG) and vector databases (Pinecone, Weaviate, Qdrant, etc.).
Proficiency in Python and solid software engineering fundamentals: APIs, testing, CI/CD, containerisation (Docker/Kubernetes).
Experience working in a consulting or client-facing environment — comfortable presenting technical approaches and adapting to ambiguous requirements.
Strong written and verbal communication skills across distributed, cross-functional teams.
Key Competencies:
Stakeholder Mentality: Treats the company's and client’s goals as their own and is genuinely motivated by its success.
Organisational Excellence : Manages time and priorities effectively, ensuring tasks are completed accurately and on time even in a fast-paced environment.
Discretion & Integrity : Handles sensitive and confidential information with professionalism and sound judgement.
Problem Solving : Approaches challenges proactively and with a solution-oriented mindset, taking initiative rather than waiting to be directed.
Collaboration : A team player who builds strong working relationships and communicates effectively with colleagues across all levels.
Why You Will Love It Here:
Work on real-world AI and advanced analytics solutions with measurable business impact.
Collaborate with a global team of engineers and data scientists.
Exposure to diverse industries, modern cloud platforms, and cutting-edge AI technologies.
A collaborative culture that values real outcomes.
Rapid learning opportunities and diverse challenges.
Flat organisational hierarchy with high visibility and accessibility to our leaders.
Senior Engineer, Compute Services (Kubernetes, Bare Metal)
Company Name
**
Livingston, NJ / New York, NY / Sunnyvale, CA / Bellevue, WA
Basic
Posted 18 days ago
CoreWeave is The Essential Cloud for AI™. Built for pioneers by pioneers, CoreWeave delivers a platform of technology, tools, and teams that enables innovators to build and scale AI with confidence. Trusted by leading AI labs, startups, and global enterprises, CoreWeave combines superior infrastructure performance with deep technical expertise to accelerate breakthroughs and turn compute into capability. Founded in 2017, CoreWeave became a publicly traded company (Nasdaq: CRWV) in March 2025. Learn more at www.coreweave.com .
What You’ll Do
As a Senior Engineer in Compute Services, you will be responsible for building fault-tolerant and reliable infrastructure to support both our internal processes and our customer platform. If you’re passionate about GitOps, KubeOps, DevOps—really, all the ops—this role could be a great fit for you!
• Design, develop, and maintain automated tooling to provision Kubernetes control planes on bare-metal
• Use Python, Golang, and Bash to create tooling and go operators
• Perform day 2 lifecycle tasks and maintenance on running clusters
• Identify gaps and implement fault-tolerant architectures
• Optimize reliability using the Grafana ecosystem
• Design automated testing to validate build quality and stability
• Participate in an on-call rotation every two months serving as point of contact
Who You Are
Investing in our people is one of our top priorities, and we value candidates who can bring their diversified experiences to our teams. Here are some qualities we’ve found compatible with our team. We'd love to talk about whether this aligns with your experience and interests and what you’re excited to work on next.
• Proven experience provisioning Kubernetes using tools such as kubeadm, Cluster API, Kubeception, Kubespray, or similar
• Demonstrated ability debugging complex kubernetes cluster issues and carrying out upgrades
• Proficiency in Golang, Bash, and Python
• Advanced Linux OS troubleshooting skills
• Extensive experience with Ansible
• Advanced DevOps experience (e.g., GitLab CI, GitHub Actions)
• Demonstrated ability to collaborate effectively on shared codebases
• Excellent documentation skills and high attention to detail
• Strong analytical and problem-solving abilities
• Experience participating in an on-call rotation to support production services
Preferred Qualifications
Bare-metal OS provisioning experience
Kubernetes operator coding experience
Advanced Linux networking expertise
AWX/Ansible tower knowledge
The base salary range for this role is $153,000 to $204,000. The starting salary will be determined based on job-related knowledge, skills, experience, and market location. We strive for both market alignment and internal equity when determining compensation. In addition to base salary, our total rewards package includes a discretionary bonus, equity awards, and a comprehensive benefits program (all based on eligibility).
What We Offer
The range we’ve posted represents the typical compensation range for this role. To determine actual compensation, we review the market rate for each candidate which can include a variety of factors. These include qualifications, experience, interview performance, and location.
In addition to a competitive salary, we offer a variety of benefits to support your needs. The benefits below reflect our US-based offerings; for roles in other locations, benefits vary and are shared during the hiring process. These include:
Medical, dental, and vision insurance - 100% paid for by CoreWeave
Company-paid Life Insurance
Voluntary supplemental life insurance
Short and long-term disability insurance
Flexible Spending Account
Health Savings Account
Tuition Reimbursement
Ability to Participate in Employee Stock Purchase Program (ESPP)
Mental Wellness Benefits through Spring Health
Family-Forming support provided by Carrot
Paid Parental Leave
Flexible, full-service childcare support with Kinside
401(k) with a generous employer match
Flexible PTO
Catered lunch each day in our office and data center locations
A casual work environment
A work culture focused on innovative disruption
California Applicants
California Consumer Privacy Act
Equal Opportunity & Accommodations
CoreWeave is an equal opportunity employer, committed to fostering an inclusive and supportive workplace. All qualified applicants and candidates will receive consideration for employment without regard to race, color, religion, sex, disability, age, sexual orientation, gender identity, national origin, veteran status, or genetic information.
As part of this commitment and consistent with the Americans with Disabilities Act (ADA) , CoreWeave will ensure that qualified applicants and candidates with disabilities are provided reasonable accommodations for the hiring process, unless such accommodation would cause an undue hardship. If reasonable accommodation is needed, please contact: careers@coreweave.com .
Export Control Compliance
This position requires access to export controlled information. To conform to U.S. Government export regulations applicable to that information, applicant must either be (A) a U.S. person, defined as a (i) U.S. citizen or national, (ii) U.S. lawful permanent resident (green card holder), (iii) refugee under 8 U.S.C. § 1157, or (iv) asylee under 8 U.S.C. § 1158, (B) eligible to access the export controlled information without a required export authorization, or (C) eligible and reasonably likely to obtain the required export authorization from the applicable U.S. government agency. CoreWeave may, for legitimate business reasons, decline to pursue any export licensing process.
Senior Software Engineer, Infrastructure Engineering
Company Name
**
Livingston, NJ / New York, NY / Sunnyvale, CA / Bellevue, WA
Basic
Posted 18 days ago
CoreWeave is The Essential Cloud for AI™. Built for pioneers by pioneers, CoreWeave delivers a platform of technology, tools, and teams that enables innovators to build and scale AI with confidence. Trusted by leading AI labs, startups, and global enterprises, CoreWeave combines superior infrastructure performance with deep technical expertise to accelerate breakthroughs and turn compute into capability. Founded in 2017, CoreWeave became a publicly traded company (Nasdaq: CRWV) in March 2025. Learn more at www.coreweave.com .
CoreWeave is seeking a highly skilled and motivated Sr. Infrastructure Engineer to join our Hardware Engineering Dev team (Metal Dev). Reporting to the Engineering Manager for Hardware Engineering Dev, you will play a crucial part in the development, deployment, and monitoring of services that manage our bare-metal infrastructure. You will collaborate closely with cross-functional teams, external vendors, and other stakeholders to ensure the successful delivery of highly performant and reliable infrastructure solutions.
Key Responsibilities:
Incident Management & Support:
Lead incident response efforts by identifying and resolving service disruptions quickly, while coaching other junior team members through resolution.
Lead the documentation of incidents, conduct in-depth root cause analysis (RCA), and drive post-incident reviews (PIRs) to identify systemic issues. Implement long term improvements that would prevent service degradation.
Own the development and continuous improvement of incident response playbooks ensuring preparedness for a wide range of failure scenarios.
Clearly communicate efforts during incidents to the management, stakeholders and the cross functional teams, during an incident. Keep clear records of incident activities.
Master clear understanding of various services on how they work in production as well as build through knowledge of the internals of these services and how they interact with the entire stack.
Operational Support & Reliability
Build a strategy around making our core services perform at its best at scale. This includes improvements to the services for robustness as well as supportability in production.
Own system observability and health leveraging tools like Prometheus and Grafana, to proactively detect performance bottlenecks and prevent incidents.
Lead automation efforts to streamline incident detection and recovery, minimizing manual intervention.
Define and drive KPIs and SLAs for incident management and ensuring alignment with the organizational reliability objectives.
Collaborate with engineers across teams to improve platform reliability, resilience improvements, and disaster recovery.
Collaborate with upstream communities, including Go and Redfish-based services.
Design and implement solutions to build operational efficiency and stability.
Document hardware automation workflows and processes.
Create CI/CD pipelines.
Ensure smooth operation of all aspects of the server hardware lifecycle, from provisioning to end-of-life, by troubleshooting bugs, automating common tasks, and documenting processes
Partner with the Fleet Operations Team to design scalable tooling and processes that enables self-service and reduction in escalation overhead.
Build out dashboards and alerts to make efficient operational troubleshooting.
Participate in on-call rotation as well as triage issues that are posted in support channels on an ongoing basis.
This role is responsible for reducing on-call queries and incidents over time.
Qualification:
7+ years of experience in cloud operations, site reliability engineering (SRE), or related technical roles.
Understanding of cloud platforms (e.g., Kubernetes, AWS, GCP) and basic knowledge of cloud infrastructure.
Familiarity with incident management practices and frameworks (e.g., ITIL, SRE best practices).
Proficiency with Go.
Prior experience with Prometheus / Grafana.
Previous experience deploying containerized applications using Kubernetes.
Excellent documentation skills and attention to detail.
Strong analytical and problem-solving abilities.
Served on an on-call rotation supporting production services.
Why CoreWeave?
At CoreWeave, we work hard, have fun, and move fast! We’re in an exciting stage of hyper-growth that you will not want to miss out on. We’re not afraid of a little chaos, and we’re constantly learning. Our team cares deeply about how we build our product and how we work together, which is represented through our core values:
Be Curious at Your Core
Act Like an Owner
Empower Employees
Deliver Best-in-Class Client Experiences
Achieve More Together
We support and encourage an entrepreneurial outlook and independent thinking. We foster an environment that encourages collaboration and provides the opportunity to develop innovative solutions to complex problems. As we get set for take off, the growth opportunities within the organization are constantly expanding. You will be surrounded by some of the best talent in the industry, who will want to learn from you, too. Come join us!
The base salary range for this role is $153,000 to $242,000. The starting salary will be determined based on job-related knowledge, skills, experience, and market location. We strive for both market alignment and internal equity when determining compensation. In addition to base salary, our total rewards package includes a discretionary bonus, equity awards, and a comprehensive benefits program (all based on eligibility).
What We Offer
The range we’ve posted represents the typical compensation range for this role. To determine actual compensation, we review the market rate for each candidate which can include a variety of factors. These include qualifications, experience, interview performance, and location.
In addition to a competitive salary, we offer a variety of benefits to support your needs. The benefits below reflect our US-based offerings; for roles in other locations, benefits vary and are shared during the hiring process. These include:
Medical, dental, and vision insurance - 100% paid for by CoreWeave
Company-paid Life Insurance
Voluntary supplemental life insurance
Short and long-term disability insurance
Flexible Spending Account
Health Savings Account
Tuition Reimbursement
Ability to Participate in Employee Stock Purchase Program (ESPP)
Mental Wellness Benefits through Spring Health
Family-Forming support provided by Carrot
Paid Parental Leave
Flexible, full-service childcare support with Kinside
401(k) with a generous employer match
Flexible PTO
Catered lunch each day in our office and data center locations
A casual work environment
A work culture focused on innovative disruption
California Applicants
California Consumer Privacy Act
Equal Opportunity & Accommodations
CoreWeave is an equal opportunity employer, committed to fostering an inclusive and supportive workplace. All qualified applicants and candidates will receive consideration for employment without regard to race, color, religion, sex, disability, age, sexual orientation, gender identity, national origin, veteran status, or genetic information.
As part of this commitment and consistent with the Americans with Disabilities Act (ADA) , CoreWeave will ensure that qualified applicants and candidates with disabilities are provided reasonable accommodations for the hiring process, unless such accommodation would cause an undue hardship. If reasonable accommodation is needed, please contact: careers@coreweave.com .
Export Control Compliance
This position requires access to export controlled information. To conform to U.S. Government export regulations applicable to that information, applicant must either be (A) a U.S. person, defined as a (i) U.S. citizen or national, (ii) U.S. lawful permanent resident (green card holder), (iii) refugee under 8 U.S.C. § 1157, or (iv) asylee under 8 U.S.C. § 1158, (B) eligible to access the export controlled information without a required export authorization, or (C) eligible and reasonably likely to obtain the required export authorization from the applicable U.S. government agency. CoreWeave may, for legitimate business reasons, decline to pursue any export licensing process.
Senior Applications Analyst
Company Name
**
Remote Indiana, Remote, IN, United States; Remote Florida, Florida, FL, United States; Remote Texas, Remote, TX, United States; Remote Michigan, Farmington HIlls, MI, United States; Remote OH, Remote
Basic
Posted 18 days ago
First Merchants Bank is seeking a Senior Applications Analyst to join our team. This position will be responsible for eliciting needs of the stakeholders, analyzing business requirements, identifying business opportunities, preparing documentation, and evaluating risk for assigned applications. Ensure assigned applications work reliably for the business and are kept current per the vendor requirements; maintain, troubleshoot, and support systems daily.
Junior FPGA Developer
Company Name
**
Chicago, IL, United States
Basic
Posted 19 days ago
company
About us
Edgehog Trading is a proprietary trading firm specializing in electronic options market making. We take a technology-driven approach, designing and operating automated, scalable systems to provide liquidity across markets.
Our team spans trading, engineering, and business operations, working together to build and support the systems that power the firm. We emphasize data-driven decision making, rigorous problem solving, and continuous improvement to navigate complex and evolving markets.
We operate in a highly collaborative environment where ideas can move quickly from concept to implementation, and where individuals are empowered to take ownership and contribute directly to the firm’s growth.
role
What you'll do:
Design, implement, and verify RTL logic in Verilog or SystemVerilog targeting ultra-low-latency trading applications Develop simulation testbenches and functional verification environments to validate FPGA designs before hardware deployment Work closely with senior FPGA engineers and the trading infrastructure team to understand latency requirements and translate them into hardware design decisions Run synthesis, place-and-route, and static timing analysis; iterate on designs to meet strict timing constraints Participate in hardware bring-up and on-board debugging, using waveform analysis and other diagnostic tools Use AI tools to accelerate your development workflow — RTL review, testbench generation, debugging, and documentation Grow your understanding of market data protocols (Ethernet, PCIe, exchange feed formats) as they relate to the systems you build Qualifications and skills:
BS or MS in Electrical Engineering, Computer Engineering, or a related field Hands-on experience with RTL design in Verilog or SystemVerilog — coursework, personal projects, or internship experience all count Familiarity with FPGA simulation tools (e.g., ModelSim, QuestaSim, Vivado, Quartus) and the synthesis-to-deployment flow Understanding of digital logic fundamentals: state machines, FIFOs, clock domain crossing, timing constraints Comfort working in a Linux environment; Python, basic C or C++ is a plus Ability and eagerness to incorporate AI tools into your development and debugging workflow A strong desire to learn the HFT domain — no prior finance knowledge required Strong analytical instincts: when something behaves unexpectedly in hardware, you dig until you find it Why Edgehog:
Small team advantage: Direct access to founders and senior team members from day one Ownership early: Manage real P&L and make meaningful impact within your first year Cutting-edge tech: Work with our proprietary models and low-latency trading systems built in-house Tight feedback loops: Weekly 1-on-1s with your mentor, quarterly reviews with leadership Chicago-based: Affordable cost of living, vibrant trading community Benefits:
Comprehensive health, dental, and vision insurance with premiums 100% covered by the firm 401(k) with a 4% company match Unlimited paid time off and sick leave Free lunch, coffee, drinks, and snacks Commuter benefits Monthly happy hours and annual team events
The base salary range for this position is listed below. Base salary represents just one part of overall compensation; all full-time, permanent roles are eligible for a discretionary bonus and benefits, including items in the above list.
The base salary for this role is 100,000 - 175,000 USD per year
Senior Storage Engineer, File & Block
Company Name
**
Livingston, NJ / New York, NY / Sunnyvale, CA / San Francisco, CA / Bellevue, WA
Basic
Posted 19 days ago
CoreWeave is The Essential Cloud for AI™. Built for pioneers by pioneers, CoreWeave delivers a platform of technology, tools, and teams that enables innovators to build and scale AI with confidence. Trusted by leading AI labs, startups, and global enterprises, CoreWeave combines superior infrastructure performance with deep technical expertise to accelerate breakthroughs and turn compute into capability. Founded in 2017, CoreWeave became a publicly traded company (Nasdaq: CRWV) in March 2025. Learn more at www.coreweave.com .
What You’ll Do:
We're looking for a Senior Storage Engineer, File & Block to help build and operate the file and block storage services at the heart of CoreWeave's Storage team — the high-performance, multi-tenant storage that our largest AI training and inference workloads and internal stateful services run on. You'll help turn CoreWeave's storage from a follower of GPU growth into a multiplier of it: owning the data path and control plane for both file and block, running them on our own hardware fleet, and shipping the customer-facing capabilities that unblock petabyte-scale deals. You'll work closely with compute, platform, and infrastructure teams to make storage fast, reliable, and effortless to consume at scale.
About the role:
Design, build, and operate highly scalable, multi-tenant file and block storage — from the data path (NFS for file; durable, high-performance block volumes) to the control plane that governs tenancy, provisioning, and data protection — running natively on CoreWeave's storage fleet and delivered through Kubernetes.
Own the file system data path : NFS hot-path reliability, IO resilience, and performance — including the small-file / IOPS-heavy workloads that AI pipelines depend on, and the throughput that keeps GPUs fed.
Build a native block storage service : durable, high-performance block volumes for customer workloads and internal stateful services — the third leg of the file/object/block triad, running on our own data path and shared storage hardware.
Deliver both file and block as first-class Kubernetes citizens: design and ship a backend-neutral CSI driver with dynamic provisioning, resize, snapshots, and RWO block volumes, backed by a public provisioning SLO — reusing shared control-plane and CSI patterns across file and block rather than rebuilding per service.
Build the control plane : multi-tenant isolation, quota, QoS, snapshots, key management (BYOK), audit, and lifecycle policy — the capabilities that unblock regulated and enterprise customers (e.g., encryption in transit, HIPAA controls).
Improve the reliability, durability, and observability of the storage stack; partner with operations to monitor, analyze, and optimize using telemetry, metrics, and dashboards to improve performance, latency, and resilience.
Work cross-functionally with platform, product, and infrastructure teams to deliver seamless storage across the stack, including tiering cold data to lower-cost object storage while preserving access.
Share your knowledge and mentor other engineers on best practices in building distributed, high-performance systems.
Work with technologies such as RDMA, GPU Direct Storage, RoCE, InfiniBand, SPDK, and distributed filesystems to optimize storage performance and efficiency.
Participate in efforts to improve the reliability, durability, and observability of our storage stack.
Collaborate with operations teams to monitor, analyze, and optimize storage systems using telemetry, metrics, and dashboards to improve performance, latency, and resilience.
Work cross-functionally with platform, product, and infrastructure teams to deliver seamless storage capabilities across the stack.
Share your knowledge and mentor other engineers on best practices in building distributed, high-performance systems.
Who You Are:
Bachelor's or Master's degree in Computer Science, Engineering, or a related field.
6–10 years building storage systems, distributed systems, or infrastructure services.
Strong hands-on experience with distributed or networked file systems and/or block storage in production — e.g., NFS/POSIX file semantics at scale, and/or block volume services (durability, snapshots, replication).
Depth in the storage data path: IO performance, caching, small-file and metadata-heavy workloads, crash consistency, and resilience under load.
Experience with multi-tenant services or storage control planes — provisioning, tenancy/quota, and data-protection features (snapshots, encryption, key management).
Proficiency in a systems/back-end language — Go strongly preferred (C or Rust a plus).
Experience building on Kubernetes — controllers/operators, CRDs, and CSI (dynamic provisioning, resize, snapshots, RWO volumes).
Familiarity with distributed databases (e.g., CockroachDB), workflow orchestration (e.g., Temporal), and gRPC/Protobuf service design.
Ideally, experience with distributed or parallel storage stacks such as Ceph/RBD, Lustre, GPFS/Spectrum Scale, BeeGFS, WEKA, VAST, or DAOS.
Familiarity with storage observability tools and telemetry pipelines (e.g., ClickHouse, Prometheus, Grafana).
Bonus: exposure to high-performance data-path technologies (RDMA, GPUDirect Storage, RoCE, InfiniBand, SPDK).
Strong debugging and problem-solving skills in distributed, high-performance environments.
Clear communicator, able to work collaboratively across teams and share technical insights effectively.
Build a native block storage service : durable, high-performance block volumes for customer workloads and internal stateful services — the third leg of the file/object/block triad, running on our own data path and shared storage hardware.
Deliver both file and block as first-class Kubernetes citizens: design and ship a backend-neutral CSI driver with dynamic provisioning, resize, snapshots, and RWO block volumes, backed by a public provisioning SLO — reusing shared control-plane and CSI patterns across file and block rather than rebuilding per service.
Build the control plane : multi-tenant isolation, quota, QoS, snapshots, key management (BYOK), audit, and lifecycle policy — the capabilities that unblock regulated and enterprise customers (e.g., encryption in transit, HIPAA controls).
Improve the reliability, durability, and observability of the storage stack; partner with operations to monitor, analyze, and optimize using telemetry, metrics, and dashboards to improve performance, latency, and resilience.
Work cross-functionally with platform, product, and infrastructure teams to deliver seamless storage across the stack, including tiering cold data to lower-cost object storage while preserving access.
Share your knowledge and mentor other engineers on best practices in building distributed, high-performance systems.
Wondering if you’re a good fit? We believe in investing in our people, and value candidates who can bring their own diversified experiences to our teams – even if you aren't a 100% skill or experience match. Here are a few qualities we’ve found compatible with our team. If some of this describes you, we’d love to talk.
Why CoreWeave?
At CoreWeave, we work hard, have fun, and move fast! We’re in an exciting stage of hyper-growth that you will not want to miss out on. We’re not afraid of a little chaos, and we’re constantly learning. Our team cares deeply about how we build our product and how we work together, which is represented through our core values:
Be Curious at Your Core
Act Like an Owner
Empower Employees
Deliver Best-in-Class Client Experiences
Achieve More Together
We support and encourage an entrepreneurial outlook and independent thinking. We foster an environment that encourages collaboration and provides the opportunity to develop innovative solutions to complex problems. As we get set for take off, the growth opportunities within the organization are constantly expanding. You will be surrounded by some of the best talent in the industry, who will want to learn from you, too. Come join us!
"The base salary range for this role is $165,000 to $242,000. The starting salary will be determined based on job-related knowledge, skills, experience, and market location. We strive for both market alignment and internal equity when determining compensation. In addition to base salary, our total rewards package includes a discretionary bonus, equity awards, and a comprehensive benefits program (all based on eligibility)."
What We Offer
The range we’ve posted represents the typical compensation range for this role. To determine actual compensation, we review the market rate for each candidate which can include a variety of factors. These include qualifications, experience, interview performance, and location.
In addition to a competitive salary, we offer a variety of benefits to support your needs. The benefits below reflect our US-based offerings; for roles in other locations, benefits vary and are shared during the hiring process. These include:
Medical, dental, and vision insurance - 100% paid for by CoreWeave
Company-paid Life Insurance
Voluntary supplemental life insurance
Short and long-term disability insurance
Flexible Spending Account
Health Savings Account
Tuition Reimbursement
Ability to Participate in Employee Stock Purchase Program (ESPP)
Mental Wellness Benefits through Spring Health
Family-Forming support provided by Carrot
Paid Parental Leave
Flexible, full-service childcare support with Kinside
401(k) with a generous employer match
Flexible PTO
Catered lunch each day in our office and data center locations
A casual work environment
A work culture focused on innovative disruption
California Applicants
California Consumer Privacy Act
Equal Opportunity & Accommodations
CoreWeave is an equal opportunity employer, committed to fostering an inclusive and supportive workplace. All qualified applicants and candidates will receive consideration for employment without regard to race, color, religion, sex, disability, age, sexual orientation, gender identity, national origin, veteran status, or genetic information.
As part of this commitment and consistent with the Americans with Disabilities Act (ADA) , CoreWeave will ensure that qualified applicants and candidates with disabilities are provided reasonable accommodations for the hiring process, unless such accommodation would cause an undue hardship. If reasonable accommodation is needed, please contact: careers@coreweave.com .
Export Control Compliance
This position requires access to export controlled information. To conform to U.S. Government export regulations applicable to that information, applicant must either be (A) a U.S. person, defined as a (i) U.S. citizen or national, (ii) U.S. lawful permanent resident (green card holder), (iii) refugee under 8 U.S.C. § 1157, or (iv) asylee under 8 U.S.C. § 1158, (B) eligible to access the export controlled information without a required export authorization, or (C) eligible and reasonably likely to obtain the required export authorization from the applicable U.S. government agency. CoreWeave may, for legitimate business reasons, decline to pursue any export licensing process.
WMS Solutions Architect
Company Name
**
Beverly Hills, California, United States; San Ramon Office, San Ramon, CA (LO-SF)
Basic
Posted 25 days ago
WHY JOIN ALO?
Mindful movement. It’s at the core of why we do what we do at ALO—it’s our calling. Because mindful movement in the studio leads to better living. It changes who yogis are off the mat, making their lives and their communities better. That’s the real meaning of studio-to-street: taking the consciousness from practice on the mat and putting it into practice in life.
OVERVIEW
We are seeking an experienced Solutions Architect to lead the design, governance, and delivery of system integrations supporting our warehouse transformation initiative, including the implementation of a Tier-1 Warehouse Management system within a highly automated apparel distribution center.
This role will be responsible for defining and governing the end-to-end integration architecture across the Warehouse Management System (WMS), Order Management System (OMS), warehouse automation platforms, shipping solutions, and enterprise systems. The Integration Architect will serve as the primary technical authority for interface design, data flows, API strategy, event processing, monitoring, resiliency, and production support readiness.
The ideal candidate possesses deep experience delivering large-scale supply chain integrations and understands the operational realities of warehouse execution, order fulfillment, transportation, inventory management, and automation systems.
Key Responsibilities
Integration Strategy & Architecture
Define the enterprise integration strategy supporting WMS and related supply chain platforms including ERP, order management, carrier and parcel systems, warehouse automation platforms (sortation, ASRS, AMR, etc.), and enterprise reporting and analytics.
Design scalable, resilient, and secure integrations using modern API and event-driven architectural patterns.
Establish architectural standards, design principles, and interface governance processes.
Develop future-state architecture diagrams, data flow models, and integration roadmaps.
Ensure solutions align with enterprise architecture, cybersecurity, and cloud platform standards.
Responsibilities include:
Inventory synchronization
Order lifecycle processing
Shipment execution
ASN processing
Master data integration
Warehouse execution events
Transportation and carrier communication
Exception handling and recovery workflows
API & Event Architecture
Define API standards and integration design patterns.
Design REST, webhook, and event-based integrations.
Establish canonical data models and message standards.
Define system contracts and data ownership models.
Create standards for versioning, monitoring, and lifecycle management.
Delivery Leadership
Partner cross-functionally with internal and vendor technology and business stakeholders.
Review technical designs and integration specifications.
Provide architectural oversight for development teams.
Participate in solution design workshops and business process reviews.
Lead architecture reviews, design approvals, and go-live readiness assessments.
Operational Support & Reliability
Define monitoring, alerting, and support models for all integrations.
Establish error handling and retry frameworks.
Create failover and business continuity strategies.
Support performance testing and volume testing activities.
Ensure operational support teams are equipped for post-go-live success.
Required Qualifications
8+ years of experience in enterprise integration architecture.
5+ years supporting supply chain, warehouse management, fulfillment, or logistics systems.
Experience leading integrations for Tier-1 WMS platforms such as:
Manhattan
Blue Yonder
Körber
SAP EWM
Oracle WMS
Strong understanding of:
API architecture
Event-driven architecture
Microservices
Message queues
Enterprise integration patterns
Experience integrating:
ERP systems
OMS platforms
Warehouse automation systems
Transportation management platforms
Carrier systems
Technical Skills
Strong knowledge of:
REST APIs
JSON
XML
OAuth
JWT
API gateways
Event streaming architectures
Middleware platforms
Integration monitoring tools
Experience with one or more:
Azure/GCP/AWS Integration Services
MuleSoft
Boomi
Kafka
Preferred Qualifications
Prior Tier-1 WMS implementation experience.
Experience integrating robotics and warehouse automation platforms.
Knowledge of apparel, retail, or omnichannel fulfillment operations.
Experience integrating and supporting cloud-native SaaS applications.
Cloud certifications.
Key Competencies
Systems thinking
Technical leadership
Influencing without authority
Vendor management
Problem solving
Executive communication
Cross-functional collaboration
Warehouse and fulfillment domain expertise
Architectural governance
Operational excellence
The base salary range for this position is $220,000-$240,000 per year which represents the current range for the base salary for this exempt position. Please note that actual salaries will vary based on factors including but not limited to location, experience, and performance. As such, on occasion and when applicable, there is the possibility that the final, agreed-upon base salary may be outside of the upper end of the range. Please also note the range listed is just one component of the company’s total rewards package for exempt employees. Other rewards may include performance bonuses, long term incentives, a PTO policy, and many other progressive benefits.
For CA residents, Job Applicant Privacy Policy HERE .
Werkstudent Softwareentwicklung (w/m/d) - Startup für sichere KI
Company Name
**
Bochum
Basic
Posted 25 days ago