> Markdown version of [/jobs/ext/1394562-developer-i-o-acceleration](https://www.wearedevelopers.com/jobs/ext/1394562-developer-i-o-acceleration). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Developer - I/O Acceleration - **Company:** IBM - **Location:** San Jose, CA, United States - **Contract:** Permanent contract - **Skills:** Application Programming Interfaces (APIs), Agile Methodology, Profiling, Codecs, Nvidia CUDA, Linux, File Systems, Distributed Data Store, General Parallel File Systems, Remote Direct Memory Access, Subsystems, Systems Integration, Apache Spark, C++14, Nvme - **Published:** July 23, 2026 - **Apply:** https://www.juju.com/job/00000000gin3vz ## About the Role * Hands-on experience in at least one of: storage I/O subsystems, decompression and codec implementation, or query-engine data paths. * Working knowledge of GPU-aware pipelines or adjacent acceleration frameworks (CUDA, GPUDirect, or similar). * Strong performance-profiling and bottleneck-isolation skills - you can read a flame graph, an nsys trace, and an fio result and know what to do next. * Familiarity with distributed data systems and the realities of running them at scale. * Track record of delivering production software in Agile, collaborative environments, including contributing to automated CI/CD pipelines. Preferred technical and professional experience * Production experience with GPUDirect Storage, RDMA, or NVMe-oF integrations. * Exposure to ESS6000, Lustre, GPFS, or other parallel and clustered file systems. * Track record of large-scale benchmarking, and publishing or presenting performance results. * Contributions to open-source data, storage, or GPU-runtime projects (Arrow, cuDF, Velox, DuckDB, Spark, and similar). ## Description As an Engineer on I/O Acceleration, you will own and optimize the path data travels from disk to GPU memory. Your work will directly determine how many queries per dollar Gala Lakehouse can deliver on Blackwell - and whether modern analytics workloads run GPU-bound (where they should) or I/O-bound (where they are today). What You'll Do * Design, build, and optimize accelerated I/O and decompression paths for data-intensive analytics workloads. * Improve end-to-end throughput across the storage * network * host * GPU boundary, eliminating copies, syscalls, and stalls. * Integrate with GPU-aware runtimes and high-bandwidth fabrics (GPUDirect Storage, RDMA, NVMe-oF) and tune for Blackwell-class hardware. * Build benchmarks and microbenchmarks that expose I/O cliffs, queue contention, and tail latency under realistic query mixes. * Instrument the data path so cost-per-query, bandwidth-per-GPU, and CPU overhead are first-class, observable metrics. * Collaborate with the query engine, storage, and hardware teams to co-design APIs that make accelerated I/O usable, not just possible. Required technical and professional expertise * Strong modern C++ and deep comfort with Linux systems internals (page cache, O_DIRECT, io_uring, NUMA, scheduling). ## Related Videos - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [Accelerating Python on GPUs](https://www.wearedevelopers.com/videos/859-accelerating-python-on-gpus) - [Docker exec without Docker](https://www.wearedevelopers.com/videos/1094-docker-exec-without-docker) - [Coffee with Developers - Stephen Jones - NVIDIA](https://www.wearedevelopers.com/videos/1303-coffee-with-developers-stephen-jones-nvidia) - [Accelerating Python on GPUs](https://www.wearedevelopers.com/videos/1112-accelerating-python-on-gpus) - [Discover the open source trio you didn’t expect: .NET and PostgreSQL on Linux](https://www.wearedevelopers.com/videos/2042-discover-the-open-source-trio-you-didn-t-expect-net-and-postgresql-on-linux) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline) - [Dev Digest 157: CUDA in Python, Gemini Code Assist and Back-dooring LLMs](https://www.wearedevelopers.com/magazine/557-dev-digest-157-cuda-in-python-gemini-code-assist-and-back-dooring-llms) - [Dev Digest 102 - Race conditions](https://www.wearedevelopers.com/magazine/386-dev-digest-102-race-conditions) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers)