AI Engineer

Haystack
Eastborough, Kansas, United StatesFull-timePosted Aug 27, 2026

About the role

We are working with a pioneering technology firm at the forefront of artificial intelligence and machine learning, revolutionising how businesses process and utilise vast amounts of data. This organisation is known for its innovative solutions and commitment to cutting-edge development, making it an exciting place to contribute your expertise.

The Role

Develop and integrate AI solutions, focusing on GPU-accelerated inference services

Build robust data pipelines for processing large volumes of documents and files

Optimise GPU configurations for efficient inference, considering memory, batch size, and concurrency

Utilise containerisation (Docker) and orchestration (Kubernetes) for scalable deployments, including GPU scheduling

Implement parallel and concurrent processing to ensure high-throughput pipelines

Contribute to a dynamic team pushing the boundaries of AI technology

What You'll Need

Bachelor's or Master's degree in Computer Science, Engineering, or a related field

Strong proficiency in Python for development and integration

Hands-on experience deploying GPU-accelerated inference services, ideally with NVIDIA NIM

Proven ability to build data pipelines and process large datasets (e.g., OCR, document AI)

Solid understanding of GPU configurations (memory, batch size, concurrency, MIG/MPS)

Experience with Docker and Kubernetes for containerisation and orchestration, particularly with GPU scheduling

What's On Offer

Opportunity to work on cutting-edge AI and GPU acceleration technologies

Engage in challenging and impactful projects

Collaborative and innovative work environment

Chance to significantly contribute to a leading tech firm's success

Apply via Haystack today!

Responsibilities

  • Develop and integrate AI solutions focusing on GPU-accelerated inference services
  • Build robust data pipelines for processing large volumes of documents and files
  • Optimise GPU configurations for efficient inference
  • Utilise containerisation and orchestration for scalable deployments
  • Implement parallel and concurrent processing for high-throughput pipelines
  • Contribute to a dynamic team pushing the boundaries of AI technology

Qualifications

  • Bachelor's or Master's degree in Computer Science, Engineering, or a related field
  • Strong proficiency in Python for development and integration
  • Hands-on experience deploying GPU-accelerated inference services
  • Proven ability to build data pipelines and process large datasets
  • Solid understanding of GPU configurations
  • Experience with Docker and Kubernetes for containerisation and orchestration

Benefits

  • Opportunity to work on cutting-edge AI and GPU acceleration technologies
  • Engage in challenging and impactful projects
  • Collaborative and innovative work environment
  • Chance to significantly contribute to a leading tech firm's success

Skills mentioned

PythonData EngineeringData PipelinesMachine LearningDeep LearningCUDAModel DeploymentModel ServingDockerKubernetes

About Haystack

Haystack combines AI & expert vetting to deliver world-class tech candidates who are engaged, aligned, and ready to interview. We're trusted by over 400,000+ tech candidates, working in Software Engineering, Data, Design, DevOps, Cloud, Tech Management, Testing, Product & Delivery, Architecture and more. 100s of employers from startups and scale-ups like Atom Bank, DuckDuckGo and Goodlord to established enterprises like American Express, Dunelm and AWS use Haystack to connect with qualified tech talent that they can't find anywhere else.

Technology51-200 employeesNewcastle upon Tyne, England