AI Engineer
About the role
We are working with a pioneering technology firm at the forefront of artificial intelligence and machine learning, revolutionising how businesses process and utilise vast amounts of data. This organisation is known for its innovative solutions and commitment to cutting-edge development, making it an exciting place to contribute your expertise.
The Role
Develop and integrate AI solutions, focusing on GPU-accelerated inference services
Build robust data pipelines for processing large volumes of documents and files
Optimise GPU configurations for efficient inference, considering memory, batch size, and concurrency
Utilise containerisation (Docker) and orchestration (Kubernetes) for scalable deployments, including GPU scheduling
Implement parallel and concurrent processing to ensure high-throughput pipelines
Contribute to a dynamic team pushing the boundaries of AI technology
What You'll Need
Bachelor's or Master's degree in Computer Science, Engineering, or a related field
Strong proficiency in Python for development and integration
Hands-on experience deploying GPU-accelerated inference services, ideally with NVIDIA NIM
Proven ability to build data pipelines and process large datasets (e.g., OCR, document AI)
Solid understanding of GPU configurations (memory, batch size, concurrency, MIG/MPS)
Experience with Docker and Kubernetes for containerisation and orchestration, particularly with GPU scheduling
What's On Offer
Opportunity to work on cutting-edge AI and GPU acceleration technologies
Engage in challenging and impactful projects
Collaborative and innovative work environment
Chance to significantly contribute to a leading tech firm's success
Apply via Haystack today!
Responsibilities
- Develop and integrate AI solutions focusing on GPU-accelerated inference services
- Build robust data pipelines for processing large volumes of documents and files
- Optimise GPU configurations for efficient inference
- Utilise containerisation and orchestration for scalable deployments
- Implement parallel and concurrent processing for high-throughput pipelines
- Contribute to a dynamic team pushing the boundaries of AI technology
Qualifications
- Bachelor's or Master's degree in Computer Science, Engineering, or a related field
- Strong proficiency in Python for development and integration
- Hands-on experience deploying GPU-accelerated inference services
- Proven ability to build data pipelines and process large datasets
- Solid understanding of GPU configurations
- Experience with Docker and Kubernetes for containerisation and orchestration
Benefits
- Opportunity to work on cutting-edge AI and GPU acceleration technologies
- Engage in challenging and impactful projects
- Collaborative and innovative work environment
- Chance to significantly contribute to a leading tech firm's success
Skills mentioned
About Haystack
Haystack combines AI & expert vetting to deliver world-class tech candidates who are engaged, aligned, and ready to interview. We're trusted by over 400,000+ tech candidates, working in Software Engineering, Data, Design, DevOps, Cloud, Tech Management, Testing, Product & Delivery, Architecture and more. 100s of employers from startups and scale-ups like Atom Bank, DuckDuckGo and Goodlord to established enterprises like American Express, Dunelm and AWS use Haystack to connect with qualified tech talent that they can't find anywhere else.