Open jobPosted 2 days ago · Closes in a month

SuperQ Quantum - AI Engineer (GPU & LLM Optimisation) – Canada / UAE / Remote

AI Engineer
Open to offers BC, Canada
Twine JobsVerified
Stretford, United Kingdom · Organization
Recently active

AI Engineer is needed in BC, Canada.

Client: SuperQ Quantum

Location: BC, ca

Contract: Freelance

Job Description

This is a remote position with a preference for candidates located in Canada, UAE, or fully remote. SuperQ Quantum Computing Inc. is looking for an AI Engineer specializing in GPU and LLM optimization to maximize the performance, scale, and efficiency of the models that power their multi-agent architecture. The engineer will play a crucial role in accelerating inference, reducing latency, and optimizing hardware utilization across various industries, including healthcare, manufacturing, and finance.

Responsibilities

  • Accelerate Inference: Optimize LLM inference and serving pipelines using high-performance frameworks (e.g., vLLM, TensorRT-LLM, TGI, Triton Inference Server).
  • Model Compression: Implement state-of-the-art model compression techniques, including quantization (e.g., GPTQ, AWQ, FP8/INT8), pruning, and knowledge distillation to reduce memory footprints.
  • Hardware Profiling: Profile GPU performance to identify and eliminate bottlenecks, optimizing memory management techniques such as KV caching and PagedAttention.
  • Custom Compute: Develop and optimize custom CUDA or OpenAI Triton kernels for specialized, computationally heavy operations where standard libraries fall short.
  • Cross-functional Deployment: Collaborate with backend and platform engineering teams to deploy highly scalable, low-latency AI endpoints within a Kubernetes-based infrastructure.
  • System Monitoring: Ensure robust monitoring of GPU health, utilization metrics, latency, and throughput in production environments.

Requirements

  • Experience: 2–4 years of experience in ML engineering, AI infrastructure, or high-performance computing (HPC) with a strong focus on Large Language Models.
  • Core Languages: Strong programming skills in Python; proficiency in C++ and/or CUDA is highly preferred.
  • Frameworks: Hands-on experience with LLM serving architectures, distributed computing, and optimization libraries (e.g., DeepSpeed, Hugging Face Accelerate, Ray).
  • Hardware Knowledge: Deep understanding of GPU architectures (NVIDIA), memory hierarchies, and parallel computing paradigms (e.g., NCCL).
  • Infrastructure: Solid understanding of containerization and orchestration (Docker, Kubernetes) tailored for GPU-accelerated workloads.
  • Bonus: Experience with hardware-aware neural architecture search (NAS), ML Ops, or working at the intersection of classical AI hardware and quantum compute interfaces.

Benefits

  • Competitive salary plus bonus based on module delivery and platform impact.
  • Stock options/equity to participate in the growth of a foundational tech platform.
  • Global remote-friendly roles with occasional team meetups and hack-weeks.
  • Learning & development stipend for conferences, AI/LLM workshops, certifications (e.g., in ML Ops, prompt engineering).
  • “Innovation time” built into schedule for passion projects and internal hackathons.
  • Cross-domain exposure to work across industries and build modules for different verticals.

Work Culture

  • Startup-scale agility within a mission-driven organization encouraging significant impact and direction shaping.
  • Collaborative environment working across disciplines with quantum engineers, UI developers, and domain experts.
  • Focus on autonomy, responsibility, and ownership over modules from conception through production and maintenance.
  • Culture of continuous learning encouraging curiosity in new models, LLM architectures, and AI trends.
  • Balanced remote-first mindset featuring flexible working hours and moments of team building.

Equal Opportunity Statement

SuperQ is dedicated to building a workplace where everyone can thrive. They are an equal opportunity employer and make employment decisions based on merit, qualifications, and business needs without regard to race, color, religion, sex, gender identity or expression, sexual orientation, national origin, age, disability, veteran status, or any other characteristic protected by law. SuperQ values diversity, equity, and inclusion and strives to create a culture of belonging for all employees.


  • Apply


    Enter your email to apply

     

    By applying, you agree to our Terms.

    Already have an account? Sign in.

  • How It Works


    🔍

    Get quality leads

    Review job leads for free, filter by local or global clients, and get real time notifications for new opportunities.


    🎉

    Apply with ease

    Pick the best leads, unlock contact details, and apply effortlessly with Twine's AI application tools.


    📈

    Grow your career

    Showcase your work, pitch to the best leads, land new clients and use Twine’s tools to find more opportunities.