Skip to content
← Back to job listings

AI Systems Engineer - Agents & Inference

axelera · Netherlands (hybrid)

RemoteImported listingfull-time10 days ago

About The Role

About UsAxelera AI is not your regular deep-tech company. We are creating the next-generation AI platform to support anyone who wants to help advancing humanity and improve the world around us. In just five years, we have raised a total of $370 million and have built a world-class team of 250+ employees (including 60+ PhDs with more than 40,000 citations), both remotely from 20 different countries and with offices in Belgium, France, Switzerland, Italy, the UK, headquartered at the High Tech Campus in Eindhoven, Netherlands.We have also launched our Metis™ AI Platform, which achieves a 3-5x increase in efficiency and performance, and have visibility into a strong business pipeline exceeding $100 million.Our unwavering commitment to innovation has firmly established us as a global industry pioneer.Are you up for the challenge?Position OverviewWe're looking for an AI System Engineer to help build and maintain Axelera Wingman and Axelera's agentic AI platform. You'll develop the agent capabilities, execution environments and software integrations that make the platform reliable and useful. Deploying models and running AI workloads will help you validate the platform, identify gaps and improve the product.Key responsibilities:Build reliable agent systems, intelligent tools and application integrationsDevelop secure, scalable execution environments for demanding AI workloadsBuild and maintain Wingman's platform capabilities, including access to accelerated computing and remote workload executionDeploy and optimize computer vision models, LLM services and inference infrastructure to validate Axelera's agentic AI platformOwn capabilities from implementation and hardware validation through production operationTake ownership of implementation, verification and ongoing reliabilityTest functionality, security boundaries, hardware behavior and failure recoveryWrite clear documentation, reproducible setup instructions and practical operational runbooksBecome productive quickly with unfamiliar tools, SDKs and systemsCollaborate effectively and exercise sound technical judgmentQualificationsRequired:Agent Systems & Effective Agent UsePractical experience building with AI agents, model APIs and toolsUnderstanding of how agents interact with applications, manage state and recover from failuresAbility to use coding agents effectively to scope, implement and verify work while retaining ownership of correctness and technical decisionsExperience with tool calling, structured outputs, streaming or MCPCompute Platforms & Secure Workload ExecutionUnderstanding of how AI workloads run across applications, hosts and accelerated hardware, including remotely hosted environmentsComfort working with Linux, processes, resource allocation and diagnosing failures across the software and compute stackApplication of sound security principles to permissions, credentials and workload isolationAbility to build reliable execution with clear progress, cancellation and recoveryModel Deployment & Platform ValidationHands-on experience deploying computer vision models and working with LLM servicesAbility to use representative applications and inference workloads to verify platform correctness on real hardwareSkills to assess accuracy, latency, throughput and resource useExperience investigating failures across model code, inference SDKs and execution environmentsWe value demonstrated ability and judgment over particular degrees or certificationsNice to have:Software Engineering: Python, Rust or another systems language. APIs, databases and native integrations. TypeScript/React and Tauri experienceWorkload Orchestration: Job submission, scheduling, quotas, concurrency. Result retrieval and reproducible environments. SDK and driver compatibility managementSecurity: Authentication, authorization, least privilege. Access revocation, sandboxing, network security. Protection of users' files and dataCloud Operations (Google Cloud/AWS): Containers, infrastructure as code, CI/CD. Monitoring, incident investigation. Safe deployments, rollback and recoveryApplied ML & Accelerated Computing: PyTorch/ONNX. Image/video processing, detection, classification, segmentation. Calibration, quantization and hardware-aware optimizationEvaluation & Retrieval: Reproducible benchmarks for model quality, agent behavior and task success. Grounded retrieval and traceable resultsAdditional Strengths:Fine-tuning, multimodal applications, distributed inferenceDesktop packagingExperience maintaining developer toolsLocation We offer a flexible working arrangement, with options to: Work from one of our Axelera AI offices (Florence and Milan in Italy, Amsterdam and Eindhoven in the Netherlands, Leuven in Belgium, Paris in France, Zurich in Switzerland, or Bristol in the United Kingdom) if you're already based in the vicinity. Work fully remotely from any European country (incl. the UK) you are already in. What we offer This is your chance to shape and be part of a dynamic, fast-growing, international organization. We offer an attractive compensation package, including a pension plan, extensive employee insurances and the option to get company shares. An open culture that supports creativity and continual innovation is awaiting you. Collaborative ownership and freedom with responsibility is characteristic for the way we act and work as a team. At Axelera AI, we wholeheartedly embrace equal opportunity and hold diversity in the highest regard. Our steadfast commitment is to cultivate a warm and inclusive environment that empowers and celebrates every member of our team. We welcome applicants from all backgrounds to join us in shaping the future of AI.

This is an external listing. JobSpring does not represent or verify the employer. Report this listing