{"id":755512,"url":"https://alion.io/job/opennebula-ai-platform-engineer-oneai","title":"AI Platform Engineer (OneAI)","company":{"id":680269,"name":"OpenNebula","domain":"opennebula.io","url":"https://alion.io/company/opennebula","size_band":null,"is_staffing_agency":false,"employer_type":"direct","is_intermediary":false,"listed_via":null,"ats_vendor":"Teamtailor","truth_index":{"grade":"B","score":75,"open_postings":3,"ghost_share":0,"stale_share":1,"repost_share":0,"time_to_fill_p50_days":null,"computed_at":"2026-10-01T05:45:00Z"}},"role":"AI/ML","role_family":"AI/ML","seniority":"middle","employment_type":"full_time","work_mode":"on_site","remote_scope":null,"remote_scope_basis":null,"remote_working_hours":null,"hiring_geo_confidence":"structured","locations":["Pozuelo de Alarcón, Spain"],"countries":["ES"],"hiring_countries":[],"hiring_countries_total":0,"salary":null,"salary_estimate":{"min_usd":50000,"max_usd":144000,"period":"year","method":"global_role_cell_scaled_by_country","sample_n":585},"experience_years_min":3,"visa_sponsorship":false,"relocation_package":false,"has_equity":false,"technologies":[{"name":"Edge AI","optional":false},{"name":"Fine-tuning","optional":false},{"name":"Hugging Face","optional":false},{"name":"Kubernetes","optional":false},{"name":"Linux","optional":false},{"name":"Machine Learning","optional":false},{"name":"PyTorch","optional":false},{"name":"Transformers","optional":false},{"name":"vLLM","optional":false},{"name":"Windows","optional":false}],"status":"live","first_seen_at":"2026-04-16T11:10:37Z","employer_posted_date":"2026-04-16","last_verified_at":"2026-09-30T08:15:15Z","board_verified":true,"closed_at":null,"days_open":167,"trust":{"level":"stale","repost_count":0,"flags":["stale"],"days_open":167},"description":"For over a decade now, OpenNebula Systems has been building the open source technology that helps organizations around the world to manage their corporate data centers and build Enterprise Clouds with unique, innovative features.\nBorn back in the day as an open source platform for Private Clouds, OpenNebula is a powerful, but easy-to-use, open source Cloud & Edge Computing Platform whose community includes nowadays leading companies and public agencies in a wide range of industry niches and countries.\nIf you want to join an established leader in the cloud infrastructure industry and the global open source community, keep reading, because you can now join exceptionally passionate, talented colleagues, and help world´s leading enterprises implement their next-generation edge and cloud strategies. We are hiring!\nCome join us in our distributed, fully-remote, international team, where we promote Creativity, Innovation, Collaboration, Open Communication and Iterative Work, and we´re seeking a AI Engineer to work in the development of the OpenNebula open-source cloud management platform and its new strategic project in AI.\nSince 2019, and thanks to the support from the European Commission, OpenNebula Systems is leading the edge computing innovation in Europe, investing heavily in research and development, and playing a key role in the key strategic initiatives of the European Union.https://opennebula.io/innovation/\nJob Description\nWe are seeking an AI Platform Engineer to contribute to the design, implementation, and deployment of advanced AI capabilities within OneAI, our platform for operating high-performance inference and training workloads across diverse cloud and edge environments leveraging OpenNebula.\nIn this role, you will be responsible for integrating cutting-edge AI frameworks, tools and engines (such as vLLM, PyTorch and Unsloth) into a secure, scalable, and highly observable AI platform that can provision compute, schedule workloads, and expose reliable APIs for the entire AI model lifecycle, serving and fine-tuning process.\nBeyond the backend infrastructure, you will help shape the OneAI user experience by designing intuitive workflows for managing models, configuring deployments, and operating inference and training jobs. A key focus will be integrating with public model repositories like Hugging Face to streamline model and dataset discovery, import, versioning, and deployment directly into OpenNebula cloud deployments.\nWorking at the intersection of applied AI, systems engineering and product design, you will ensure that inference and training are efficient at scale. Key focus areas include establishing deployment strategies, optimizing performance and cost efficiency, implementing GPU-aware operations, and building comprehensive observability tools to track latency, throughput, utilization, and failure rates.\nAs we are an international team, please submit your CV in English.\nCore Responsibilities\nDesign, implement, and deploy advanced AI capabilities within the OneAI platform.\n\nShape the end-user experience by designing intuitive workflows for model management, deployment configuration, and job operation.\n\nStreamline the model lifecycle by integrating public repositories (e.g., Hugging Face) for seamless discovery, import, versioning, and deployment.\n\nBridge the gap between systems engineering and product design to ensure a seamless transition from backend infrastructure to user features.\n\nAI Platform & Systems Engineering\nIntegrate cutting-edge AI frameworks and engines, such as vLLM, NVIDIA Dynamo and Unsloth, into a secure and scalable environment.\n\nLeverage OpenNebula to orchestrate high-performance inference and training workloads across diverse cloud and edge environments.\n\nDevelop and maintain reliable APIs for compute provisioning and workload scheduling.\n\nImplement GPU-aware operations to ensure optimal resource allocation and hardware utilization.\n\nBuild comprehensive observability suites to monitor and track critical metrics, including latency, throughput, utilization, and failure rates.\n\nResearch, Optimization & Innovation\nEstablish and refine deployment and workflow strategies to ensure AI workloads remain efficient and stable at scale.\n\nOptimize system architecture to balance high performance with cost efficiency.\n\nResearch and integrate emerging AI tools and engines to keep the OneAI platform at the forefront of the industry.\n\nAnalyze performance bottlenecks to iterate on the efficiency of both training and inference processes.\n\nExperience Required\nAcademic Background and Certifications\nBachelor’s or Master’s degree in Computer Science, Information Technology, or Engineering.\n\nProfessional Experience\n3+ years of experience in applied AI, machine learning, or software engineering, with hands-on delivery of AI/ML solutions in production environments\n\nDemonstrated experience designing and deploying high-performance AI infrastructure, specifically focusing on the scalability and reliability of inference and training workloads.\n\nProven track record of deploying Large Language Models (LLMs) at scale, with deep knowledge of serving engines (e.g., vLLM) and fine-tuning tools (e.g., Unsloth).\n\nExperience building AI-centric platforms or toolchains that manage the model lifecycle (versioning, deployment, and discovery).\n\nExperience with GPU orchestration and optimizing workloads for cloud, distributed or large-scale environments and collaborating with platform or infrastructure teams.\n\nTechnical Experience\nHands-on experience with high-throughput inference engines (e.g., vLLM) and fine-tuning tools (e.g., Unsloth)\n\nProficiency in integrating with the Hugging Face ecosystem (Transformers, Hub, Datasets) for model and data management.\n\nExperience implementing monitoring tools to track system-level AI metrics such as token throughput, latency, GPU utilization, and failure rates.\n\nExperience designing and implementing scalable, reliable APIs for compute provisioning and workload scheduling.\n\nExperience working with cloud platforms and containerized environments (e.g., OpenNebula, Kubernetes)\n\nLanguage Skills\nAdvanced English level (B2 or higher) is required.\n\nSoft Skills & Collaboration\nStrong analytical and problem-solving skills with a practical, experimental mindset\n\nAbility to work independently in complex, fast-moving technical environments\n\nComfortable collaborating in distributed teams and engaging with open-source communities\n\nWhat's in it for me?\nSome of our benefits and perks vary depending on location and employment type, but we are proud to provide employees with the following;\nCompetitive compensation package and flexible remuneration: Meals, Transport, Nursery/Childcare\n\nCustomized workstation (macOS, Windows, Linux)\n\nPrivate health insurance\n\nPaid time off: Holidays, Personal Time, Sick Time, Parental leave\n\nAfternoon-off working day every friday and during summer\n\nRemote company with bright HQ centrally located in Madrid; offices in Boston (USA), Brussels (Belgium) and Brno (Czech Republic); and access to office space near your location when needed. During the first year, for onboarding purposes, and for participation on certain projects, employees should be able to attend events and face-to-face meetings in our Madrid offices and other European cities. All employees are also required to attend our company-wide face-to-face all-hands meetings twice a year\n\nHealthy work-life balance: We encourage the right for Digital Disconnecting and promote harmony between employees personal and professional lives\n\nFlexible hiring options: Full Time/Part Time, Employee (Spain/USA) / Contractor (other locations)\n\nWe are building an awesome, Engineering First Culture and your opinion matters: Thrive in the high-energy environment of a young company where openness, collaboration, risk-taking, and continuous growth are valued\n\nBe exposed to a broad technology ecosystem. We encourage learning and researching new technologies and methods as part of your everyday duties","description_format":"text","description_chars":8030,"description_truncated":false,"requirements":{"experience_years_min":3,"management_years_min":null,"team_size_min":null,"manages_managers":false,"education":{"level":"bachelor","optional":false},"security_clearance":false,"languages":[{"language":"English","level":"Upper-Intermediate (B2)","optional":false}]},"benefits":["Health insurance","Parental leave"],"hiring_locations":[],"hiring_excludes":[],"relocation_offered":false,"industries":["Cloud Platforms (IaaS & PaaS)","Virtualization & Containers"],"lifecycle":[{"event":"open","at":"2026-09-11T17:05:44Z"}],"liveness":{"score":7,"band":"cold","label":"Long shot","p_open":1,"p_active":0.249,"p_room":0.28,"age_days":167,"expected_fill_days":18,"reasons":["conf:21","win:tail","crowd:"],"computed_at":"2026-10-01T05:45:00Z"},"pay":null,"html_url":"https://alion.io/job/opennebula-ai-platform-engineer-oneai","json_url":"https://alion.io/job/opennebula-ai-platform-engineer-oneai.json","meta":{"generated_at":"2026-10-01T09:44:45Z","cache_seconds":300,"methodology":"https://alion.io/methodology","terms":"https://alion.io/terms","contact":"https://alion.io/contact","api":"https://alion.io/developers","usage":{"tier":"crawler","counted_by":"address","units_charged":1,"used_today":480,"day_limit":5000,"remaining_today":4520,"minute_limit":60,"resets_at":"2026-10-02T00:00:00Z"}}}