Relocantly← All jobs

🇳🇱 Amsterdam, Netherlands · 1h ago

Multimodal Data & AI Infrastructure Expert

Sparagus

LinkedInseniorEnglish-friendly
Apply on LinkedIn →Get jobs like this daily
Location: Amsterdam, the NetherlandsEmployment Type: Full-time, PermanentAbout the RoleWe are looking for a senior Multimodal Data & AI Infrastructure Expert to help shape the next generation of data infrastructure for the AI Agent era.You will drive the architecture and core technology development of a next-generation multimodal intelligent data platform, working across heterogeneous computing, multimodal computing engines, vector storage and retrieval, distributed caching, AI4DB, and Data Agent technologies.The role focuses on building end-to-end capabilities for massive-scale heterogeneous and unstructured data — spanning resource scheduling, computation, retrieval, storage, and intelligent autonomous management.You will work at the intersection of Data Infrastructure and AI, developing core technologies optimized for large language models, multimodal workloads, and AI Agent ecosystems.What You Will DoMultimodal Data InfrastructureDesign and develop a unified scheduling and management platform for heterogeneous computing resources, including CPUs, GPUs, and NPUsDevelop serverless resource pooling and management capabilities across multiple computing enginesOptimize heterogeneous resource scheduling, system performance, and overall resource utilizationMultimodal Computing EnginesDevelop and optimize computing engines for large-scale unstructured and multimodal dataEnable efficient processing and analysis of text, images, video, and other data typesDesign next-generation query optimizers and hybrid execution enginesSupport high-performance retrieval and processing of vector, textual, geospatial, and other dataVector Storage & RetrievalDesign and develop systems for multimodal vector retrieval, storage, metadata management, and access controlDevelop intelligent storage optimization and index lifecycle management capabilitiesIntegrate data infrastructure with relevant open-source ecosystemsDistributed CachingBuild high-performance distributed caching services for multimodal data platformsDevelop high-speed near-compute caching capabilities for computing enginesOptimize efficient east-west data transfer across distributed computing enginesAI4DB & Data AgentsExplore and apply AI4DB (AI for Databases) and LLM Agent technologiesDevelop intelligent Data Agents and Skills for next-generation data platformsEnable greater automation and intelligence across workload development, data storage, data analysis, operations, and system maintenanceRequired QualificationsStrong programming skills in languages such as C, C++, Python, or JavaStrong R&D background and technical expertise in one or more of the following areas:Database SystemsBig Data SystemsDistributed SystemsHigh-Performance ComputingStrong understanding of heterogeneous computing architectures and performance considerations involving CPUs, GPUs, and NPUsHands-on experience with heterogeneous resource management and schedulingExperience in cloud computing platform development and maintenanceExperience with DevOps-related engineering practicesPreferred QualificationsExperience in one or more of the following would be highly relevant:Query optimizersExecution enginesStorage enginesDistributed storage systemsVector databases, vector indexing, or high-performance retrieval systemsLarge-scale multimodal or unstructured data processingLLM fine-tuning and reinforcement learningNLP, computer vision, or time-series data processingIntegration of AI technologies with database or data systemsAI4DBLLM Agents / Data AgentsHigh-performance distributed cachingServerless and heterogeneous computing infrastructureWhat We OfferA highly competitive compensation packageRelocation allowance for candidates relocating for the positionAnnual performance bonusAttractive short-term and long-term incentive programsThe opportunity to work on next-generation data infrastructure at the intersection of AI, database systems, distributed computing, and multimodal technologiesWho Should Apply?We are particularly interested in experienced database, distributed systems, data infrastructure, and high-performance computing experts who have worked on core system technologies rather than only application-level data engineering.If you have built database engines, distributed data systems, multimodal computing infrastructure, vector retrieval systems, or AI-native data platforms and are interested in building the next generation of infrastructure for LLM and AI Agent ecosystems, we would be very happy to connect.

Sourced from LinkedIn. Relocantly aggregates public job postings; apply on the original site.