Description
A client of byteSpark.ai is seeking a Principal Applied AI Engineer to provide senior technical direction and hands-on expertise across speech, NLP, large language models and modern applied AI capabilities within secure, high-volume production environments. The role will focus initially on improving speech-to-text quality and reliability, including the evaluation, adaptation and fine-tuning of ASR models such as Whisper or equivalent technologies, while also supporting NLP-driven transcript analysis and the development of AI assistant capabilities using current agentic AI and MCP-based integration approaches. The successful candidate will work closely with existing architects, developers and platform teams to guide model selection, solution architecture, evaluation frameworks, inference optimisation and infrastructure dimensioning, with particular emphasis on on-premises, offline and air-gapped deployment. This is a hands-on principal-level individual contributor role for someone who can combine deep technical judgement with practical model work, mentor engineering teams, and help turn complex AI requirements into reliable production capabilities.
Requirements
1. Hands-on production experience in applied AI or machine learning, with strong depth in speech, NLP or related model-driven systems.
2. Strong Python capability with practical experience using PyTorch and the Hugging Face ecosystem or comparable modern deep-learning frameworks.
3. Production experience with automatic speech recognition and speech-processing pipelines using Whisper, WhisperX, Wav2Vec2 or equivalent technologies.
4. Demonstrated ability to adapt, fine-tune, evaluate and improve model performance using measurable quality metrics such as word error rate, latency, throughput or resource efficiency.
5. Strong understanding of transformer architectures, LLMs, embeddings, tokenisation, sequence modelling and NLP techniques relevant to classification, entity extraction and conversational AI.
6. Practical exposure to modern Agentic AI approaches, tool integration, API or database actions, with current hands-on understanding of MCP or closely related integration patterns.
7. Experience deploying or operating open-source AI models in on-premises, offline, self-hosted or air-gapped environments.
8. Ability to evaluate model size, hardware requirements and inference trade-offs, with practical exposure to optimisation techniques such as quantisation, ONNX, TensorRT, batching, pruning or distillation.
9. Strong technical judgement with the ability to guide architects and engineering teams on model selection, solution design, evaluation strategy and infrastructure requirements.
10. Strong English communication, technical documentation, stakeholder collaboration, design review and mentoring capability.
Desirable
1. Hands-on experience with WhisperX or equivalent advanced ASR frameworks in production environments.
2. Experience with multilingual speech recognition, particularly Arabic, regional dialects or other challenging language and accent variations.
3. Experience with speaker diarisation, noisy or degraded audio, telephony voice streams, far-field audio or other complex speech-processing conditions.
4. Practical experience with real-time or near-real-time speech inference and optimisation for latency, throughput and resource efficiency.
5. Experience with NLP techniques for transcript classification, named-entity recognition, sentiment or threat-related analysis using BERT-family models, spaCy or comparable technologies.
6. Hands-on experience with quantisation, ONNX, TensorRT or other techniques for optimising models for constrained or specialised hardware.
7. Experience delivering AI systems in telecommunications, government, cybersecurity, defence, public-sector or other regulated and security-sensitive environments.
8. Experience contributing to AI research, open-source machine-learning projects, technical publications, patents or other evidence of advanced technical depth.
9. Exposure to non-generative image or video analytics, multimodal AI or related media-analysis capabilities relevant to future product development.
Role Highlights
💰 Compensation
USD 80,000–90,000 gross annual base + performance bonus + local benefits
📍 Location
Europe (Remote)
💼 Work Location Type
Remote
🏢 Department
Information Technology
🏭 Industry
Technology, Information & Media
🔹 Sub-Industry
AI Development