KIEFER Logo

KIEFER

Senior ML Engineer (LLM)

Posted 24 Days Ago
Be an Early Applicant
Remote or Hybrid
Hiring Remotely in Athens
Senior level
Remote or Hybrid
Hiring Remotely in Athens
Senior level
Develop and continuously improve a Greek-focused large language model across pre-training, training from scratch, fine-tuning, evaluation, and production deployment. Build ML pipelines for inference, serving, monitoring, and lifecycle management, while optimizing latency, throughput, cost, quantization, and GPU utilization. Work with datasets, benchmarks, and experiments to improve model quality and domain performance.
The summary above was generated by AI

About the company:

KIEFER is building Greece's integrated AI ecosystem. From renewable energy infrastructure and AI systems to robotics and enterprise applications, we connect the technologies that power Europe's intelligent future. Founded in 2014, KIEFER has delivered 600MW+ of energy projects and is now developing sovereign AI infrastructure, enterprise AI products and physical AI systems for Greece and Southeast Europe.

About the role:

We are looking for a Senior Machine Learning Engineer to join Kiefer Tech and strengthen our ML Engineering team.

In this role, you will work on the development and continuous improvement of Sophea AI, our Greek-focused Large Language Model. You will be deeply involved in LLM pre-training, training from scratch, fine-tuning, model evaluation, inference optimization, and production-grade ML systems.

This is a hands-on engineering role for someone who has already worked directly with language models and understands how to improve their quality, performance, and reliability in real production environments.

Important: this role requires strong practical experience with LLM development. Classical ML, computer vision, basic RAG, or high-level AI tools alone will not be enough for this position.

What you will do:

  • Work on Sophea AI across LLM pre-training, training from scratch, fine-tuning, evaluation, and continuous model improvement

  • Build production-grade ML pipelines for inference, serving, deployment, monitoring, and model lifecycle management

  • Optimize model performance in production, including latency, throughput, cost efficiency, quantization, and GPU workload usage

  • Work with datasets, experiments, benchmarks, and evaluation methods to improve language model quality and domain-specific performance

What you will need:

  • Strong hands-on experience with LLMs, including pre-training, training from scratch, fine-tuning, evaluation, and performance improvement

  • Strong ML engineering background, including Python, PyTorch, Docker, and production ML practices

  • Experience with model serving, inference optimization, quantization, GPU workloads, and frameworks such as vLLM, SGLang, NVIDIA Triton, TensorRT, TGI, or similar tools

  • Ability to build production-grade ML systems, not only research prototypes, scripts, basic RAG applications, or high-level AI integrations

Nice to have:

  • Experience with ASR systems, speech models, or speech-to-text pipelines

  • Experience working with non-English language models, multilingual models, or low-resource language adaptation

  • Experience with MLOps infrastructure, experiment tracking, model serving pipelines, and GPU workload management

  • Contributions to open-source ML projects or published research in AI/ML

What is there for you:

  • Compensation: competitive package aligned with talent benchmarks

  • Impact: hands-on role working on Sophea AI, one of the most ambitious Greek-focused AI products in the market

  • Work format: remote work option, with relocation support available for candidates open to working from our Athens office

  • AI-native environment: real challenges across LLMs, training, fine-tuning, inference optimization, GPU workloads, and production AI systems

  • NVIDIA ecosystem: access to related conferences, certifications, internal knowledge sharing, and advanced AI infrastructure through Kiefer’s strategic collaboration

  • Culture: engineering-first, high autonomy, low bureaucracy, and space to build meaningful AI products

Similar Jobs

8 Days Ago
Remote
Senior level
Senior level
Artificial Intelligence • Information Technology • Consulting
Own production optimization of LLM and VLM inference systems, improving latency, throughput, memory efficiency, GPU utilization, quality, reliability, and cost. Deploy and benchmark inference engines, build compression and quantization workflows, implement acceleration techniques, diagnose serving bottlenecks, and collaborate with kernel, platform, research, product, and customer teams.
Top Skills: CudaFlashinferKserveKubernetesLmcacheNvidia DynamoPythonPyTorchRayRay ServeSglangTensorrt-LlmTritonTriton Inference ServerVllm
56 Minutes Ago
Remote or Hybrid
Junior
Junior
Consumer Web • Digital Media • Edtech • Information Technology • Social Impact • Software
Oversee scientific manuscript peer review from reviewer selection through editorial decisions, revisions, production, and publication. Coordinate with authors, editors, publishing managers, and production providers to ensure timely, accurate, ethical, and high-quality publication of manuscripts. Manage high volumes independently, meet strict deadlines, and ensure compliance with journal and industry standards.
Top Skills: HTMLLatexExcelMS OfficeMicrosoft PowerpointMicrosoft WordPdfXML
Yesterday
Remote or Hybrid
Senior level
Senior level
Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
Leads incentive compensation strategy, quota-setting, recognition programs, analytics, governance, implementation, and continuous improvement across Pfizer’s international markets. Partners with business leaders and cross-functional teams to design compliant, motivating, and financially responsible programs, deliver data-driven recommendations, oversee payouts and operational timelines, and maintain audit readiness. The role requires strong quantitative analysis, commercial operations expertise, project management, stakeholder influence, and experience supporting pharmaceutical or healthcare organizations.

What you need to know about the Edinburgh Tech Scene

From traditional pubs and centuries-old universities to sleek shopping malls and glass-paneled office buildings, Edinburgh's architecture reflects its unique blend of history and modernity. But the fusion of past and future isn't just visible in its buildings; it's also shaping the city's economy. Named the United Kingdom's leading technology ecosystem outside of London, Edinburgh plays host to major global companies like Apple and Adobe, as well as a growing number of innovative startups in fields like cybersecurity, finance and healthcare.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account