Admir Krnic

Data Engineer owning pharma supply chain pipelines

Podgorica, Montenegro

#OpenToWork

About

I started out building things end to end. My first job, at Logate in Podgorica, was full-stack: content management systems for internet service providers, a mail administration tool, frontend through backend through database schema. It taught me how a whole system fits together, which has mattered more than any particular framework I picked up there. Then I moved to Munich for a master's in Informatics at TUM, focusing on databases and artificial intelligence. Partway through I joined Qyobo, a pharmaceutical market-analytics company, and found the problem I have been working on ever since. Qyobo's product is built on regulatory and supply-chain data collected from hundreds of government sources across dozens of countries. That data is unstructured, contradictory, inconsistently formatted, and often simply wrong. I started on collection, writing scrapers, and moved gradually down the pipeline until I owned the regulatory and supply-chain domains end to end: petabytes per sync, billions of rows, and no canonical schema or identifier anywhere to anchor it. What I found is that I like this. The interesting problems there are not the ones with documented answers. They are the ones where I take a question to someone more senior and they do not know either, and the only way forward is analysis and judgment. Rebuilding the drug parsing module from scratch. Resolving hundreds of thousands of supply-chain entities across three sources with nothing to join on. Deciding when the honest answer is a visible gap rather than a confident guess. In 2025 I took on a second role, part-time, researching neural audio watermarking at DeepMark. I am first author on two papers, including what I believe is the first systematic study of backward compatibility in the field, a problem the EU AI Act has made suddenly urgent for anyone deploying generative audio. It is a different discipline from data engineering, but it comes from the same instinct: find the question nobody has answered yet, and answer it properly. That combination is what I would point to. Plenty of engineers build reliable pipelines. Plenty of researchers publish. Not many are doing both at the same time, and the overlap shows up as an unusual tolerance for problems that have no ground truth and no documentation, which in my experience describes most of the problems worth solving.

What I'm looking for

I am looking for a fully remote role, working from Montenegro on CET or hours that overlap it, as a data engineer or a machine learning engineer. Full-time. I can be hired through an employer of record or invoice as a registered contractor, so the arrangement is straightforward from either side. What I want from the work itself is ownership of a problem rather than a queue of tickets. The three years I have spent owning data domains end to end, from acquisition through to what the product serve

Experience

DeepMark

AI Research Engineer

DeepMark

Jun 2025 – Present

Neural audio watermarking research. Part-time, held concurrently with the role below. • Formalized backward compatibility in zero-bit neural audio watermarking — to our knowledge the first systematic study of the problem, and a requirement under EU AI Act Article 50 and its Code of Practice. Introduced an Angular Diversity Loss that restores true-negative rates to 99–100% against out-of-chain embedders with no measurable cost to detection accuracy or audio quality, and showed the same mechanism enables revocation of leaked embedders. Evaluated across 8 model versions spanning DenseNet, ViTMAE, ResNet and AST architectures. First author; under review at IEEE WIFS 2026. • Improved watermark robustness using contrastive learning, adding an InfoNCE objective and projection head to Timbre Watermarking training. Achieved 95–99% bitwise recovery accuracy across 25 distinct distortion and attack types while preserving imperceptibility. First author; presented at a Montenegro AI Association (MAIA) workshop. • Currently extending the backward-compatibility framework to multi-bit watermarking. • Stack: PyTorch, AWS.

Pytorchaws
Qyobo

Data Scientist / Developer

Qyobo

May 2023 – Present

Pharmaceutical market analytics delivered as a platform-as-a-service; ~27 people. Previously Data Scientist. • Own the regulatory and supply-chain data domains end to end — from acquisition through SQL and Python transformation to the outputs served to the product frontend — across a pipeline processing petabytes per sync and billions of rows. • Rebuilt the drug-data parsing module from scratch, replacing a brittle approach that failed unpredictably on real-world inputs. The current system reaches 70–100% extraction success across covered sources and removed a recurring class of silent parsing failures. • Expanded regulatory coverage by 10+ countries and raised data quality across many existing ones, working from unstructured, contradictory web sources with no shared schema or canonical identifiers. • Redesigned the finished-dosage-form marketing authorization dataset from a flat to a nested structure and built the processing pipeline around it, enabling correct mapping of manufacturers to packs and substances — approaching 100% mapping accuracy on priority markets. • Unified US supply-chain regulatory data scattered across 3 sources with no join keys, resolving hundreds of thousands of entities through iterative analysis and heuristics. Cited by the company owner as materially increasing confidence in the US dataset. • Combine SQL transformation pipelines with Python and LLM-based processing into single reconciled outputs, using Postgres, Trino, dbt, SQLMesh and Temporal. • Interviewed and helped select a working student who was hired on my recommendation and now contributes to a project I initiated and continue to own.

SQLPythonTemporalsqlmeshdbtPostgreSQL
Logate

Fullstack Developer

Logate

Sep 2021 – Sep 2022

• Built a content management system for internet service providers that controlled which banners were served based on the network gateway a user connected through — frontend, backend services and database schema. • Developed frontend and backend services for a mail administration application, letting clients manage mail domains, users and permissions. • Stack: Java (Spring), PHP, Angular, SQL, Cassandra, REST APIs.

AngularJavaPHPSQLCassandraSpring

Education

Technical University of Munich (TUM)

MSc · Informatics

2022 – 2025
Technical University of Munich

Technical University of Munich

Master of Science · Informatics

2022 – 2025
University of Montenegro

University of Montenegro

BSc · Computer Science

2018 – 2021
University of Montenegro

University of Montenegro

Bachelor of Science · Computer Science

2018 – 2021

Skills

Analytics EngineeringLarge Language Models (LLDeep LearningMachine LearningPyTorchWeb ScrapingTemporalSQLMeshTrinoData PipelinesData QualityEntity ResolutionAWSdbtPostgreSQLData Modelingetl/eltsqlPythonData Engineering

Languages

English (Professional working proficiency)Serbian (Native or bilingual proficiency)Croatian (Native or bilingual proficiency)Bosnian (Native or bilingual proficiency)Montenegrin/Serbian/Bosnian/Croatian (Native or bilingual proficiency)