Applied ML Engineer
Build production ML systems from research methods to user-facing products
As an Applied ML Engineer at Sentient, you will reproduce and evaluate research methods using open-weight and API-accessible models, design evaluation datasets and scoring methods, and turn research workflows into product experiences. You will build and extend evaluation infrastructure, ship production-quality systems with APIs and observability, and work across model internals, inference infrastructure, and backend systems to turn promising m...
Why This Role?
Turn promising ML research into reliable, measurable, and usable products with real user impact
Key Responsibilities
- Reproduce and evaluate research methods using open-weight and API-accessible models
- Design evaluation datasets, probes, scoring methods, baselines, and experiment harnesses
- Build and extend evaluation infrastructure including runners, judges, persistence, and reporting
- Turn research workflows into product experiences including experiment configuration, runs, traces, and reports
- Ship production-quality systems with APIs, background jobs, observability, testing, and documentation
Requirements
- Strong Python engineering skills
- Hands-on experience with PyTorch and Hugging Face Transformers
- Strong understanding of ML evaluation including dataset design, baselines, metrics, calibration, and reproducibility
- Ability to read ML research papers and implement methods from first principles
- Experience building production software beyond notebooks including APIs, asynchronous jobs, databases, logging, testing, and deployment
Required Skills
Indonesia Context
- Working Hours Overlap:
- Flexible — work your own hours
View Original Description from Ashby Job BoardsShow more
Original description from Ashby Job Boards
APPLIED ML ENGINEER THE ROLE We’re looking for an Applied ML Engineer to build systems at the intersection of machine learning research and production software. This is an end-to-end engineering role. You should be comfortable reading a research paper, identifying what is actually testable, building the smallest useful experiment, evaluating it rigorously, and turning the result into a production system that users can interact with. You’ll work across model evaluation, model internals, inference infrastructure, backend systems, and product interfaces. The goal is not simply to reproduce research. It is to turn promising methods into reliable, measurable, and usable products. WHAT YOU’LL DO - Reproduce and evaluate research methods using open-weight and API-accessible models. - Design evaluation datasets, probes, scoring methods, baselines, calibration tests, and experiment harnesses. - Work directly with model weights, logits, hidden states, activations, model APIs, and inference infrastructure when required. - Build and extend our evaluation infrastructure, including runners, judges, persistence, experiment orchestration, and reporting. - Turn research workflows into product experiences, including experiment configuration, runs, traces, comparisons, reports, and review workflows. - Investigate how verification methods behave under model modification, including fine-tuning, merging, quantization, distillation, safety removal, and deliberate evasion. - Design controlled experiments that separate meaningful signals from artifacts or confounders. - Write clear technical reports that distinguish measured evidence, interpretation, and hypotheses. - Ship production-quality systems with APIs, background jobs, observability, testing, and documentation. WHAT WE’RE LOOKING FOR - Strong Python engineering skills and hands-on experience with PyTorch and Hugging Face Transformers. - A strong understanding of ML evaluation, including dataset design, baselines, metrics, calibration, false positives, false negatives, statistical uncertainty, and reproducibility. - Ability to read ML research papers and implement methods from first principles rather than relying entirely on existing packages. - Experience building production software beyond notebooks, including APIs, asynchronous jobs, databases, logging, testing, and deployment. - Comfort working with open-weight models and understanding how modern LLM inference systems operate. - Ability to work across backend and frontend boundaries. Our product surface is primarily React/TypeScript, and you should be able to make complex experiments and results understandable to users. - Strong technical judgment about what experimental evidence does and does not support. For example, evidence that one model was derived from another is not necessarily evidence that it was directly trained on that model's outputs. - High agency and a strong sense of ownership. You are comfortable identifying problems, proposing solutions, and driving work forward without waiting for detailed instructions. - Comfortable working in a fast-moving startup environment where priorities can evolve quickly and individuals are expected to operate across functions. USEFUL EXPERIENCE Experience in any of the following is a plus: - Model provenance, fingerprinting, watermarking, distillation detection, red-teaming, safety evaluations, or interpretability. - Activation and representation analysis, probing, model hooks, logits, hidden states, or other model-internals work. - Evaluation and inference infrastructure such as DSPy, LiteLLM, Temporal, Ray, vLLM, PostgreSQL/pgvector, or similar systems. - Next.js, React, TypeScript, data visualization, or experiment dashboards. - Running and serving open-weight models on GPUs and reasoning about latency, throughput, memory, precision, and cost tradeoffs. - Designing adversarial evaluations or testing systems against deliberate attempts to evade detection. WHAT SUCCESS LOOKS LIKE IN THE FIRST SIX MONTHS You will: - Reproduce at least one published model-provenance or verification method and clearly document its capabilities, assumptions, and limitations. - Build a repeatable model-verification runner with versioned inputs, artifacts, metrics, and reports. - Add at least one verification workflow to Construct and make it accessible through the Eldros UI. - Run controlled experiments across base models, fine-tuned models, merged models, quantized models, and known distilled models. - Improve our ability to understand when verification methods succeed, when they fail, and why. - Leave behind production-quality code, tests, tooling, and documentation that another engineer can confidently operate and extend. THIS ROLE IS NOT - A pure research role where work ends with a paper or notebook. - A generic model-training or fine-tuning position. - A frontend-only or backend-only engineering role. - A role where benchmark scores are accepted at face value without understanding how they were produced. - A role for someone who wants to stay within a single layer of the stack. We are looking for someone who enjoys moving between research, experimentation, engineering, and product, and who cares about building systems that produce evidence people can actually trust.
Salary Context
Similar Engineering roles on LokerDollar pay around $170k/yr (range $4.7k–1000k/yr, n=802 active listings).
Market context
- ESTIMATEEstimated pay is 3% above the role median of $170,000/year (n=802 pay-disclosing listings).
- VERIFIEDSentient: 1 postings in the last 3 months, 4 all-time on LokerDollar.
- VERIFIEDCompany first seen May 7, 2026.
- VERIFIEDThis listing first seen Sep 24, 2026.
- VERIFIEDLast verified live Sep 27, 2026.
Openness not stated by employer — check the listing
Frequently asked questions
- Is Applied ML Engineer at Sentient a remote job?
- Yes, Applied ML Engineer at Sentient is remote, but the employer did not state which countries can apply. Check the listing before applying.
- What type of employment is Applied ML Engineer at Sentient?
- This is a full time position.
- How do I apply?
- Click the "Apply" button on this page to go to the official application at Sentient.
Explore related
Market data & reports
Salary & skill-demand research built from our own listings data.
- Indonesia IT Jobs vs Global Remote (2026)Primary analysis of 2,049 listings: methodology, classification rules, downloadable datasets.
- AI-Skill Demand: Indonesia vs Global Remote (2026)10,000+ postings, taxonomy-first classifier, Wilson CIs, pre-registered before analysis.
- Remote ≠ Remote: The Skills That Open Global Work to Indonesians (2026)12,891 remote listings: the highest-paid coding skills are the most geo-locked for Indonesia-based applicants. CC BY 4.0 aggregate dataset.
- The Compliance Layer of the AI Hiring Stack (2026)6,349 remote listings: 77.2% never state who may apply. Methodology and a CC BY 4.0 aggregate dataset.
- Indonesia Hiring Report: Tech vs Non-TechJob demand by field from aggregate open-job counts — never individual listings.
- Indonesia Salary BenchmarkAggregate salary ranges across roles, with open methodology and dataset.
- Indonesian Remote Work Salary & Demand IndexHow much of the global remote job corpus is open to Indonesia, and what it pays (USD) by role.
- Indonesia Quarterly Labor Market ReportLayoffs, funding, salaries & skills per quarter — open aggregates.
- Remote Market Reports by RoleAuto-generated per role family — skills, seniority, companies, salary.
- Global Remote Salary BenchmarkAnnual salary by role & currency, plus the share of listings open worldwide.
From the blog
- Navigating the Global Remote Job Market: Where the Real Opportunities AreWith 5,119 active remote listings, here is how to position yourself for USD-paying roles. Insights on skills, demand, and compensation.
- How to Land a Global Remote Job Paying in USD in 2026With 5,119 active remote listings, here is how to navigate the current job market and what skills are actually driving hiring decisions today.
- Remote Work Realities: What You Need to KnowExplore the realities of remote work and USD-paying jobs. Learn about entry-level positions, salary expectations, and how to find legitimate opportunities.