inhousefyi
← Back to listings

Senior Machine Learning Engineer

Evolution Cloud Services (EVOCS)United States · Posted 4 months ago
Full-timeEst. 232,500 USD
Apply now

Description

EVOCS OVERVIEW

EVOCS was founded with a clear purpose: to help businesses operate more effectively, solve complex challenges, and create opportunities for growth through practical expertise and technology solutions.

As an IT consulting firm, we work with our clients to understand their needs, identify the right technologies, and deliver solutions that improve performance and support their business objectives.

Today, EVOCS is a trusted technology partner to a growing number of organizations and industry leaders. Our team combines technical expertise, business understanding, and a commitment to quality to deliver effective solutions and build lasting client relationships based on responsiveness, consistency, and results.

🎯 Role Overview

As a Senior Machine Learning Engineer, you will be the person we trust with the training side of our AI work. You’ll decide what to build, how to build it, and whether to build it at all. You will be responsible for the quality of the models we ship: the data they learn from, the pipelines that produce them, and the judgment calls that separate a useful model from an expensive one. You’ll be mentoring engineers who’ve never watched a loss curve diverge and felt something.

đŸ§© What you will do

In this role, you will:

  • Be the primary person responsible for the data piece of our AI initiatives. This could be called the unglamorous stuff that decides whether the model works, but this is your passion. You put the enthusiast in MLE.
  • Build and maintain training and retraining pipelines in Azure AI Foundry. We’re talking working on the model catalog, fine-tuning workflows, deployment, drift monitoring, and closing the loop when production data reveals the eval set was lying to you.
  • Make the real model design calls because you know best which way to go: full fine-tune vs. LoRA/QLoRA vs. DPO vs. “better prompting would save us three weeks.” Know when not to train.
  • Run hyperparameter work that isn’t a grid search copied from a 2021 Medium post.
  • Operate distributed training setups and know what breaks at scale. Pick your poison: FSDP, DeepSpeed, Megatron, accelerate, etc.
  • Design eval harnesses that catch what’s actually wrong, with a skeptical eye on benchmark contamination.
  • Ship models into production as the load-bearing piece of the product, not a feature slapped on the side.
  • Mentor engineers who can call an inference endpoint but have never trained one themselves.

🧠 What you will bring

The top candidate will have the following skills:

  • 5+ years of ML engineering experience, with meaningful time spent fine-tuning transformer models end-to-end. We’re not talking notebook demos, we mean real runs with real eval harnesses where you worked through and found the problems and fixed them.
  • Strong Python and PyTorch, plus fluency with the Hugging Face stack (transformers, datasets, accelerate, peft, trl). Bonus for JAX; extra bonus for having read a CUDA kernel and not flinched.
  • You’ve already built or seriously operated a distributed training setup and know how to set that up.
  • Azure AI Foundry experience (or strong Azure ML adjacent and willingness to get deep), plus SQL and at least one data pipeline tool (dbt, Airflow, Dagster, or Spark. We’re not religious).
  • Experiment tracking discipline (W&B, MLflow, or a spreadsheet you defend philosophically) and the usual engineer stuff. You should know Git, Docker, and have the ability to actually ship.
  • Fluent in English (written and spoken) – bilingual or near-native level
  • Strong interpersonal and communication skills – this is a client-facing role that involves frequent interaction via email, calls, and meetings

Ideally you have


  • Run quantized models locally — GGUF, GPTQ, AWQ, MLX — and know what K-quants are and why Q4_K_M is usually the sweet spot.
  • Familiarity with the whisper.cpp / flash-moe universe — efficient inference on hardware that shouldn’t be able to do that. (Spoiler: our projects will go here.)
  • A strong take on MoE routing, speculative decoding, or why KV-cache management is more interesting than it has any right to be.
  • RLHF, DPO, or preference data curation experience.

#LI-DNI

Pay Range for jobs in the US.

Pay Range
$210,000—$255,000 USD

đŸ‘„ Our Values

We are privileged to serve our loyal customer base in our mission to build lasting relationships with our clients based on trust and mutual success. We strive to deliver exceptional quality and consistency through a white-glove approach. By empowering businesses with tailored solutions and insights, we help them achieve their goals and navigate the ever-evolving tech landscape.

The values we live by:

  • Customer-centric Solutions
  • Innovation & Excellence
  • Integrity & Transparency
  • Data-driven Decision Making

📝 Need to Know

The posting will be active for a minimum of 3 days. The active posting will continue to extend by 3 days until the position is filled.

All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, age, disability or protected veteran status, or any other legally protected basis, in accordance with applicable law.

Similar jobs

Echo NeurotechnologiesSan Francisco, California, United States

Est. 205,000 USD

Company Overview Echo Neurotechnologies is an exciting new startup in the Brain-Computer Interface (BCI) space, driving innovation through advanced hardware engineering and AI solutions. Our mission is to deliver cutting


Full-time
XPENGSanta Clara, California, United States

Est. 235,200 USD

XPENG is a leading smart technology company at the forefront of innovation, integrating advanced AI and autonomous driving technologies into its vehicles, including electric vehicles (EVs), electric vertical take-off and


Full-time
XPENGSanta Clara, California, United States

Est. 289,800 USD

XPENG is a leading smart technology company at the forefront of innovation, integrating advanced AI and autonomous driving technologies into its vehicles, including electric vehicles (EVs), electric vertical take-off and


Full-time
PhysicsXLondon, United Kingdom

Est. 120,000 GBP

About us PhysicsX is a deep-tech company with roots in numerical physics and Formula One, dedicated to accelerating hardware innovation at the speed of software. We are building an AI-driven simulation software stack for


Full-time
XPENGSanta Clara, California, United States

Est. 235,200 USD

XPENG is a leading smart technology company at the forefront of innovation, integrating advanced AI and autonomous driving technologies into its vehicles, including electric vehicles (EVs), electric vertical take-off and


Full-time
PhysicsXNew York, New York, United States

Est. 225,000 USD

About us PhysicsX is a deep-tech company with roots in numerical physics and Formula One, dedicated to accelerating hardware innovation at the speed of software. We are building an AI-driven simulation software stack for


Full-time
PhysicsXSingapore, Singapore

About us PhysicsX is a deep-tech company with roots in numerical physics and Formula One, dedicated to accelerating hardware innovation at the speed of software. We are building an AI-driven simulation software stack for


Full-time
VannevarRemote

Est. 182,500 USD

Vannevar is a defense technology company building AI to deter our adversaries. In the 21st century, conflict moves at algorithmic speed and foresight equals firepower. Our agentic AI is purpose-built to compete with Chin


Full-timeRemote
CoreViewItaly

Est. 60,000 EUR

About CoreView CoreView is the global leader in Microsoft 365 (M365) tenant resilience, serving over 23 million users worldwide. We empower the world’s leading organizations to master the complexity of Microsoft M365. Th


Full-time
Scale AILondon, United Kingdom

Est. 120,000 GBP

Machine Learning EngineerLondon, UK About the role Applied Intelligence Systems (AIS) is part of the Scale Generative AI Platform (SGP), focused on pushing the frontier of what agentic applications can do across diverse


Full-time
Particle41Remote

Data Science & Engineering Lead Lead the charge in AI and data innovation as our Data Science & Engineering Lead. Here, you’ll work hands-on with our talented team to build and deploy high-impact ML models, optim


Full-timeRemote
Movable InkRemote

Movable Ink scales content personalization for marketers through data-activated content generation and AI decisioning. The world’s most innovative brands rely on Movable Ink to maximize revenue, simplify workflow and boo


Full-timeRemote

Robots & Pencils is an applied AI engineering firm building the next frontier of business architecture. We design and ship AI co-workers that integrate into enterprise operations and deliver measurable results for ou


Full-timeRemote
Atomic MachinesEmeryville, California, United States

Est. 225,000 USD

Atomic Machines is ushering in a new era of micromanufacturing with its Matter Compilerℱ technology platform. This platform enables new classes of micromachines to be designed and built by providing manufacturing process


Full-time
CaylentRemote

Est. 120,000 USD

Caylent is an AI-first cloud services company that helps organizations turn ambitious ideas into meaningful business impact. As an AWS Premier Tier Services Partner and a charter member of Anthropic’s Claude Partner Netw


Full-timeRemote
AI/ML Developer4 months ago

Est. 120,000 PLN

We are launching a strategic pool of AI specialists to accelerate the growth of our AI Practice and turn strong market demand into scalable, repeatable capabilities. This role goes beyond project delivery-it is about sha


Full-timeRemote
Lightning AIRemote

Est. 237,500 USD

Who We Are Lightning AI is the company behind PyTorch Lightning. Founded in 2019, we build an end-to-end platform for developing, training, and deploying AI systems—designed to take ideas from research to production with


Full-timeRemote
CaylentRemote

Caylent is an AI-first cloud services company that helps organizations turn ambitious ideas into meaningful business impact. As an AWS Premier Tier Services Partner and a charter member of Anthropic’s Claude Partner Netw


Full-timeRemote
PhysicsXSingapore, Singapore

About us PhysicsX is a deep-tech company with roots in numerical physics and Formula One, dedicated to accelerating hardware innovation at the speed of software. We are building an AI-driven simulation software stack for


Full-time
PhysicsXLondon, United Kingdom

Est. 90,000 GBP

About us PhysicsX is a deep-tech company with roots in numerical physics and Formula One, dedicated to accelerating hardware innovation at the speed of software. We are building an AI-driven simulation software stack for


Full-time
VerveRemote

Who We Are Verve has created a more efficient and privacy-focused way to buy and monetize advertising. Verve is an ecosystem of demand and supply technologies fusing data, media, and technology together to deliver result


Full-timeRemote
Echo NeurotechnologiesSan Francisco, California, United States

Est. 195,000 USD

Company Overview Echo Neurotechnologies is an exciting new startup in the Brain-Computer Interface (BCI) space, driving innovation through advanced hardware engineering and AI solutions. Our mission is to deliver cutting


Full-time
Parallel SystemsLos Angeles, California, United States

Est. 200,000 USD

Parallel Systems is pioneering autonomous battery-electric rail vehicles designed to transform freight transportation by shifting portions of the $900 billion U.S. trucking industry onto rail. Our innovative technology o


Full-time
LivefrontRemote

Est. 152,500 USD

At Livefront, we help companies design and build world-class digital products that command attention and inspire joy. We’ve helped household names like CVS, Samsung, General Mills, and Optum create experiences that have


Full-timeRemote
CaylentRemote

Caylent is an AI-first cloud services company that helps organizations turn ambitious ideas into meaningful business impact. As an AWS Premier Tier Services Partner and a charter member of Anthropic’s Claude Partner Netw


Full-timeRemote
Scale AIWashington, District of Columbia, United States

Est. 254,100 USD

The goal of a Senior Machine Learning Engineer at Scale is to leverage techniques in the fields of generative AI, computer vision, reinforcement learning, and agentic AI to improve Scale's products and customer experienc


Full-time
XPENGSanta Clara, California, United States

Est. 328,650 USD

XPENG is a leading smart technology company at the forefront of innovation, integrating advanced AI and autonomous driving technologies into its vehicles, including electric vehicles (EVs), electric vertical take-off and


Full-time

Est. 150,000 USD

Staff Platform Engineer Location: This position is a remote role based in the US Company Overview Robots & Pencils is an applied AI engineering firm building the next frontier of business architecture. We design and


Full-timeRemote
XPENGSanta Clara, California, United States

Est. 328,650 USD

XPENG is a leading smart technology company at the forefront of innovation, integrating advanced AI and autonomous driving technologies into its vehicles, including electric vehicles (EVs), electric vertical take-off and


Full-time
PhysicsXNew York, New York, United States

Est. 170,000 USD

About us PhysicsX is a deep-tech company with roots in numerical physics and Formula One, dedicated to accelerating hardware innovation at the speed of software. We are building an AI-driven simulation software stack for


Full-time