Deployment and MLOps

Shipping models to production and keeping them healthy.

  • 40 Tracked terms
  • Last 30 days Feed window

What this topic collects on

An article joins this feed when it matches these terms. Each one is also a search of its own.

Latest in Deployment and MLOps

DEV Community
dev.to > darkpandawarrior > one-global-threshold-is-how-you-delete-valid-data-303h

One global threshold is how you delete valid data

1+ hour, 46+ min ago   (388+ words) Our GPS filter worked perfectly, right up until someone sat in Bangalore traffic. A parked phone does not sit still in the data. The reported position wanders a few metres in every direction, and if you naively sum the gaps…...

DEV Community
dev.to > arhuman > the-logging-dilemma-58po

The Logging Dilemma

3+ hour, 43+ min ago   (123+ words) Every developer has lived through this scene: the adrenaline spike when a production incident is... Tagged with go, monitoring, logging, module....

DEV Community
dev.to > sergeisolod > my-fdcservers-vps-was-online-disk-reads-were-taking-187-seconds-4796

My FDCServers VPS Was Online. Disk Reads Were Taking 18.7 Seconds

9+ hour, 9+ min ago   (1599+ words) I actually liked FDCServers. That is probably the strangest way to start an article about a VPS that eventually became unusable enough for me to cancel it. For quite a while, the server worked normally. It handled real production traffic....

DEV Community
dev.to > gitgo_5662 > first-rollback-revert-the-agent-pr-you-cannot-explain-3akh

First Rollback: Revert the Agent PR You Cannot Explain

9+ hour, 10+ min ago   (783+ words) Your first AI pull request will often need rollback. Plan that rollback before you merge anything. You lack repo history on day one. Agents still produce large and confident diffs today. A rollback plan keeps that blast radius tiny. First…...

Medium
medium.com > @optimzationking2 > when-autoscaling-makes-an-outage-worse-bd77a264dabc

When Autoscaling Makes an Outage Worse

6+ hour, 30+ min ago   (30+ words) The traffic suddenly spikes. CPU reaches 85%. The autoscaler notices. More instances appear. Then more. Then more. Your dashboard looks like the system is …...

Medium
medium.com > serverinspector > what-is-uptime-monitoring-a-practical-look-at-synthetic-checks-14fdf59e71c6

What Is Uptime Monitoring? A Practical Look at Synthetic Checks

7+ hour, 8+ min ago   (350+ words) Originally published at https://serverinspector.com/blog/what-is-uptime-monitoring-synthetic-checks Imagine your website going dark at …...

DEV Community
dev.to > sergueyasaelshinder > the-meter-is-running-on-every-request-2pc2

The Meter Is Running on Every Request

12+ hour, 11+ min ago   (338+ words) The demo cost four pence. You ran it thirty times while building it, glanced at the total, and stopped thinking about it, because four pence is not a number anybody worries about. Then it shipped, and the cost stopped being…...

Monte Carlo
montecarlo.ai > blog-five-failure-modes-evals-wont-catch

Five Failure Modes Evals Won't Catch And What To Do About Them

14+ hour, 25+ min ago   (1076+ words) Evals are a critical part of every data and AI team’s agent development process. An engineer builds an eval, defines what a bad answer looks like, runs a judge against a test set, and ships when the score looks good....

The New Stack
thenewstack.io > pg-99-conf-2026-inference-costs

Chip Huyen explains how to cut inference costs without new hardware

14+ hour, 48+ min ago   (775+ words) Ahead of her return to P99 CONF on October 21–22, Chip Huyen’s advice on cutting AI inference costs gets a fresh look in the age of agents....