Nov 2026
9 Mon
10 Tue
11 Wed
12 Thu
13 Fri 09:00 AM – 06:00 PM IST
14 Sat 09:00 AM – 06:00 PM IST
15 Sun
Alosh Denny
Submitted Sep 29, 2026
One-line summary. Quantisation changes the model, but the decision usually sits with whoever owns the serving cost. I want to compare notes on who actually makes that call across teams, and what evidence they require before it ships.
Precision is one of the few changes that crosses every team boundary at once. It is a cost decision, a platform decision, a model quality decision and a hardware decision, and in most organisations nobody owns all four. The result is that it either never happens, or it happens without a validation story.
I have measured the technical side extensively and I have very little visibility into how other teams organise the decision. That asymmetry is exactly why this should be a discussion rather than a talk.
Platform engineering, MLOps, SRE, engineering leaders.
Level: intermediate.
About ten minutes of measurements to give the room a common factual footing, including the failure modes that are not obvious, and then I facilitate rather than present.
Open source project. Research or investigation.
On my own side, treating precision as a purely numerical question. Almost every real difficulty turned out to be about which artefact you are allowed to change and who has to sign off, which is not something I can answer from my own work.
Ask the people who ship this in production rather than infer it from papers.
Running this as a talk against a BOF. I have enough material for a talk, but the interesting half of the question is the half I have no data on, and a room full of platform engineers does.
A set of operational practices, and a way of framing a decision that currently falls between teams.
Current state: in progress.
Tags: #bof #mlops #platformengineering #inference #governance #modelserving
{{ gettext('Login to leave a comment') }}
{{ gettext('Post a comment…') }}{{ errorMsg }}
{{ gettext('No comments posted yet') }}