Skip to content
F Fadhil Mochammad ML engineer · Stockholm
About Posts Projects Notes
EXPLORER 24 notes
Agents & LLMs9
  • Agents
  • Evaluation-driven development
  • Golden datasets
  • Groundedness
  • Human review
  • Online evaluation
  • RAG
  • Retrieval recall
  • Feedback triage
Experimentation5
  • A/B testing
  • Multi-armed bandits
  • Metric semantic layer
  • Self-service tooling
  • Thompson sampling
ML platform4
  • Kubernetes
  • SDK design
  • Model serving
  • Distributed traces
System design3
  • Go worker pools
  • P95 latency
  • Observability
Photography1
  • Exposure triangle
Music1
  • FM synthesis
Meta1
  • Start here

No notes match that search.

vault / systems / observability.md

Observability

System design 51 words 3 outgoing 5 backlinks

For feedback investigation, record enough of the request to reconstruct what happened: the question, retrieval, response and relevant configuration. A trace helps connect those steps. Dashboards can show a pattern, but a concrete replay is often what tells us which component or document needs attention.

Read the full example: related post.

LINKS IN THIS NOTE
[[ Distributed traces ]] A trace follows a request through the services and steps it touches. It lets a slow or failed prediction be inspected as one chain rather than unrelated log lines. In a serving interface, tracing is useful as a shared platform... [[ P95 latency ]] P95 describes the point below which 95% of measured request times fall. It says more about slow requests than an average, but still needs a workload and measurement window. A load-test result is not a promise about production traffic. Bounded... [[ Feedback triage ]] A thumbs-down starts an investigation, not a diagnosis. The workflow I built replays the question, collects evidence, distinguishes a content gap from a configuration issue, and routes the case to an owner. Routing is ordinary code; the model proposes the...
GRAPHdrag · scroll · click
BACKLINKS Go worker pools A bounded worker pool separates accepting work from doing it. The queue and worker limit make overload a visible choice: wait, reject, or shed work rather than spawn without limit. The firmware comparison in the post is about keeping handlers... Kubernetes Kubernetes stores desired state while controllers keep trying to make the running system match it. A Deployment manages replicas, a Service gives traffic a stable destination, and Helm packages configuration. That mental model is more useful than memorising object names... Online evaluation Offline examples help compare changes under controlled conditions. Production feedback shows questions and failures the reference set missed. The two complement each other. Online signals need investigation: a thumbs-down tells us that something went wrong, not whether retrieval, content, configuration... Distributed traces A trace follows a request through the services and steps it touches. It lets a slow or failed prediction be inspected as one chain rather than unrelated log lines. In a serving interface, tracing is useful as a shared platform... Feedback triage A thumbs-down starts an investigation, not a diagnosis. The workflow I built replays the question, collects evidence, distinguishes a content gap from a configuration issue, and routes the case to an owner. Routing is ordinary code; the model proposes the...
OUTGOING
Distributed traces P95 latency Feedback triage
F
© 2026 Fadhil Mochammad · Stockholm ● agents ● experimentation ● data ● ml-platform ● systems ● creative