All systems operational. This profile is monitored 24/7 by caffeine-driven synthetic checks.
apiVersion: human/v1
kind: SiteReliabilityEngineer
metadata:
name: fabian-m
namespace: santiago-chile π¨π±
labels:
role: sre
specialty: observability
experience: 3y+
spec:
education: Computer Engineer @ Duoc UC π
languages:
- Spanish (native) π¨π±
- English (professional) πΊπΈ
focus:
- High-Availability Systems
- Golden Signals
- SLI/SLO & MTTR Optimization
- Incident & Root Cause Analysis
status:
phase: Running β
workingOn: π Observability at scale
askMeAbout:
- SLOs
- Grafana
- Incident Management
- Alert noise reductionflowchart TD
A["βοΈ Services<br/>K8s Β· RabbitMQ Β· Redis Β· AWS"] --> B["π‘ Metrics<br/>Prometheus Β· Datadog"]
A --> C["π Logs<br/>Loki Β· Splunk"]
B --> D["π Grafana dashboards"]
C --> D
D --> E{"π― SLO breached?"}
E -- no --> F["π΄ Sleep well"]
E -- yes --> G["π¨ Correlate Β· BigPanda"]
G --> H["π§βπ Respond Β· ServiceNow"]
H --> I["π Postmortem & RCA"]
I -. "improve" .-> A
classDef ok fill:#064E3B,stroke:#10B981,color:#fff
classDef alert fill:#7F1D1D,stroke:#EF4444,color:#fff
classDef core fill:#4C1D95,stroke:#A855F7,color:#fff
class A,B,C,D core
class F ok
class G,H,I alert
Observability architecture with dashboards for K8s, RabbitMQ, Redis, AWS RDS & S3.
Root cause analysis, SLI/SLO framework & alert noise reduction.
High-availability architecture with centralized logging & synthetic monitors.
Infrastructure monitoring across multiple regions with IaC deployment.
π A snake eating my contributions β the only acceptable kind of data loss.
π¨ Need an SRE? Click to open the runbook
# Step 1 β Identify the issue
$ echo "Need better observability"
# Step 2 β Escalate to the right person
$ page --to "FabiΓ‘n M." --severity "let's-talk" \
--channel linkedin.com/in/fabianimv \
--fallback fabianignaciomv@gmail.com
# Step 3 β Resolution
β Dashboards built
β SLOs defined
β Alert noise reduced
β Team sleeps well




