Task: Develop and add website performance and operation monitoring
Develop and add website performance and operation monitoring
03.10.2026haih.site
Select and implement lightweight website monitoring without Sentry: logs, metrics, dashboards, and basic alerts.
Objective
Add a lightweight observability system for the website to quickly understand:
- whether the website and its key services are running;
- where errors occur;
- how performance is changing;
- whether there is degradation in response time, load, or resources.
Constraints and Preferences
- Do not use Sentry as the primary tool.
- Avoid heavy and redundant infrastructure.
- Preference for simple, clear, and inexpensive/self-hosted solutions.
- Consider a structured logs + Prometheus + Grafana stack and lighter alternatives if they solve the task more simply.
What to Develop
- Determine the minimum set of observability for the current website architecture.
- Configure centralized collection and convenient viewing of application/web server logs.
- Add metrics at least for:
- availability;
- HTTP status / error rate;
- latency / response time;
- request rate;
- CPU, RAM, disk, and other important host/container resources, if applicable.
- Add application metrics for critical website operations if necessary.
- Build a compact health and performance dashboard.
- Add basic alerts for critical failure states and noticeable degradation.
- Check the impact of monitoring itself on resources and prevent unjustified operational complexity.
Candidates
Baseline option for evaluation:
- structured application/web server logs;
- Prometheus;
- exporters / application with
/metrics; - Grafana;
- lightweight log backend (e.g., Loki or equivalent) if necessary and justified.
A simpler stack is acceptable if it provides sufficient diagnosability and lower operational costs.
Deliverables
- minimum monitoring stack selected and justified;
- monitoring deployed and connected to the website;
- logs and main metrics available;
- dashboard with key performance indicators available;
- minimum necessary alerts configured;
- documentation on where to check system health and how to diagnose typical issues.