Daniel Andrzejewski
6c5f708179
node_disk_write_time_seconds_total is in seconds, not in milliseconds. node_disk_write_time_seconds_total should be grater than 0, otherwise you get +Inf result.
2020-09-17 15:13:42 +02:00
Yashar Nesabian
d6b39a7f3f
More accurate alerts
...
added `mondodb instance down` alert and changed the `too many
connections` alert to fire when the connections are more than 80% of the
available connections.
removed `mongodb_replset_member_state` based alerts as I don't have
enough information on them
2020-08-09 10:35:39 +04:30
Yashar Nesabian
3ce1084f5b
Added percona mongodb alert rules
2020-08-03 10:45:32 +04:30
kaifen.xie
a04eef39c0
add istio
2020-07-25 23:24:36 +08:00
Nirav Chotai
8fb5da83de
Fix HPA alerts
...
- Fixing KubernetesHpaMetricAvailability
- Fixing KubernetesHpaScalingAbility
2020-07-24 13:32:44 +08:00
Ozarklake
88e812c78e
add sql server rules
2020-07-17 15:02:41 +08:00
Ozarklake
4e66d17d01
add sql server rules
2020-07-17 14:58:26 +08:00
Ozarklake
e009c5d8b5
Optimizing mysql slow query alert rules
2020-07-14 12:55:17 +08:00
Mansur Marvanov
05e521c0a8
Fix PrometheusJobMissing alert
2020-07-09 16:36:45 +09:00
tux
add6d9c2f3
Add official rabbitmq exporter rules
2020-06-30 15:48:42 +02:00
Nabil BENDAFI
edbc9cac2b
fix: remove unnecessary test
2020-06-24 14:53:03 +02:00
Nabil BENDAFI
42b2dc07a6
Fix numbering
2020-06-24 14:22:31 +02:00
Nabil BENDAFI
b324c6f32f
feat(traefik): add rules for Traefik v2
...
Fixes #7
2020-06-23 13:40:01 +02:00
Nabil BENDAFI
5e51c3daef
Fix data-clipboard-target-id unicity
2020-06-17 14:57:06 +02:00
Mickaël Canévet
24f7095cd5
Fix HAProxy rules
2020-05-29 10:11:54 +02:00
dependabot[bot]
2f1a1b4670
Bump activesupport from 6.0.2.1 to 6.0.3.1
...
Bumps [activesupport](https://github.com/rails/rails ) from 6.0.2.1 to 6.0.3.1.
- [Release notes](https://github.com/rails/rails/releases )
- [Changelog](https://github.com/rails/rails/blob/v6.0.3.1/activesupport/CHANGELOG.md )
- [Commits](https://github.com/rails/rails/compare/v6.0.2.1...v6.0.3.1 )
Signed-off-by: dependabot[bot] <support@github.com>
2020-05-26 16:18:47 +00:00
Ilya Kisleyko
663b0e94da
check free space for all mountpoints
2020-05-20 20:04:32 +03:00
Anton Smolkov
bbbe14f2bd
Update rules.yml
...
WMI memory alert had opposite meaning, triggered on 90% free instead of 90% used
2020-05-19 11:07:11 +03:00
Fernando Carletti
e6de413146
fix: container ContainerMemoryUsage alert
2020-05-18 17:38:05 -05:00
Rob Brown
5050fd64d5
Correct "device" to "interface"
2020-05-14 16:57:19 +01:00
Samuel Berthe
4cd3ff1d4a
Merge pull request #117 from samber/replace-severity-error-critical
2020-05-14 17:22:37 +02:00
Samuel Berthe
da1e4f6301
💄 replacing "error" severity by "critical", repo wide
2020-05-14 17:20:19 +02:00
Rob Brown
5d3e812fd7
Add HostNetworkNot1GbSpeed rule
2020-05-14 15:00:24 +01:00
Samuel Berthe
7293bca720
Merge pull request #107 from robert-will-brown/NetworkTransmitErrors
2020-05-09 21:32:40 +02:00
Samuel Berthe
b081f28f5d
Merge pull request #112 from robert-will-brown/SpeedTestExporter
2020-05-09 21:31:33 +02:00
Samuel Berthe
660312d0ea
fix OOM killer threshold
2020-05-09 21:25:13 +02:00
Samuel Berthe
6d6b41e241
Merge pull request #108 from robert-will-brown/EdacMemoryErrors
2020-05-09 21:23:01 +02:00
Rob Brown
8faa295745
Add SpeedTest stanza
2020-05-09 10:20:55 +01:00
Rob Brown
ee4e046c66
Add "> 0" at the end of NetworkTransmitErrors queries
2020-05-09 10:18:21 +01:00
Samuel Berthe
d5f6388899
renaming some mysql alerts
2020-05-09 02:11:18 +02:00
Rob Brown
5d83e393cc
Add initial Speedtest Exporter rules
2020-05-08 15:25:54 +01:00
Rob Brown
8912db93bc
Fix "greater than" value
2020-05-04 19:04:52 +01:00
Rob Brown
4b22c078ea
Align EDAC errors with comments
2020-05-04 18:47:20 +01:00
Samuel Berthe
718cd2188c
shame on me
2020-05-04 00:10:43 +02:00
Samuel Berthe
eb8dc736a3
improve acuracy for context switching query
2020-05-04 00:05:33 +02:00
Samuel Berthe
790139211e
fix typo: postgresql replication lag
2020-05-03 23:23:21 +02:00
Samuel Berthe
773b3456d2
renaming sms to pager
2020-05-03 21:40:45 +02:00
Samuel Berthe
648b83250a
improve accuracy "Kubernetes Pod not healthy" query
2020-05-03 18:01:25 +02:00
Samuel Berthe
81349d939f
Merge pull request #111 from zales/master
2020-05-03 18:00:41 +02:00
Ondrej Zalesky
d3d13946e6
fix "Kubernetes Pod not healthy" query
2020-04-30 22:53:25 +02:00
Rob Brown
981e82d649
Add HostEDACUncorrectableErrorsdetected and HostEDACCorrectableErrorsdetected rules
2020-04-30 13:27:30 +01:00
Rob Brown
f87e6d300d
Added spacing as per standard
2020-04-30 12:39:12 +01:00
Rob Brown
c57a5e6e36
Add HostNetworkReceiveErrors and HostNetworkTransmitErrors rules
2020-04-30 12:38:23 +01:00
Samuel Berthe
951d80121f
Merge branch 'master' of github.com:samber/awesome-prometheus-alerts
2020-04-06 09:13:29 +02:00
Samuel Berthe
e97023d2a4
linkerd2: adding first rule
2020-04-06 09:01:51 +02:00
Samuel Berthe
a8a8950c01
Merge pull request #101 from Seljuke/master
...
FIX KubernetesPodnothealthy Alert
2020-04-02 21:22:20 +02:00
Selçuk Arıbalı
c98a04784e
FIX KubernetesPodnothealthy Alert
...
Kube state metrics assigns value of current pod phase with 1, so according to that Kubernetes Pod not healthy fixed.
2020-04-02 21:01:04 +03:00
Samuel Berthe
c20227b458
oops: adding one-to-one vector matching to mysql subqueries
2020-03-31 16:02:28 +02:00
Samuel Berthe
7f05b0cbc4
Merge pull request #98 from mcrauwel/extra-mysql-checks
...
added some extra MySQL checks
2020-03-31 16:00:13 +02:00
Matthias Crauwels
79b5ad3b5d
removed avg grouping where possible
2020-03-31 11:42:05 +02:00