* added workqueue related and apiserver_storage_db_total_size_in_bytes (available since K8s/EKS v1.26+) metrics into kube-admin scrape job
* chnanged OTEL scrape config and recording rules file to make apiserver Grafana dashboards working
* added APISERVER Grafana dashboards into variables, Flux kustomization
* added original kube-prom-stack kube-apiserver scrape config into OTEL
* Revert "added original kube-prom-stack kube-apiserver scrape config into OTEL"
because this scrape config does not work :-(
This reverts commit 715db895c656af09430e3d6ada2087bd1828413f.
* added API server troubleshoting dashboard
* removed empty line for clarity
* make API serve rmonitoring default to true but can be disabled as well
* Update eks-apiserver.md
* updated eks-monitoring/README.md according to pre-commit
---------
Co-authored-by: Jens-Uwe Walther <waltju@amazon.com>
* Typo
* Remove Grafana provider
* Temp: move dashbaords to gitOps
* Move external labels to resource attributes
* Avoid DDoS with using 0.0.0.0
* Pre-commit
* Transition in two steps
Will need to remove provider in a separate version to provide a transition path as removing this will break terraform and leave orphans in the state
* Move patterns' dashboards creation to gitOps
Standardize config objects for patterns as well
* Pre-commit
* Create AMP dashboard from external source with Grafana provider
* Fix deprecated option
* Fix Flux requirements
* Run pre-commit
* Update example with operator
* Cleanup examples
* Update multicluster example
* Update multicluster example
* Drop dead variable
* Update docs
* Change GitOps branch name
* Update docs
* Replacing Secrets Manager to SSM to store Grafana API Key (#178)
* Fixing SSM
* Fixing SSM
* Replacing Secrets Manager with SSM
* Replacing Secrets Manager with SSM
* Update architecture diagram
* Update architecture diagram
* Update README.md
* Update index.md
* Fixing Grafana Operator Version
* Fix multicluster example
* Update docs
---------
Co-authored-by: Ela AWS <51791117+elamaran11@users.noreply.github.com>
Co-authored-by: Elamaran Shanmugam <elamaran.shan@gmail.com>
* Adding EKS multicluster observability example
* EKS multicluster example - precommit fix
* Corrected the path to an example
* Made the multicluster example simpler
* Pre-commit changes
* Saved the images to Github and linked them in docs
* Comments and language edits
* Fix trailing spaces
* Non-controversial naming and simpler variables
* Added region to the data gathering
* Formatting changes
---------
Co-authored-by: Rodrigue Koffi <bonclay7@users.noreply.github.com>
* Import and customize fluenbit add-on
* Enable fluent bit logs
* Bump helm addon version
* Dropping account id as it seems to create scraping errors
* Create separate log groups per namespace
* Apply pre-commit
* Remove conflicting global label
* Add config object for logs
* Enable logs in examples
* Add logs docs
* Fix broken link
* Add screenshots
* Update docs
* Typos
* Managed Grafana Workspace with Identity Centre Users (#83)
* update kuberenetes and instance type
* initial setup of managed grafana workspace and identity centre identities
* cleanup
* run precommit
* output grafana workspace ID
* add identity store id variable
* remove API key
* update outputs naming convention as per terraform guidelines
* update docs and add versions
* update variables type
* update naming conventions
* add managed policy arn for querying promethues
* update readme
* update workshop references to this
* add role arn type
* cleanup and simplification
* remove workshop
---------
Co-authored-by: charlie keegan <chakeega@amazon.com>
Co-authored-by: Rodrigue Koffi <bonclay7@users.noreply.github.com>
* Rename example
* Update grafana example and base module references
* Update example's reference
* Cleanup and docs ref
* Add docs
* Update docs
* TODO: add link after merge
* Update managed-grafana.md
---------
Co-authored-by: Charlie Keegan <91210223+charliekeeegan@users.noreply.github.com>
Co-authored-by: charlie keegan <chakeega@amazon.com>
Co-authored-by: Mark Beacom <7315957+mbeacom@users.noreply.github.com>
* Added amp+xray images for o11y-accelerator
Added draw.io and png images with AMP and X-Ray images for AWS Observability Accelerator
* updated doc with images
Updated doc with latest Observability accelerator architecture image and minor changes to draw.io and architicture diagram.
* added light and dark version of architecture diagram
Added draw.io and png files - both dark and light background version
* updated light background diagrams
updated light background diagrams
* Delete o11y-accelerator-amp-xray.drawio
* Delete o11y-accelerator-amp-xray.png
* Update dark-o11y-accelerator-amp-xray.png
* update
* moved images to docs
moved images to docs
* Delete dark-o11y-accelerator-amp-xray.drawio
* Delete dark-o11y-accelerator-amp-xray.png
* Delete light-o11y-accelerator-amp-xray.drawio
* Delete light-o11y-accelerator-amp-xray.png
---------
Co-authored-by: Rodrigue Koffi <bonclay7@users.noreply.github.com>
* Added parameter for instance type
See #89.
Added a variable to change the default instance type for the EKS Cluster
* Added additional parameters for size and version