* Adding Module and Example for ECS cluster monitoring with ecs_observer
* Adding Module and Example for ECS cluster monitoring with ecs_observer
* Incorporating PR comments
* Restructuring Examples and modules folder for ECS, Added content in main Readme
* Fixing path as per PR comments
* Parameterzing the config files, incorporated PR review comments
* Adding condition for AMP WS and fixing AMP endpoint
* Adding Document for ECS Monitoring and parameterized some variables
* Added sample dashboard
* Adding Document for ECS Monitoring and parameterized some variables
* Fixing failures detected by pre-commit
* Fixing failures detected by pre-commit
* Fixing failures detected by pre-commit
* Pre-commit fixes
* Fixing failures detected by pre-commit
* Fixing failures detected by pre-commit
* Pre-commit
* Fixing HIGH security alerts detected by pre-commit
* Fixing HIGH security alerts detected by pre-commit
* Fixing HIGH security alerts detected by pre-commit, 31stOct
* Add links after merge
* 2ndNov - Added condiotnal creation for Grafana WS and module versions for AMG, AMP
---------
Co-authored-by: Rodrigue Koffi <bonclay7@users.noreply.github.com>
* Update docs structure with regrouped patterns
* Rename multi cluster to avoid ambiguity with cross account
* Separate troubleshooting for eks-monitoring
* Added example for multi-cluster eks-monitoring and made changes to eks-monitoring module to allow cross-cluster IRSA
* Updated eks-monitoring module's README to add the variable description
* Hard-coded grafana license type in eks-cross-cluster-with-amp/main.tf
* Added README page for eks-cross-account-with-central-amp
* Extracted out the eks/amg creation and modified it to use existing resources
* Update README.md to add multiaccount dashboard png
* Updated README.md and multiaccount.md to add eks multiaccount png
* Updated eks-cross-cluster-with-amp example to disable dashboard creation for cluster 2
* Pre commit changed committed
* Fixed cross-account-observability docs and README.md and added variable for amp_workpace_alias
* Removed extra spacing in multiaccount.md and added precommit suggested changes
* Modified cross-account-observability example to change cross-account-amp-role to snake_case
* Capitalized Terraform string in multiaccount.md and converted iam-role-attach to snake_case
---------
Co-authored-by: Rodrigue Koffi <bonclay7@users.noreply.github.com>
* added docs for ADOT health monitoring
* Update index.md - Added screen-shots of the dashboard
* Update index.md -Corrected some typos and formatting
* Update index.md - Made corrections based on comments
* Update index.md
---------
Co-authored-by: Mathews <kurampil@amazon.com>
* adot-container-insight-tf-code by Rajat Omar
* Documentation and naming convention added
* PR review fixes
* PR CI pipeline fixes
* pr ci run fixes
* added the variables mentioned in ci build
* ci variable fixes
* changed the module name
---------
Co-authored-by: Omar <merajat@3c0630162a5a.ant.amazon.com>
* added workqueue related and apiserver_storage_db_total_size_in_bytes (available since K8s/EKS v1.26+) metrics into kube-admin scrape job
* chnanged OTEL scrape config and recording rules file to make apiserver Grafana dashboards working
* added APISERVER Grafana dashboards into variables, Flux kustomization
* added original kube-prom-stack kube-apiserver scrape config into OTEL
* Revert "added original kube-prom-stack kube-apiserver scrape config into OTEL"
because this scrape config does not work :-(
This reverts commit 715db895c656af09430e3d6ada2087bd1828413f.
* added API server troubleshoting dashboard
* removed empty line for clarity
* make API serve rmonitoring default to true but can be disabled as well
* Update eks-apiserver.md
* updated eks-monitoring/README.md according to pre-commit
---------
Co-authored-by: Jens-Uwe Walther <waltju@amazon.com>
* Typo
* Remove Grafana provider
* Temp: move dashbaords to gitOps
* Move external labels to resource attributes
* Avoid DDoS with using 0.0.0.0
* Pre-commit
* Transition in two steps
Will need to remove provider in a separate version to provide a transition path as removing this will break terraform and leave orphans in the state
* Move patterns' dashboards creation to gitOps
Standardize config objects for patterns as well
* Pre-commit
* Create AMP dashboard from external source with Grafana provider
* Fix deprecated option
* Fix Flux requirements
* Run pre-commit
* Update example with operator
* Cleanup examples
* Update multicluster example
* Update multicluster example
* Drop dead variable
* Update docs
* Change GitOps branch name
* Update docs
* Replacing Secrets Manager to SSM to store Grafana API Key (#178)
* Fixing SSM
* Fixing SSM
* Replacing Secrets Manager with SSM
* Replacing Secrets Manager with SSM
* Update architecture diagram
* Update architecture diagram
* Update README.md
* Update index.md
* Fixing Grafana Operator Version
* Fix multicluster example
* Update docs
---------
Co-authored-by: Ela AWS <51791117+elamaran11@users.noreply.github.com>
Co-authored-by: Elamaran Shanmugam <elamaran.shan@gmail.com>
* Adding EKS multicluster observability example
* EKS multicluster example - precommit fix
* Corrected the path to an example
* Made the multicluster example simpler
* Pre-commit changes
* Saved the images to Github and linked them in docs
* Comments and language edits
* Fix trailing spaces
* Non-controversial naming and simpler variables
* Added region to the data gathering
* Formatting changes
---------
Co-authored-by: Rodrigue Koffi <bonclay7@users.noreply.github.com>
* Import and customize fluenbit add-on
* Enable fluent bit logs
* Bump helm addon version
* Dropping account id as it seems to create scraping errors
* Create separate log groups per namespace
* Apply pre-commit
* Remove conflicting global label
* Add config object for logs
* Enable logs in examples
* Add logs docs
* Fix broken link
* Add screenshots
* Update docs
* Typos
* Managed Grafana Workspace with Identity Centre Users (#83)
* update kuberenetes and instance type
* initial setup of managed grafana workspace and identity centre identities
* cleanup
* run precommit
* output grafana workspace ID
* add identity store id variable
* remove API key
* update outputs naming convention as per terraform guidelines
* update docs and add versions
* update variables type
* update naming conventions
* add managed policy arn for querying promethues
* update readme
* update workshop references to this
* add role arn type
* cleanup and simplification
* remove workshop
---------
Co-authored-by: charlie keegan <chakeega@amazon.com>
Co-authored-by: Rodrigue Koffi <bonclay7@users.noreply.github.com>
* Rename example
* Update grafana example and base module references
* Update example's reference
* Cleanup and docs ref
* Add docs
* Update docs
* TODO: add link after merge
* Update managed-grafana.md
---------
Co-authored-by: Charlie Keegan <91210223+charliekeeegan@users.noreply.github.com>
Co-authored-by: charlie keegan <chakeega@amazon.com>
Co-authored-by: Mark Beacom <7315957+mbeacom@users.noreply.github.com>
* Added amp+xray images for o11y-accelerator
Added draw.io and png images with AMP and X-Ray images for AWS Observability Accelerator
* updated doc with images
Updated doc with latest Observability accelerator architecture image and minor changes to draw.io and architicture diagram.
* added light and dark version of architecture diagram
Added draw.io and png files - both dark and light background version
* updated light background diagrams
updated light background diagrams
* Delete o11y-accelerator-amp-xray.drawio
* Delete o11y-accelerator-amp-xray.png
* Update dark-o11y-accelerator-amp-xray.png
* update
* moved images to docs
moved images to docs
* Delete dark-o11y-accelerator-amp-xray.drawio
* Delete dark-o11y-accelerator-amp-xray.png
* Delete light-o11y-accelerator-amp-xray.drawio
* Delete light-o11y-accelerator-amp-xray.png
---------
Co-authored-by: Rodrigue Koffi <bonclay7@users.noreply.github.com>
* Added parameter for instance type
See #89.
Added a variable to change the default instance type for the EKS Cluster
* Added additional parameters for size and version