Raymond Ops
Author

Raymond Ops

Linux ops automation, cloud-native, Kubernetes, SRE, DevOps, Python, Golang and related tech discussions.

723
Articles
0
Likes
5.5k
Views
0
Comments
Recent Articles

Latest from Raymond Ops

100 recent articles max
Raymond Ops
Raymond Ops
Jul 17, 2026 · Operations

Linux Disk Space Alerts? Locate the Problem in 3 Quick Steps

When a Linux disk space alert fires, this guide walks you through three rapid steps—identifying the affected partition, deep‑diving with df, du, ncdu and custom scripts, and cleaning up logs, caches, inodes, LVM, Docker, and quotas—to quickly pinpoint and resolve the root cause.

DockerLVMLinux
0 likes · 44 min read
Linux Disk Space Alerts? Locate the Problem in 3 Quick Steps
Raymond Ops
Raymond Ops
Jul 16, 2026 · Operations

Nginx Configuration Optimization: Mastering Worker Processes for Performance Tuning

This guide explains Nginx's multi‑process architecture, shows how to bind worker processes to CPU cores, tune worker connections, configure upstream load‑balancing, enable proxy buffering, keepalive, gzip/Brotli compression, SSL/TLS settings, and provides testing and troubleshooting scripts for high‑performance deployments.

GZIPNginxSSL
0 likes · 35 min read
Nginx Configuration Optimization: Mastering Worker Processes for Performance Tuning
Raymond Ops
Raymond Ops
Jul 14, 2026 · Cloud Native

Kubernetes Networking: From CNI Basics to Troubleshooting

This article explains Kubernetes' three‑principle network model, compares the leading CNI plugins (Flannel, Calico, Cilium), details pod communication paths, Service and Ingress mechanisms, DNS and NetworkPolicy implementations, and provides step‑by‑step troubleshooting cases with performance data and concrete configuration examples.

CalicoDNSFlannel
0 likes · 34 min read
Kubernetes Networking: From CNI Basics to Troubleshooting
Raymond Ops
Raymond Ops
Jul 14, 2026 · Operations

Network Troubleshooting with tcpdump & Wireshark: Step‑by‑Step Guide and Ready‑to‑Use Scripts

This comprehensive guide walks you through using tcpdump and Wireshark for network fault isolation, covering core concepts, capture filters, detailed analysis techniques, performance tuning, expert information interpretation, automation scripts, and best‑practice recommendations for efficient packet‑level troubleshooting.

LinuxNetwork TroubleshootingWireshark
0 likes · 51 min read
Network Troubleshooting with tcpdump & Wireshark: Step‑by‑Step Guide and Ready‑to‑Use Scripts
Raymond Ops
Raymond Ops
Jul 13, 2026 · Operations

Scaling Prometheus to Thousands of Nodes with Thanos: Architecture, Storage, and HA Practices

The article analyzes the storage, query performance, high‑availability, and data‑loss challenges of running Prometheus on a 1,000‑node Kubernetes cluster and demonstrates how a Thanos‑based architecture—Sidecar, Query, Store Gateway, Compactor, Receiver, and object‑storage back‑ends—can be designed, tuned, and operated to achieve horizontal scalability, efficient down‑sampling, and reliable fault recovery.

High AvailabilityKubernetesObject Storage
0 likes · 35 min read
Scaling Prometheus to Thousands of Nodes with Thanos: Architecture, Storage, and HA Practices
Raymond Ops
Raymond Ops
Jul 13, 2026 · Operations

Systematic Root‑Cause Analysis for NFS Mount Failures

This guide presents a systematic, layer‑by‑layer methodology for diagnosing and resolving NFS share mount failures, covering version differences, architecture layers, key daemons, RPC mechanisms, common error messages, detailed server and client checks, firewall and SELinux considerations, performance tuning, high‑availability setups, and monitoring.

KerberosLinuxNFS
0 likes · 50 min read
Systematic Root‑Cause Analysis for NFS Mount Failures
Raymond Ops
Raymond Ops
Jul 12, 2026 · Operations

Essential Port Connectivity Troubleshooting: A Complete Step‑by‑Step Guide

This guide walks you through a systematic, seven‑layer approach to diagnosing port connectivity failures on Linux systems, covering service listening checks, local firewall rules, SELinux policies, network path analysis, cloud security groups, and application‑level protocols, with concrete commands, scripts, case studies, best‑practice recommendations, and monitoring tips.

LinuxSELinuxfirewall
0 likes · 40 min read
Essential Port Connectivity Troubleshooting: A Complete Step‑by‑Step Guide
Raymond Ops
Raymond Ops
Jul 11, 2026 · Cloud Native

Kubernetes HPA & VPA Auto-Scaling: Elastic Strategies for Traffic Spikes

An in‑depth comparison of Kubernetes Horizontal and Vertical Pod Autoscalers—including algorithms, configurations, performance benchmarks, mixed‑mode trade‑offs, custom‑metric integrations, and real‑world case studies—demonstrates how to choose and tune HPA, VPA, and KEDA for rapid traffic spikes while avoiding conflicts.

Cloud NativeHPAKEDA
0 likes · 47 min read
Kubernetes HPA & VPA Auto-Scaling: Elastic Strategies for Traffic Spikes