Veille Technologique — Bilan hebdomadaire

Semaine 2026-W19 · du 2026-05-04 au 2026-05-10

27 feeds · 203 articles traités · 40 sélectionnés · gemma4:e2b · 2026-05-04T08:29:18.863Z

Divers

Looking for hands-on DevOps experience — happy to contribute to real projects

r/devops (Tier 1) · Publié : 2026-05-02 · 77.3/10

Hi everyone, I’m a QA engineer (~3 YOE) based in Pune, India, transitioning into DevOps. Over the last 6–7 months, I’ve completed 100+ hands-on labs (KodeKloud) and worked with tools like Kubernetes, Docker, Linux, Terraform, AWS, Jenkins, ArgoCD, Python, Grafana, and Prometheus. I’m looking for opportunities to contribute to real-world projects (personal, open-source or professional) to gain practical experience before applying for DevOps roles. I’m happy to help, learn, and collaborate — compe

How to monitor your Kubernetes cluster with the OpenTelemetry Collector using the agent + gateway pattern

r/devops (Tier 1) · Publié : 2026-05-04 · 76.1/10

I've worked in the observability industry for a while and set up a lot of collectors for customers. Wanted to put together an end to end writeup covering the things most blogs skip when monitoring a Kubernetes cluster with OTel. Covers the full agent + gateway pattern. Agent runs as a DaemonSet for node local stuff (container logs, kubelet stats, hostmetrics, OTLP from local pods). Gateway is a Deployment for cluster-wide telemetry(k8s_events, k8s_cluster, and the only connection to your bac

Implementing Security-First CI/CD: A Hands-On Guide to DevSecOps Automation

DZone DevOps (Tier 1) · Publié : 2026-04-28 · 72.6/10

Editor’s Note: The following is an article written for and published in DZone’s 2026 Trend Report,  Security by Design: AI Defense, Supply Chain Security, and Security-First Architecture in Practice . DevSecOps means security is part of software delivery from the beginning, where security is built into planning, coding, building, testing, releasing, and operations. As pipelines become faster and more automated, security checks should run inside the CI/CD pipeline and be enforceable across d

How to build CI/CD observability at scale

GitLab Blog (Tier 1) · Publié : 2026-04-28 · 72.5/10

CI/CD optimization starts with visibility. Building a successful DevOps platform at enterprise scale should include understanding pipeline performance, job execution patterns, and quantifiable operational insights — especially for organizations running GitLab self-managed instances. To help GitLab customers maximize their platform investments, we developed the GitLab CI/CD Observability solution as part of our Platform Excellence program, which transforms raw pipeline metrics into actionable ope

Customize preconfigured views for AWS, Azure, and Google Cloud with Cloud Provider Observability in Grafana Cloud

Grafana Blog (Tier 1) · Publié : 2026-04-27 · 67.5/10

Part of what makes Cloud Provider Observability in Grafana Cloud really useful is that it gives you prebuilt dashboards and drill-downs for AWS, Azure, and Google Cloud. Out of the box you get service overviews, instance-level views, and quick links to explore your data. However, you might already have dashboards you trust, want a view tailored to your team’s workflow, or need to change which panels show up when you drill into a single instance. The good news: you can now customize all of that w

I started a DevOps YouTube channel and would love feedback / ideas on learning content

r/devops (Tier 1) · Publié : 2026-05-02 · 55.7/10

Hello DevOpsers! I am a big video learner, and learned lots about devops through videos in the last 10 years or so. There are some nice channels like learndevopswithnana for example. This influenced me to start a devops channel of my own on YouTube, and would love to hear your thoughts on what direction I could go in.. I am big on analogies to explain things. For example trying to explain what cloud networks are using analogies. What are some of the topics you really struggle to understand still

Is "building a Docker image" during the CI pipeline considered a best practice?

r/devops (Tier 1) · Publié : 2026-05-01 · 54.9/10

Hi everyone! I am new to DevOps and trying to better understand CI/CD best practices. I often hear that, during the CI flow, we should “compile the source code” and run unit tests. However, I am not completely sure what “compile the source code” means in this context. For context, my app should deliver a Docker image with a Python app running on it. We use GitHub Actions for CI/CD. Questions: Should the pipeline simply check out the source code from the repository and compile/build it directly o

AI Agents for DevOps on Kubernetes Need Real Engineering, Not Magic

DZone DevOps (Tier 1) · Publié : 2026-04-30 · 47.8/10

In a real Kubernetes cluster, incidents rarely appear as a single, clean alert. They arrive as waves of Kubernetes events, latency spikes, pod restarts, rollout failures, and unpredictable autoscaling behavior all at once. The hard part is usually not “Can we fix it?” but “Can we understand what’s happening fast enough to make a safe decision?” AI agents for DevOps can help here — but only when they sit on solid engineering foundations. They should compress the early correlation and triage phase

Radar, the “yet another Kubernetes UI” project, now at 1.4k stars after a couple of months

r/devops (Tier 1) · Publié : 2026-05-03 · 46.8/10

A couple of months ago I posted here about Radar, the OSS Kubernetes UI we had just released after getting frustrated with Lens / FreeLens / Headlamp / Kubernetes Dashboard / k9s. That post got a lot more attention than we expected, and the repo is now at ~1.4k GitHub stars ⭐. So first: thanks. A lot of the feedback from that thread shaped what we shipped next. Radar is still fully open source, Apache 2.0. It runs locally as a single Go binary using your existing kubeconfig. No account or cloud

Copy Fail - Une IA trouve la faille Linux que personne n'a vue

Korben (Tier 1) · Publié : 2026-04-30 · 46.8/10

732 octets, c'est tout ce qu'il faut pour passer de simple utilisateur à root sur n'importe quel Linux non patché compilé depuis 2017, soit la quasi-totalité des kernels. Cette faille béante s'appelle Copy Fail (CVE-2026-31431), elle a été dénichée par Taeyang Lee de chez Theori avec leur outil d'audit IA Xint Code. Et comme elle vient d'être divulguée hier sur la liste oss-security et qu'en plus, ils ont fait un joli petit site qui explique tout comme ça fonctionne, je vais essayer de tout vous

Where can I find DevOps tutors at an affordable rate?

r/devops (Tier 1) · Publié : 2026-05-02 · 43.9/10

Hi all, I’ve been working as a SCADA engineer for 2 years now in the Oil and Gas industry and I’m not sure I want to do it long term. Before that I work in an internship at an IT company where I got to learn a bit about DevOps CI/CD Pipelines (Jenkins), APIs (Python, Flask), Micro service architectures, configuration management (Chef Cookbooks), containers (Docker), BDD testing (Cucumber, Pytest), Ruby, etc. I did fine, however, I was laid off after 2 years and had some challenges finding other

AI coding tools are now a CVSS 10.0 CI/CD supply chain vector - patch Gemini CLI and update Cursor

r/devops (Tier 1) · Publié : 2026-05-03 · 43.0/10

Two critical AI coding tool vulns landed last week with the same root cause: agents that autonomously execute OS operations trust their environment in ways that weren't designed for automated use. Gemini CLI (CVSS 10.0, no CVE): In headless/CI mode, it automatically trusted workspace folders for config loading - no sandboxing, no explicit consent. Attack vector: submit a PR to any project running Gemini CLI in CI, plant a crafted .gemini/ config, and you get RCE on the CI host before the san

Is Docker still used in industry or is orchestration the way to go?

r/devops (Tier 1) · Publié : 2026-05-03 · 42.6/10

For most of the time I've had a home lab, I've used Docker to set up services in my network. I don't see much online for examples of actual businesses using Docker in a significant capacity though. Does anybody still use Docker in the industry or has everything switched over to Kubernetes? submitted by /u/ferriematthew [link] [comments]

Kubernetes v1.36: Staleness Mitigation and Observability for Controllers

Kubernetes Blog (Tier 1) · Publié : 2026-04-28 · 42.0/10

Staleness in Kubernetes controllers is a problem that affects many controllers, and is something may affect controller behavior in subtle ways. It is usually not until it is too late, when a controller in production has already taken incorrect action, that staleness is found to be an issue due to some underlying assumption made by the controller author. Some issues caused by staleness include controllers taking incorrect actions, controllers not taking action when they should, and controllers ta

Fresh data has us asking, does AI demand Kubernetes?

The New Stack (Tier 1) · Publié : 2026-05-01 · 41.4/10

Kubernetes is becoming the de facto operating system for AI. Two-thirds of organizations running generative AI models use Kubernetes for The post Fresh data has us asking, does AI demand Kubernetes? appeared first on The New Stack .

Announcing the new Partner Premier tier for the Terraform Registry

HashiCorp Blog (Tier 1) · Publié : 2026-04-30 · 41.1/10

HashiCorp is excited to announce the launch of a new Partner Premier status on the Terraform Registry.

Grouping CI test failures by error signature, is this the right approach?

r/devops (Tier 1) · Publié : 2026-05-01 · 41.1/10

I have a hobby project that I would like to get input on. Im a software engineer, much less in devops, but one of the things I work with at my job is GitHub test pipelines. We have multiple long-running pipelines. Some take several hours, and some can take days. They run 1,300+ test cases, and it is common for a run to have 200+ failed tests. Many of the tests are flaky or unreliable. Its a mess, I know, but im sure its not uncommon. That makes debugging painful. A failed run with 200–300 failed

KULA - Le monitoring serveur Linux qui tient dans un seul binaire

Korben (Tier 1) · Publié : 2026-05-01 · 40.9/10

Ouais, je sais, on est le 1er mai, et je suis pas censé bosser mais que voulez-vous on ne se refait pas ^^. Et si j'ai ouvert l'ordi ce matin, c'est pour vous parler de KULA ! KULA est un binaire tout simple qui permet de monitorer très facilement votre serveur Linux en temps réel, sans aucune dépendance. c0m4r , le dev derrière le projet, l'a codé en Go avec une obsession claire : Que ça marche partout sans rien installer à côté !

OpenTelemetry Japanese Community Survey

OpenTelemetry Blog (Tier 1) · Publié : 2026-04-28 · 40.9/10

This report presents findings from the OpenTelemetry Japanese Community Survey, conducted to understand the current landscape of OTel awareness, adoption, and community engagement among developers and engineers in Japan. The survey targeted practitioners across roles such as development, SRE, DevOps, and Platform Engineering, distributed through CNCF community channels and Japanese social platforms like X (formerly Twitter), Qiita , and Zenn. The goal was to develop data-driven strategies that c

Project Yellow Olive - Pokemon Yellow inspired Kubernetes TUI game

r/devops (Tier 1) · Publié : 2026-05-01 · 40.9/10

Hello r/devops, Hope you're all doing well! A while back I posted here about my side project Project Yellow Olive - a retro-styled TUI game inspired by Pokémon Yellow. The initial feedback was trending on the positive side, so I kept building it. A bit about Project Yellow Olive : The game is all about turning the pain of learning K8s into a fun TUI game. We explore regions, battle with Posemons (container-based creatures), use kubectl-like commands as moves, and complete quests that actuall

4 YOE DevOps Engineer — Can someone review my resume? A senior told me I need 3+ pages to get offers but I kept it to 2 . can some give any suggestions on this.

r/devops (Tier 1) · Publié : 2026-05-03 · 39.2/10

One of my senior colleagues suggested I need at least 3 pages in my resume to get offers. However, from everything I've read online, 1–2 pages is the standard — especially in tech. I kept mine to 2 pages and focused on quality over quantity. I have 4 years of experience in DevOps, specializing in Kubernetes, cloud infrastructure, and CI/CD automation. Currently working at Infosys as a Senior Associate Consultant and actively looking for new opportunities. Key highlights: CKA + CKS certified

Kubernetes v1.36: Pod-Level Resource Managers (Alpha)

Kubernetes Blog (Tier 1) · Publié : 2026-05-01 · 39.2/10

Kubernetes v1.36 introduces Pod-Level Resource Managers as an alpha feature, bringing a more flexible and powerful resource management model to performance-sensitive workloads. This enhancement extends the kubelet's Topology, CPU, and Memory Managers to support pod-level resource specifications ( .spec.resources ), evolving them from a strictly per-container allocation model to a pod-centric one. Why do we need pod-level resource managers? When running performance-critical workloads such as mach

Kubernetes v1.36: Tiered Memory Protection with Memory QoS

Kubernetes Blog (Tier 1) · Publié : 2026-04-29 · 38.7/10

On behalf of SIG Node, we are pleased to announce updates to the Memory QoS feature (alpha) in Kubernetes v1.36. Memory QoS uses the cgroup v2 memory controller to give the kernel better guidance on how to treat container memory. It was first introduced in v1.22 and updated in v1.27. In Kubernetes v1.36, we're introducing: opt-in memory reservation, tiered protection by QoS class, observability metrics, and kernel-version warning for memory.high . What's new in v1.36 Opt-in memory reservation wi

Java in a Container: Efficient Development and Deployment With Docker

DZone DevOps (Tier 1) · Publié : 2026-04-28 · 37.3/10

There is a specific kind of frustration reserved for Java developers who have just containerized their application. You spend hours optimizing your Spring Boot microservice, ensuring your logic is sound and that your tests pass. You wrap it in a Docker container, push it to the registry, and deploy. Then the reality sets in. Your image is 800MB, your startup time is 40 seconds, and during load testing, the container is killed silently by the OS. In my recent work, migrating a monolithic Java app

Anthropic Brings AI-Powered Security Scanning to Enterprise Teams With Claude Security

DevOps.com (Tier 1) · Publié : 2026-05-01 · 37.2/10

Anthropic's Claude Security is now in public beta for Enterprise customers, offering AI-powered codebase scanning and patch generation for security teams.

Secure performance testing at scale: Introducing secrets management for Grafana Cloud k6

Grafana Blog (Tier 1) · Publié : 2026-04-28 · 36.6/10

To simulate real user behavior, performance tests often rely on API keys, tokens, or credentials to interact with real systems. But as your testing suite grows, this sensitive data can start to sprawl across scripts, configs, and environments, increasing the risk of exposure and making tests harder to manage and maintain. To address this challenge, we’re rolling out secrets management for Grafana Cloud k6 , the fully managed performance testing platform powered by k6 OSS. Secrets management allo

Kubernetes v1.36: Mutable Pod Resources for Suspended Jobs (beta)

Kubernetes Blog (Tier 1) · Publié : 2026-04-27 · 36.3/10

Kubernetes v1.36 promotes the ability to modify container resource requests and limits in the pod template of a suspended Job to beta. First introduced as alpha in v1.35, this feature allows queue controllers and cluster administrators to adjust CPU, memory, GPU, and extended resource specifications on a Job while it is suspended, before it starts or resumes running. Why mutable pod resources for suspended Jobs? Batch and machine learning workloads often have resource requirements that are not p

Java Backend Development in the Era of Kubernetes and Docker

DZone DevOps (Tier 1) · Publié : 2026-04-28 · 35.3/10

We moved our monolithic Java application to Kubernetes last year. The promise was scalability and resilience. The reality was a series of silent failures during deployments. Users reported dropped connections every time we pushed a new version. Our monitoring showed zero downtime, but the customer experience told a different story. Requests vanished into the void during rolling updates. We spent weeks chasing network ghosts before finding the root cause. The issue was not the network. It was how

Get observability in the terminal, for you and your agents, with the gcx CLI tool

Grafana Blog (Tier 1) · Publié : 2026-04-28 · 34.7/10

The way you write code is changing, which means the way you observe your systems and respond to issues needs to change, too. Engineers today spend much of their day working via command line, as agentic tools like Cursor and Claude Code have become highly effective at handling many day-to-day engineering tasks. This greatly accelerates code generation, but it doesn't solve for the context switching that comes when you have to jump into another tool that's not part of this new, faster workflow. Mo

DevOps/SRE autonomous agent permissions

r/devops (Tier 1) · Publié : 2026-05-03 · 33.6/10

After the story of claude destroying their whole data with a token it found laying around, how much would you trust a DevOps/SRE agent sitting in the cloud tasked with some autonomous tasks like remediating alerts? how much autonomy/permissions would you give such an agent? https://x.com/lifeof_jer/status/2048103471019434248 submitted by /u/Nash0o7 [link] [comments]

Kubernetes v1.36: In-Place Vertical Scaling for Pod-Level Resources Graduates to Beta

Kubernetes Blog (Tier 1) · Publié : 2026-04-30 · 33.5/10

Following the graduation of Pod-Level Resources to Beta in v1.34 and the General Availability (GA) of In-Place Pod Vertical Scaling in v1.35, the Kubernetes community is thrilled to announce that In-Place Pod-Level Resources Vertical Scaling has graduated to Beta in v1.36! This feature is now enabled by default via the InPlacePodLevelResourcesVerticalScaling feature gate. It allows users to update the aggregate Pod resource budget ( .spec.resources ) for a running Pod, often without requiring a

Anthropic’s Claude Security emerges from closed preview to scan your codebases for vulnerabilities

The New Stack (Tier 1) · Publié : 2026-04-30 · 32.3/10

On Thursday, Anthropic took Claude Security, a defensive security tool in Claude Code on the web that scans codebases for The post Anthropic’s Claude Security emerges from closed preview to scan your codebases for vulnerabilities appeared first on The New Stack .

Faster fixes, less context sharing: how Grafana Assistant learns your infrastructure before you even ask

Grafana Blog (Tier 1) · Publié : 2026-04-30 · 32.1/10

When an unexpected alert fires these days, most engineers' first move is to ask their AI assistant for help.You ask why your checkout service is slow and the assistant gets to work, but it can't get any meaningful insights—at least not quickly—without the proper guidance. So, the next thing you know you're sharing deals about your existing data sources, the services you have running, how they connect, which labels and metrics matter, and on and on. Every conversation starts from scratch, and tha

For Transitioning to DevOps - Does this looks a good plan?

r/devops (Tier 1) · Publié : 2026-05-03 · 31.6/10

https://preview.redd.it/1csfwfxqywyg1.png?width=944&format=png&auto=webp&s=f89aef501d4a03c2a43c0c8d0264abbd2d593f2b I am helping an x-colleague transitioning to DevOps, does this looks kind of a good plan? Transitioning from a Telecom Engineer to a Senior DevOps Engineer presents a moderate gap, primarily in cloud and containerization skills. With your strong problem-solving abilities and systems integration experience, you are well-positioned to bridge this gap with targeted learnin

Should I take the AWS SAA certificate?

r/devops (Tier 1) · Publié : 2026-05-01 · 30.3/10

I’ve always been against certification over experience, but recently and with the way the market is, I unfortunately think a certificate gives a proper push. For context, I am a DevOps engineer and have slight hands on experience with AWS, and growing. As the company isn’t huge, I have bigger ownerships and opportunities to dive deep into the usual DevOps and also AWS. I decided to ask here because I noticed a lot of people are taking it, and while it’s not opening doors from 0, it may be benefi

Built a Jenkins plugin that tracks all 4 DORA metrics natively

r/devops (Tier 1) · Publié : 2026-05-02 · 29.9/10

Been working on getting DORA metrics visibility for our pipelines without setting up external infrastructure like Prometheus/Grafana or paying for commercial tools. Ended up building a Jenkins plugin that does it all inside Jenkins itself. It tracks Deployment Frequency, Lead Time for Changes, MTTR, and Change Failure Rate. Also does pipeline rankings (slowest, most failing, flakiest), stage-level analytics, and has a REST API if you want to pull data into other tools. Everything runs on an embe

SRE Weekly Issue #515

SRE Weekly (Tier 1) · Publié : 2026-05-04 · 29.8/10

View on sreweekly.com A message from our sponsor, atscaleconference.com: Building scalable, high-performance infrastructure for AI is one of today’s toughest challenges. Join @Scale: Systems & Reliability on June 25 in Bellevue, WA to learn how leading engineers are solving it. Secure your seat today! The Silent Failure of Reliability Metrics at Scale: Lessons Learned from […]

Kubernetes for platform teams: Leveraging k0s and k0rdent

CNCF Blog (Tier 1) · Publié : 2026-04-27 · 28.4/10

In our previous blog, we explored a GitOps use case for on-premises infrastructure, managing multiple clusters hosted on the k3s Kubernetes distribution using k0rdent.  But the platform engineering ecosystem is vast, and one blog barely scratches...

Kloak : injection de secrets en kernel-space via eBPF sur Kubernetes

Une tasse de café (Tier 1) · Publié : 2026-05-02 · 26.5/10

Comment Kloak intercepte le trafic TLS de vos pods au niveau kernel avec des uprobes eBPF pour injecter vos secrets de façon transparente, sans modifier vos applications ni déployer de sidecar.

Arm Adds Free Toolkit to Analyze AI Agent Performance

DevOps.com (Tier 1) · Publié : 2026-04-30 · 26.3/10

Arm this week made available a free toolkit for analyzing agentic artificial intelligence (AI) workloads as they are being developed by DevOps and platform engineering teams. Earlier this year, Arm unveiled a 3nm processor based on its Neoverse V3 architecture that is specifically designed for AI workloads. The Arm Performix toolkit provides system-wide analysis across […]

Généré par veille-auto · Modèle : gemma4:e2b