[ 0.000000]Linux version 6.18.33 (kernelpanic@lima) #1 SMP PREEMPT_DYNAMIC
[ 0.000000]Command line: BOOT_IMAGE=/vmlinuz root=/dev/aws-prod ro quiet
[ 0.812345][ OK ] Started KernelPanic on-call engineer.

We keep
production
alive.

Your on-call SRE for AWS: one senior engineer who knows the environment before anything breaks, reviews it, diagnoses the whole system, and answers for it. 100% AWS infrastructure.

[ 1.200000][ OK ] 100% AWS · Production infrastructure
[ 1.313000][ OK ] LatAm + US · Experience in both markets
[ 1.426000][ OK ] On call · One engineer who knows the environment
[ 1.539000][ OK ] No surprises · A quote before the work · a price that does not change
[ 1.711111][FAILED] Failed to start Keeping production alive: nobody's job. See 'journalctl -u production' for details.
[ 1.900000][ OK ] Reached target 01 / The gap

Keeping production alive is nobody's job. Until it goes down.

Engineers ship features; that is what they are paid for. The AWS partner that sends the invoice opens a ticket and answers in office hours. AWS Support supports AWS, not the company's systems.

[ more ][ less ]

When it breaks, the fix depends on the few people who hold the whole system in their heads. Nothing gets written down, so the same outage comes back.

If a partner already invoices AWS, they stay: the invoice is their job; reliability is KernelPanic's.

[ 2.000000][ OK ] Reached target 02 / How it works

Three steps. One engineer.

01

Start on the retainer

An engineer on call who knows production before anything breaks: access, an alert channel, a runbook. Written reviews every week and every month. When something is wrong, you hear it from us first.

02

Fix what the reviews find

Each fix is quoted before the work and the price does not change: hardening, architecture changes, load testing to the root of the latency, migrations onto AWS, the infrastructure for a new system.

03

Call when it breaks

Incident response by the engineer who already knows the system, closed with a written root cause.

A quote before the work. A price that does not change. First step: write to info@kernelpanic.pe. One conversation, then the retainer terms.

[ 2.050000][ OK ] Reached target 03 / The retainer

An on-call engineer who already knows the environment.

The retainer is attention on the environment: one engineer, who becomes the company's on-call production engineer.

[ more ][ less ]

It buys four things: the environment known before anything breaks; an on-call engineer who already knows the system; written reviews every week and every month, with an alert when something is wrong; and fast quotes, because the context already exists.

The AWS account stays the company's: root is never handed over, access is read-only wherever possible, and findings are delivered privately.

Talk about the retainer

[ 2.104000][ OK ] Reached target 04 / The work

What comes out of the reviews, and what happens when it breaks.

[ 2.200000][ OK ] Reached target Diagnose — 01/03
[ 2.421000]

[ OK ] Started Architecture review

The whole environment read against AWS's Well-Architected pillars: what holds, what does not, and why.

[ 2.738000]

[ OK ] Started Load testing

The system proven against the next campaign, before the peak proves it. One scenario, one target, a written read of the results.

[ 2.900000][ OK ] Reached target Fix — 02/03
[ 3.055000]

[ OK ] Started Hardening & remediation

The weak points left by reviews and incidents, closed. The findings from a review get one quote, and each one is closed with evidence.

[ 3.372000]

[ OK ] Started Migration & new infrastructure

Moving a system onto AWS or building the infrastructure for a new one, for companies on the retainer. Infrastructure only.

[ 3.500000][ OK ] Reached target Respond — 03/03
[ 3.689000]

[ OK ] Started Incident response

When production fails, the engineer who already knows the environment resolves it and closes with the written root-cause analysis.

A quote before the work. A price that does not change. The price does not grow with AWS consumption: a smaller bill and fewer incidents suit KernelPanic.

[ 3.900000][ OK ] Reached target 05 / Proof

We read the whole system, not one layer.

The freeze that started three layers away

The ticketing system froze every peak day, and everyone looked at the ticketing system. The load balancer logs led to dashboard code calling an internal API on every login, and behind it, an API that could not scale. Symptom in one system, cause in another, evidence in a third. Code rewritten, API scaled, season saved.

SYMPTOM → LOGS → CODE → CAUSE

When adding servers made nothing faster

A legacy application autoscaled onto a shared FSx cache and latency never moved, while AWS's dashboards called the storage healthy. The bottleneck was in metadata operations, visible only when CloudWatch agent metrics were read against a load-test baseline. The cache moved to NVMe instance volumes with stickiness, and the system held 170,000 requests per minute, 10,000 concurrent users in a 15-minute window.

170,000 REQ/MIN · 10,000 CONCURRENT USERS
[ more ][ less ]

Production on AWS for a US tax-filing platform, where the year's traffic arrives in one season. One season it kept failing on infrastructure; the next, zero infrastructure incidents. The only failures were application code that arrived broken, each one named with its cause in writing. Before that, two AWS consulting partners, with companies in Chile, Colombia, Argentina, Peru, and the US. AWS Certified Solutions Architect. The same engineer on every engagement.

[ 4.200000][ OK ] Reached target 06 / FAQ

Asked before the conversation.

Who keeps production on AWS alive?

KernelPanic: one senior engineer who knows the company's production infrastructure, reviews it, and answers for it. Write to info@kernelpanic.pe.

How do we start?

On the retainer: an engineer on call, written reviews, and an alert channel. The fixes come out of the reviews, each one quoted before it starts. One conversation by mail, then the terms.

How does the retainer work?

One engineer knows the environment, is on call, and reviews production in writing every week and every month. When work is needed, it is quoted before it starts and the price does not change.

We already have a partner that invoices our AWS. Now what?

Keep them. That contract covers the invoice; production reliability usually sits outside it. KernelPanic works alongside the billing partner, not against it.

How much does it cost?

Every engagement is quoted before it starts and the price does not change. The retainer terms are discussed directly.

Do you do migrations or new infrastructure?

Yes, for companies on the retainer: moving a system onto AWS or building the infrastructure for a new one. Infrastructure only; the application stays with the company's team.

Doesn't AWS Support cover this?

No. The AWS Support FAQ says it does not include software debugging, system administration, or remote access to customer systems. AWS supports AWS. KernelPanic answers for the company's infrastructure.

Do you touch application code?

Diagnosis reads the whole system. The scope is AWS infrastructure: when the cause lives in code or in a data platform, the findings name it and the company's team fixes it.

Do you serve companies outside Peru?

Peru, Latin America, and the US, in English and Spanish. Lima is on UTC-5, the US East Coast clock.