Files
benvin c05ccfcb5d
ci/woodpecker/pr/build Pipeline was successful
ci/woodpecker/pr/pre-commit Pipeline was successful
ci/woodpecker/pr/test Pipeline was successful
Initial implementation: NATS->S3 archiver + search/retrieve CLI
logarchiver replaces the plain Vector archiver leg of the centralized
logging stack (argocd-apps #296) with a Go service that archives raw logs
from NATS JetStream to S3 as zstd-compressed, OpenPGP-encrypted, indexed
objects, plus an operator CLI to search the index and retrieve/decrypt
archived logs. It adds the things that outgrew Vector: zstd compression,
encryption keyed from Ben's Vault GPG secrets engine, a searchable
ClickHouse index, and sink-conditional acks (a batch is acknowledged to
JetStream only after the object is durably in S3 AND indexed).

Service (`logarchiver run`):
- Durable JetStream pull consumer (stream LOGS, durable archiver, subject
  filter default logs.k8s.vault.>), explicit acks, independent offsets.
- Batch per subject by size/count/time -> NDJSON -> zstd -> encrypt -> S3
  PUT -> ClickHouse index row -> ack. On any failure the batch is Nak'd and
  redelivered, so nothing is lost on a sink outage.
- Encryption is a wrapped-DEK envelope (container LARC1): the bulk is
  AES-256-GCM framed under a random data key, and only that 32-byte key is
  OpenPGP-encrypted to the engine's public key. This is because the Vault
  GPG engine does whole-payload decrypt only; retrieval round-trips just the
  tiny wrapped key regardless of object size. Public key fetched from the
  engine or a mounted file (configurable); key fingerprint recorded per
  object; periodic pubkey refresh for rotation.
- Prometheus metrics, structured slog, graceful drain on shutdown.

CLI:
- `search` queries the index (subject/host/time) and lists matching objects.
- `fetch` downloads, decrypts via the Vault GPG engine, unzstds and emits
  NDJSON (optionally re-filtered by host/time).
- `init-schema` creates/prints the ClickHouse archive_index DDL.
- cobra `completion` subcommands.

Config via file+env (k8s-friendly, secrets from env), boundaries (NATS/S3/
ClickHouse/Vault) behind interfaces with unit tests (config, batching,
host/subject extraction, crypto roundtrip with a test key, ack-after-persist
with fakes, search query building). go build/vet/test -race clean;
golangci-lint v2 clean. Woodpecker CI: build/test/pre-commit on PR; on v*
tag a container image plus a Gitea binary release + rpm-internal RPM. Docs
per subcommand + architecture + retrieval runbook + deployment drop-in.

Claude-Session: https://claude.ai/code/session_015ur3i7D2azsMAWTSVABApv
2026-07-27 23:22:40 +10:00

36 lines
1.9 KiB
SQL

-- logarchiver archive index schema.
--
-- One row is written per stored S3 object. In-cluster this DDL is owned by the
-- argocd bootstrap Job (a ClickHouse PostSync hook, like the logging stack's
-- clickhouse-schema job); `logarchiver init-schema` applies the same statements
-- for local/dev use, and `logarchiver init-schema --print` emits them.
--
-- Keep this file in sync with internal/index/ddl.go (the source of truth used by
-- init-schema). The database/table names below match the config defaults
-- (database `logs`, table `archive_index`).
CREATE DATABASE IF NOT EXISTS logs;
CREATE TABLE IF NOT EXISTS logs.archive_index
(
object_key String, -- S3 key of the stored object
bucket LowCardinality(String), -- S3 bucket (e.g. logs-archive)
subject LowCardinality(String), -- NATS subject the batch came from
hosts Array(LowCardinality(String)),-- distinct source hosts in the object
min_ts DateTime64(3), -- earliest event time in the object
max_ts DateTime64(3), -- latest event time in the object
event_count UInt64, -- number of events
raw_bytes UInt64, -- pre-compression NDJSON bytes
stored_bytes UInt64, -- stored object bytes (post zstd+encrypt)
compression LowCardinality(String), -- 'zstd'
cipher LowCardinality(String), -- 'AES-256-GCM'
container_format LowCardinality(String), -- 'LARC1'
key_name LowCardinality(String), -- Vault GPG engine key name
key_fingerprint String, -- 40-hex OpenPGP fingerprint (engine %X)
created_at DateTime64(3) DEFAULT now64(3),
INDEX idx_hosts hosts TYPE bloom_filter GRANULARITY 1
)
ENGINE = MergeTree
PARTITION BY toYYYYMM(min_ts)
ORDER BY (subject, min_ts, object_key);