Open source · Self-hosted · Spec-driven with Specstride
// 01 · the loop
intent ─▶ ┌────────┐ CRs ┌─────────────┐ reconcile ┌────────┐
│ agents │ ────▶ │ controllers │ ────────▶ │ fabric │
└───▲────┘ └─────────────┘ └───┬────┘
│ gNMI telemetry │
└──────────────────────────────────────────┘
- agentic-netops: AGNTCY and LangGraph agents turn intent into Kubernetes CRs, controllers reconcile them onto a live SONiC EVPN/VXLAN fabric, and gNMI telemetry closes the loop. Runs in containerlab and kind. Built with Specstride.
- agentic-netops-srl: the same approach on Nokia SR Linux: a declarative EVPN/VXLAN fabric control plane with a guarded multi-agent intent tier. Built with Specstride.
- agentic-ops-bench: model-by-harness benchmark on real ops tasks (AIOps, NetDevOps, HPC, inference, RAG, tool loops). Local 30B-class models on one RTX 3090 against hosted frontier models.
// 02 · harnesses
- specstride: spec-driven autonomous coding orchestrator. Drives an agent through Spec Kit phases behind an LLM critic gate, diagnoses stuck phases, and tunes its own per-phase settings.
- mixture-of-loops: generates provenance-bound, unattended Specstride pipelines from Spec Kit artifacts, and ships specstride-batch to run that pipeline over batches of features at once. Built with Specstride.
- agent-observability-stack: self-hosted observability for LLM agents and the Linux host running them. Prometheus, Grafana, Loki, Tempo and OTel, with Arize Phoenix as a second trace backend for LLM calls. Metrics and traces for OpenClaw, LiteLLM, Claude Code and RAG.
- qmd-memory-stack and qmd-gateway: local GPU-accelerated three-tier RAG memory for coding agents, and a shared writable MCP memory gateway for a multi-agent fleet.
- claude-plugins: Claude Code plugin marketplace, including antares-scan, first-pass security triage of a code folder with a local Granite-4.0 1B model.
// 03 · labs
Containerlab labs for Nokia SR OS and SR Linux, built at Nokia and published under srl-labs.
- srl-sros-telemetry-lab: interactive streaming telemetry lab with SR Linux and SR OS, gNMI into Prometheus and Grafana.
- sros-anysec-lab: quantum-safe ANYsec encryption demo on SR OS FP5 vSIMs.
- sros-anysec-macsec-lab: quantum-safe ANYsec and MACsec together in one lab.
// 04 · the fleet
Skills, curated per session through skillOverrides profiles — these are the ones that stay on; the host-ops ones live in agent-skills:
- gpu-ops: operates and inspects the RTX 3090 eGPU — status, free VRAM, drain, safe power cycles — through proven scripts, never ad-hoc commands.
- speckit-batch: runs a Spec Kit command over a range of specs on a chosen harness and model, several features in parallel.
- Specstride batch: runs the full Specstride / Mixture of Loops pipeline over a range of specs — one launch contract per feature, derived and executed unattended on Claude Code or dsh, model and reasoning pinned per run, a shared queue so two harnesses drain batches side by side.
- mixture-of-loops: derives the provenance-bound, unattended Specstride pipelines that Specstride batch launches.
- specstride-curate: curates and reconciles a Spec Kit spec corpus for a target host, unattended end to end — auto-research agents verify the host facts and distill aesthetic references into cited briefs, a deterministic script backs up and cross-checks, curation agents apply the retargeting rules, analyze reports refresh.
- jupyter-pull: pulls every file off a remote Jupyter server I am logged into, given only the pasted session cookie.
- spec-reconcile: writes a dated spec-vs-deployed sheet, read-only and in a fixed vocabulary — the drift audit between what a spec says and what is running.
- qmd-recall: recall from and write to the shared fleet memory, so every agent on the host inherits prior context and durable gotchas.
- fleet-control: brings the LLM and observability stack up or down and returns a fleet health digest.
- proxmox-ops and proxmox-triage: guarded lifecycle operations (snapshot before risky changes) and read-only diagnosis of Proxmox VMs and containers.
- local-model-ops: which local model to pick on the 3090, when to think longer, and when to escalate to a hosted frontier model.
Agents:
- Linky: drafts a LinkedIn post in my voice every morning and publishes only after I approve it (the link is the public demo).
- Local AI analyst: a weekly buy-or-wait read on local inference hardware against hosted frontier models, checked against written buy triggers and a ledger.
- Observability digest: a daily Telegram summary of the agent fleet and its host.
- Portfolio sync: keeps mairp.ai and its digital twin's knowledge base in step with my public repos, daily.
- Relay: a NetOps and infrastructure agent I reach over Telegram; it works cards from the kanban board.
- Eterna: agentic RAG over InfiniBand, RDMA, RoCE and HPC material.
- Havan: hunts flight deals and returns a ranked shortlist (the link is the public demo).
- Release watcher: polls upstream release tags every 15 minutes and fast-forwards my local clones, never merging or running repo code.
- video-to-deck: turns videos into Marp slide decks using local Whisper, OCR and a fresh agent per video.
- kanban: persistent kanban board with a Claude and Telegram front end.
- Digital twin: answers questions about my work on mairp.ai.
// 05 · tooling
- kind-cilium-hubble-cluster: observable Kubernetes cluster with kind, Cilium and Hubble from scratch.
- gpu_rtx_3090: safe power-cycle scripts for an RTX 3090 Thunderbolt eGPU.
- certforge: small OpenSSL CLI for CSRs and self-signed certificates from a .cnf.





