STAFF / PRINCIPAL / FORWARD DEPLOYED AI SYSTEMS

I build, break, and deploy AI systems.

Hands-on engineer across model tuning, inference infrastructure, offensive security, and regulated production. I lead from the code, the architecture, and the customer room.

Model work that ships

I transform model behavior, quantize for real hardware, validate serving targets, and publish the artifacts with explicit evaluation limits.

Gemma 4 31B uncensored · format family

Problem

Remove refusal behavior from a 30.7B multimodal model without losing a deployable format set.

Evidence: BF16
Technical decision

Applied weight-level abliteration, then carried the same model lineage through BF16, ModelOpt NVFP4, and a GGUF K-quant ladder. For multimodal NVFP4 serving, the vision tower, embedder, and lm_head stayed in BF16.

Evidence: BF16NVFP4GGUF
Result

BF16 evaluation: 0/686 effective refusals across four datasets. The 20 GB NVFP4 build, reduced from 59 GB, was validated in vLLM and SGLang. GGUF: 12–31 GB with 0–1/100 hard refusals across quant levels.

Evidence: BF16NVFP4GGUF

Ornith 1.0 35B · NVFP4

Problem

Fit a 34.7B Qwen3.5-MoE reasoning checkpoint onto a single Blackwell workstation without quantizing every subsystem.

Evidence: Ornith-1.0-35B-NVFP4
Technical decision

Quantized routed experts to NVFP4 with NVIDIA ModelOpt; kept attention QKV, shared experts, and the vision encoder at higher precision.

Evidence: Ornith-1.0-35B-NVFP4
Result

About 23 GB versus 69 GB BF16, validated to load and generate with vLLM on an RTX PRO 6000 Blackwell.

Evidence: Ornith-1.0-35B-NVFP4

MiniMax-M3-uncensored · 428B MoE

Problem

Study refusal behavior in a 428B multimodal MoE while preserving its reasoning and routing structure.

Evidence: MiniMax-M3-uncensored
Technical decision

Modified attention o_proj and every residual-writing down_proj across all 60 layers; released the result as 796 GB BF16 in 59 shards.

Evidence: MiniMax-M3-uncensored
Result

0/16 hard refusals on the published sample; multimodal behavior, reasoning, and MoE routing were smoke-tested, not fully benchmarked.

Evidence: MiniMax-M3-uncensored
View all Hugging Face releases (opens in a new tab)

Attack the agent. Harden the system.

I connect model-behavior research with the controls needed around agents: red teaming, prompt injection, MCP attack paths, containment, policy enforcement, and auditability.

Model behavior

Abliteration, alignment analysis, refusal measurement, and adversarial evaluation across released checkpoints.

Agent attack surface

LLM red teaming, prompt injection, MCP scanning, tool poisoning, exfiltration paths, and tool-chain abuse.

Containment & audit

Runtime policy enforcement, egress controls, human approval, signed audit trails, and compliance evidence.

Production systems, end to end

The model is only one component. I build the Linux, GPU, Kubernetes, inference, networking, observability, and security path around it.

  1. WEIGHTSPublished checkpoints and model behavior
  2. TUNELoRA · behavior work
  3. QUANTIZENVFP4 · GGUF
  4. SERVEvLLM · SGLang
  5. ATTACKMCP · agents
  6. HARDENPolicy · audit

AI inference platform

I build and operate a GitOps-managed Talos/Kubernetes inference platform across NVIDIA and AMD GPUs. An OpenAI-compatible router coordinates vLLM model swaps, backend readiness, and streaming; request, token, latency, concurrency, and swap telemetry make operation measurable.

Private platform · architecture available in interview

Audit-grade delivery

Technical delivery and evidence in FINMA, DORA, and TLPT environments.

FINMA · DORA · TLPT

Lead from the difficult technical work

I lead by reducing ambiguity, making the hard technical decisions, and staying with the problem until the system ships.

Selected work

Open-source AI-security tools and systems I build, publish, and operate.

Rust · React · PostgreSQL · AI workflows

paperless-archivist

Security-first AI automation for Paperless-ngx: resumable OCR and tagging workflows, validation-gated autopilot, human review, cited document chat, RBAC, and auditable actions.

Product code, shipped quietly

Uhr Lernen Einfach

Learn to read clocks. 7 game modes, 8 themes, Game Center.

iOS · Swift / SwiftUI

ABC Lernen

Learn the alphabet. 8 game modes — writing, matching, syllables.

iOS · Swift / SwiftUI

Pixel Privacy

Blur faces & license plates in photos. 100% offline, on-device.

iOS · Swift / SwiftUI

CraftFace

Turn your selfie into a Minecraft character. 3D preview, 7 outfits.

iOS · Swift / SwiftUI

Twenty years of increasing technical scope

From Linux and high-availability systems to cloud, offensive security, model work, and global customer delivery.

Global offensive-security delivery

100+ engagements across 11 countries while remaining hands-on.

Cloud architecture to security practice

Practice building, APT delivery, and a 40% delivery improvement.

Infrastructure to automation

600 nodes migrated, 35+ engineers trained, and 80% less manual work.

Linux engineering to solution architecture

99.99% HA systems and ERP RPO reduced from one day to 15 minutes.

Show full career history
  1. Associate Director, Offensive Security Global @ Kyndryl

    Embedded with C-Level at Tier-1 banks, insurers, automotive across 11 countries. 100+ on-site engagements. Custom tooling shipped under engagement deadlines. DORA / FINMA / TLPT.

  2. Senior Red Team Lead Engineer, Global @ Kyndryl

    Forward-deployed APT simulation for Fortune 500. Custom tooling and rapid prototyping in-engagement — 40% efficiency gain in delivery.

  3. Associate Partner Security @ Kyndryl

    Built consulting practice and GTM strategy after IBM spin-off. Pre-sales scoping and customer-embedded delivery model.

  4. Security Consultant & Cloud Architect @ IBM

    Cloud transformations — AWS, GCP, OpenShift, K8s. Founded consulting security division.

  5. DevOps Engineer Expert @ Avectris AG

    Founded Linux department. Backend automation in Go (80% workload reduction). Scrum Master.

  6. DevOps Engineer @ Swisscom AG

    Automation for Swisscom TV (300+ channels). Voice control system. 35+ engineers trained.

  7. Linux Architect @ EveryWare AG

    Enterprise infrastructure for Allianz, JobCloud, UBS AG.

  8. Linux DevOps Engineer @ CTBTO (United Nations)

    Migrated 600 nodes from CFEngine to Puppet. Near 100% config automation.

  9. Lead Solution Architect @ Schrack Seconet AG

    Rebuilt global IT (HQ + 16 branches). 100% security audit pass. ERP RPO: 1 day → 15 min.

  10. Linux Engineer @ Freelance

    HA systems in Interxion datacenters. 99.99% uptime.

Credentials

39 credentials across offensive security, Linux, automation, containers, and cloud.

Show all 39

Security & Offensive

CompTIA

Red Hat

Linux Foundation & LPI

LFCSLinux Foundation & LPI (opens in a new tab)LPI-3 (304)Linux Foundation & LPILPI-3 (303)Linux Foundation & LPI

Cloud & ITSM

Chef / DevOps

Extending ChefChef / DevOpsCertified Chef DeveloperChef / DevOpsDeploying CookbooksChef / DevOpsLocal Cookbook DevChef / DevOpsBasic Chef FluencyChef / DevOps

Interactive terminal

Explore the Evidence CLI, then optionally run a fictional local delivery mission. Commands and file content are never transmitted; writable files use origin-private storage only after you enable it.

Terminal output

Commands run only in this browser. Command and file content is never sent to a server.

Commands run only in this browser. Command and file content is never sent to a server.

Writable files are saved locally only after you enable them.

Try: help, pwd, ls, about, models, security, network, storage status

Interactive features are unavailable. All evidence and contact paths remain available; only the terminal and display settings are inactive.

Build the hard system with me.

I am interested in hands-on Staff, Principal, and Forward Deployed roles at AI companies, plus selective technical collaboration.

Display settings
Matrix background