Capybara Consulting

AI · Security · Infrastructure

Calm, senior engineering for systems where AI meets production.

Everyone in this space sells fear. We'd rather be the calm in the room: fixed-scope audits, private LLM builds, and infrastructure that stays up — delivered by people who operate what they build, and handed over completely when we're done.

Tell us what hurts →

Fixed-Scope Services

AI Security & Readiness Audit

from $12,500· 2 weeks

For teams whose AI feature just hit an enterprise security questionnaire. OWASP LLM Top 10 and NIST AI RMF mapped assessment, delivered as evidence your buyer's CISO will accept.

  • Threat model of your LLM stack
  • Prompt-injection & data-exfil review
  • Remediation backlog, prioritized

Agent & MCP Security Review

Inquire· scoped per engagement

Your AI agents touch dozens of systems with credentials. We map every tool, permission, and exfiltration path before someone else does.

  • Agent permission & access audit
  • MCP server and tool-surface review
  • Runtime guardrail recommendations

Private LLM: Fine-Tune + Deploy

from $8,000· 3–4 weeks

A domain-specialized model fine-tuned on your literature and workflows, running on hardware you control. No per-token fees, no data leaving your walls.

  • Data curation & LoRA/DPO fine-tune
  • Quantized serving (vLLM / GGUF)
  • Docs, pipeline, and handover — yours

Infrastructure Audit + Monitoring

$3,500· 1 week

Kubernetes and storage health audit with Prometheus/Grafana observability deployed, alert rules configured, and a written report your team can act on.

  • Cluster + storage assessment
  • Monitoring stack deploy
  • Hardening & security review

Ongoing? Advisory retainer from $2,500/mo — re-testing on model updates, questionnaire response support, incident triage.

For Universities & Research Groups

A private research assistant for your group

We build LLMs fine-tuned on your field's literature and your group's workflows — retrosynthesis planning, spectroscopy interpretation, method development, grant-writing support. Students query it like any other tool.

  • Runs on hardware you control — or we provision a GPU server for you
  • Unpublished results and IP-sensitive research never leave the building
  • No API dependency, no per-query cost

We start with a free proof-of-concept

Before anyone commits to anything, we fine-tune a demonstration model on your subfield's literature and bring it to a 20-minute demo. If it doesn't clearly beat a general-purpose chatbot on your workflows, there's nothing to discuss.

Collaboration framing welcome — joint research evaluations, student-involved builds, and HPC integrations are all on the table.

For Startups

Get through enterprise security review

You shipped the AI feature; now Fortune-500 procurement is asking for prompt-injection defenses and model supply-chain evidence. The AI Security Audit above exists exactly for this — two weeks from "questionnaire panic" to a report your sales team can forward.

Fractional senior AI/infrastructure engineering

Production LLM pipelines, Kubernetes and distributed storage, inference serving, zero-trust networking — 10–20 hrs/week of senior-level help without a senior-level headcount. Available through the A.Team network or directly.

Production Work Behind This

How It Goes

Reach out

Tell us what hurts. A paragraph is plenty.

Scope it

30-minute call, fixed scope, fixed price. No discovery-phase billing.

We deliver

Work happens on your infra or ours. You see progress, not a black box.

You own it

Configs, code, models, docs — everything, handed over. No hostage-taking.

Contact

Email: hello@capybaraconsulting.agency
XMPP: capybara-consulting@5222.de — OMEMO welcome.