Skip to content

Security and compliance

AI system security

AI system security addresses the risks specific to systems built on language models and agents: prompt injection, data exfiltration through model outputs, over-privileged tools, and model access control. Hexploits reviews and hardens AI systems using the OWASP guidance for LLM applications and the governance patterns it built for swarmd.ai.

A named engineer replies within one working day. A written scope and an indicative price within two.

  • Businesses about to connect an AI assistant or agent to internal systems.
  • AI vendors answering enterprise security questionnaires.
  • Security teams asked to approve a system they have not seen before.

Deliverables, not slogans. Each one appears in the statement of work.

  • A threat model for the AI system: inputs, tools, data sources and trust boundaries.
  • Controls: input and output filtering, scoped tool credentials, rate limits, tenant isolation, logging.
  • Testing for prompt injection and data leakage, with results.
  • Model access control and secrets management.
  • Documentation for security review and customer questionnaires.

Our engineers work across the major languages, frameworks and cloud platforms. We build on the stack you already run, with technology choices explained in writing before work begins.

The same four stages as every Hexploits engagement, applied to this capability.

  1. Stage 1

    Model the threats

    What the system can reach, who can influence its inputs, and what an attacker would gain.

  2. Stage 2

    Harden

    Least-privilege tools, filtering, isolation and logging built in.

  3. Stage 3

    Test

    Adversarial testing against the deployed system, repeated on change.

  4. Stage 4

    Monitor

    Anomalous prompts, tool calls and outputs alert an engineer.

Every engagement agrees its measures and the measurement period in writing before work starts.

  • Adversarial test results before and after hardening.
  • Tools and data sources reachable by the system, minimised and documented.
  • Security questionnaires answered from evidence.

Case studies with numbers, and reviews linked to Google where they were left there.

  • Director, IO Solutions

    Fantastic to work with. High level of attention to detail and flawless communication throughout. Would recommend to anyone looking to develop or improve a software product.

    Christian LorzaDirector, IO SolutionsRead the review
  • Director, Lothbury

    Top quality delivery, and reasonable price. Will be using again.

    Peter DentonDirector, LothburyRead the review
What is prompt injection?
An attack where content the model reads, such as a document or a web page, contains instructions the model follows. The defence is to limit what the model can do, filter what it reads and writes, and log everything.
Can an agent leak our data?
It can if it is over-privileged and unmonitored. Scoped credentials, output filtering and an immutable log make leakage both unlikely and visible.
Do you follow a standard?
The OWASP Top 10 for LLM applications, NCSC guidance on secure AI development, and ISO 42001 for governance.

Request a proposal.

Tell us about the system and the sector. A named engineer replies within one working day. A written scope and an indicative price within two working days of a short scoping call.