# Have a custom AI agent built for your team

> An AI agent is software that handles a recurring job inside your systems on its own: it reads what comes in, works through tools like your CRM, ERP or ticketing system, produces a result you can check, and hands over to a person the moment its remit ends. twigbit builds agents like that as software in your own stack, with evals, approvals and one metric, and keeps them running: model migrations, regressions and maintenance included.

URL: https://www.twigbit.ai/en/ki-agenten
Published: 2026-09-10
Author: Moritz Morgenroth (Co-Founder & Managing Director)
Area served: Germany, Austria, Switzerland

---

## Results

- **3.9×** Return on the support-triage agent at Projektron in its first year.
- **−58 %** Median time to first response on support tickets at Projektron.
- **27 %** Cost-saving potential the agent found for A2Mac1 across the supply chains analysed.
- **€2,900** Fixed price for the discovery sprint, one week, credited in full against the deployment.

## Five steps to an agent in production

1. **Pick a workflow and measure it** One inbox, one intake channel, one job that comes up many times a day. We record today's volume, error rate and handling time, so that “better” ends up being a number.
2. **Discovery sprint, one week, €2,900** We cut the role: the job, the inputs, the expected result, the exceptions, the access it needs. Plus the evaluation criteria the deployment will be measured against, and a go/no-go recommendation with a costed roadmap.
3. **Build and integrate** The agent is built as software in your stack, connected to your CRM, ERP, ticketing system, Microsoft 365, Slack or Langdock. Least-privilege access, documented data flows, and evals against historical cases before the first live run.
4. **Go live with approvals** Every consequential action is approved by a person first. Autonomy comes later, per action type, against measured error rates. Responsibility stays where it belongs, and the agent earns its room to move.
5. **Operations and maintenance** Monitoring with alerts, evals on every change to prompt, model or upstream system, model migrations, cost and token budgets, and a monthly review against the agreed metric. Scope grows when the number holds.

## What the price covers after go-live

- **Monitoring and alerts** Every run leaves behind its input, context sources, tool calls, model version, result and approval. Deviations report themselves before your customers notice them.
- **Evals as a regression test** A few hundred historical cases with known-good results. The suite runs on every change to prompt, model or upstream system, and no change ships until it passes.
- **Model migrations** Providers deprecate model versions and change prices. We migrate, measure the candidate against the eval suite, and keep a rollback ready before the new version goes live.
- **Cost and token control** Budgets per run and per month, model choice matched to the job, a cost report in the review. An agent that quietly gets more expensive is a maintenance case, not a law of nature.
- **Access, data and approvals stay with you** Identities, permissions, data and approvals live in your house. You are the deployer under the EU AI Act. We answer for the code, the evals, the migrations and the regressions, like any software vendor under a maintenance contract.
- **Monthly review** One metric, the month's error classes, the open proposals. Then you decide whether the agent runs more actions autonomously, takes on more cases, or stays exactly as it is.

## What an AI agent is, and what separates it from a chatbot

A chatbot answers a question, and then the case is over. An AI agent does a job. It often starts without any chat, from an incoming email or a new ticket, reads the case, pulls context out of your systems, runs several steps and leaves a result where your work actually happens: a draft order in the ERP, a classified ticket in the right queue, a prepared reply in the inbox.

Technically an agent has four parts. The language model supplies language understanding and judgement. The tools give it access to your systems, a CRM, a ticketing system, a price list. The memory holds what has already happened in the case at hand. The rules set out what it may do without asking. In production this runs as a loop: goal, next step, tool, check the result, continue or hand over. That is where it differs from classic automation. The goal is fixed in advance; the path to it is not.

Most of the "we need an agent" conversations we have end in a workflow with a single, tightly bounded agentic step. For most business processes that is the better design. Autonomy is a capability and a risk at once, and it is earned per action type, against measured error rates. The distinction in detail is in the glossary under [AI agent](/glossary/ai-agent) and [agentic workflow](/glossary/agentic-workflow).

## Which workflows suit a first agent

A good candidate has three properties: it happens often, it has a recognisable intake, and it produces a result someone can check. The value rarely appears where a whole role is meant to be replaced. It appears in the cases that come up many times a day and touch the same two or three systems every time.

- **Support triage.** Read every incoming request, classify it, draft a first answer from the knowledge base, route it to the right queue, escalate to a person when confidence is low. At Projektron the agent classifies 71% of tickets automatically and has cut median time to first response by 58%.
- **Order entry.** Orders arrive as PDFs, free-text emails or scans, every customer in their own format. The agent reads the order, matches line items against the article master, flags discrepancies and creates a draft order. A person confirms it.
- **Supplier and customer email in procurement.** Order confirmations, delay notices, price changes and complaints get classified, enriched with context from ERP and CRM, and passed on as a draft reply or an internal escalation.
- **Quote preparation.** Read the enquiry, pull master data and the price list, prepare a draft quote with its open questions for sales to check and send.
- **CRM upkeep and follow-ups.** Create contacts, log interactions, research new leads, chase what is due. At twigbit that is [Chameleon](/agents/revenue-operations).
- **Content research and drafts.** Research topics, gather sources, prepare drafts in the house voice for a person to approve. At twigbit that is [Woodpecker](/agents/marketing).
- **Reporting and anomalies.** Watch agreed metrics, investigate material changes, present findings with sources and open questions. At twigbit that is [Seagull](/agents/data-analyst).

The warning signs are unclear goals, irreversible decisions without approval, and cases too rare to fill an eval suite. An agent that decides about people, ranking applicants or assessing creditworthiness, is a high-risk system under the EU AI Act and not a first agent.

We bring four agents out of our own operation: [Woodpecker](/agents/marketing) for content, [Chameleon](/agents/revenue-operations) for the CRM, [Seagull](/agents/data-analyst) for reporting and [Lynx](/agents/seo-accessibility) for SEO and accessibility. They run here daily, and their roles, evals and permission models are the starting point when your workflow resembles one of them.

## Integrating into IT that has grown over years

An agent is software in your stack, not another platform with its own login. We build it onto the systems you have: CRMs like HubSpot, Salesforce or Pipedrive, your ERP, ticketing systems like Zendesk, Jira or Freshdesk, Microsoft 365 or Google Workspace, Slack or Teams. Where you already run Langdock, ChatGPT Enterprise or Copilot, we deliver the agent into those, through MCP servers and skills, so your people stay in the interface they know. How that works technically is under [Model Context Protocol](/glossary/model-context-protocol).

Three principles hold in every project. First, least privilege: the agent gets exactly the access its job requires, and identity and permissions live in your IAM, not with us. Second, documented data flows before the first live run: which source, which model, which region, which retention. Third, no lock-in: code, prompts, evals and configuration live in your repository. We work like any software vendor under a maintenance contract, and the contract settles the exit before the agent goes live.

For hosting there are three routes: model providers with an EU region and a processing agreement, European providers, or [models on your own infrastructure](/glossary/on-premise-ai) where your data and the job call for it. The choice is made during the sprint, together with your IT and your data protection officer, and it is written into the documentation afterwards.

## Data protection, GDPR and the EU AI Act

An agent that reads email and queries your CRM processes personal data all day. The questions that follow are well known, and we answer them in the design rather than afterwards. The legal basis under Art. 6 GDPR is documented, usually through a legitimate-interest assessment. Decisions with legal or similarly significant effect keep effective human review, as Art. 22 requires, and "effective" means something: whoever nods through 400 results a day is not reviewing. The model provider is engaged as a processor under Art. 28, with a contract, a sub-processor list and a clear answer to where inference runs. Where customer or employee data is processed systematically, a data protection impact assessment under Art. 35 comes before go-live. Retrieval stays scoped to the case at hand, and every run is logged in full. Those same logs satisfy the GDPR's accountability duty, the AI Act's obligations and your own debugging needs.

The [EU AI Act](/glossary/eu-ai-act) applies in stages. The AI-literacy duty under Art. 4 has applied since February 2025, the transparency duties under Art. 50 since August 2026, and the transition periods for high-risk systems run into 2027 and beyond. A triage or order-entry agent with human review is not a high-risk system. High risk is AI that decides about people. If an agent drifts from sorting the applications inbox to ranking candidates, it has changed risk class, and the role is cut precisely to prevent that.

You are the deployer under the regulation. That means identity, permissions, data and approvals stay with you, and the people supervising the agent understand what it does. We answer for the code, the evals, the model migrations and the regressions. That split is written into the contract.

## What operations and maintenance actually mean

Go-live is where the work starts. Model providers deprecate versions. An upstream system changes a field. A colleague edits a template and the agent suddenly produces something different. None of that is a bug in the agent, and all of it is a maintenance case.

That is why every agent we build comes with an eval suite of a few hundred historical cases with known-good results, which runs on every change to prompt, model or upstream system. Plus monitoring with alerts, cost and token budgets, a rollback to the last working version, and a monthly review against the agreed metric. What that means in detail, and how responsibility splits between deployer and vendor, is in the two guides [operating AI agents](/blog/ki-agenten-betreiben) and [AI agent maintenance](/blog/ki-agenten-wartung).

## What it costs

Our prices are on the page. The first call takes 30 minutes and costs nothing. The discovery sprint is €2,900 at a fixed price and takes a week; it ends in the agent role, the evaluation criteria and a go/no-go recommendation with a costed roadmap. Deployments start at €7,900 setup, operations and maintenance at €2,400 a month. The sprint is credited in full against a deployment within 90 days.

What drives the price: the number of integrations, the state of your data, how much control you need — that is, how many action types require approval — and the volume. What does not drive it: the size of your company. The proposal after the sprint makes scope, setup, running costs, metric and exit explicit.

## Why agent projects fail

It is almost never the model. MIT's Project NANDA examined 300 enterprise deployments and found that around 95% of GenAI pilots show no measurable effect on the P&L. RAND, after interviews with 65 experienced engineers, estimates that more than 80% of AI projects fail, twice as often as IT projects without AI. The reasons repeat. The agent lacks clean access to the systems where the work actually happens. Nobody defined what a good result is. The handover to a person is missing, and the first edge case leaves the job stuck. After go-live, nobody looks after it.

Our approach is cut to those four causes: the sprint supplies the definition of success, the integration supplies the access, the approval mode supplies the handover, the maintenance contract supplies the operation. What the engineering discipline behind it looks like is in [shipping AI that survives production](/blog/shipping-ai-that-survives-production).

## Why an engineering team

twigbit is a team of senior software engineers in Berlin that runs itself on agents. Sales, marketing, reporting and parts of development go through the same agents we ship. What we learn there about evals, context upkeep and permission models goes into the next project.

What we sell is software that moves a number, and the responsibility for it in production. At Projektron the support-triage agent returned 3.9× its cost in the first year; at A2Mac1 the agent found 27% cost-saving potential across the supply chains analysed. Both cases are below, with names and contacts. Ask them.

## How to start

Bring a recurring workflow, or a product idea. In 30 minutes we work out whether it is clear, frequent and controllable enough for an agent, what integration or engineering could make of it, and what the smallest sensible first step would be: sprint, prepare, or deliberately wait.

<Sources
  title="Sources"
  items={[
    {
      label: "The GenAI Divide: State of AI in Business 2025",
      href: "https://fortune.com/2025/08/18/mit-report-95-percent-generative-ai-pilots-at-companies-failing-cfo/",
      source: "MIT Project NANDA, via Fortune",
    },
    {
      label: "The Root Causes of Failure for Artificial Intelligence Projects and How They Can Succeed",
      href: "https://www.rand.org/pubs/research_reports/RRA2680-1.html",
      source: "RAND Corporation",
    },
    {
      label: "Regulation (EU) 2024/1689 (AI Act)",
      href: "https://eur-lex.europa.eu/legal-content/EN/TXT/?uri=CELEX:32024R1689",
      source: "EUR-Lex",
    },
    {
      label: "Guidance: Artificial intelligence and data protection",
      href: "https://www.datenschutzkonferenz-online.de/media/oh/20240506_DSK_Orientierungshilfe_KI_und_Datenschutz.pdf",
      source: "Datenschutzkonferenz (DSK)",
    },
    {
      label: "Building effective agents",
      href: "https://www.anthropic.com/engineering/building-effective-agents",
      source: "Anthropic",
    },
  ]}
/>

## What companies ask before they have an AI agent built

### What does it cost to have an AI agent built?

The first call is free. The discovery sprint is €2,900 at a fixed price: one week, ending in the agent role, the evaluation criteria and a go/no-go recommendation with a costed roadmap. Deployments start at €7,900 setup, operations and maintenance at €2,400 a month. The sprint is credited in full against a deployment within 90 days. What drives the price is integrations, the state of your data and how much control you need — not the size of your company.

### How long until the first AI agent is in production?

The sprint takes a week. A first deployment with an approval loop usually needs four to eight weeks after that; the range depends on the number of integrations and the state of your data. The proposal after the sprint names the duration for your case. Autonomous actions follow step by step, once the error rates allow it.

### Who runs and maintains the agent?

We do, under a maintenance contract from €2,400 a month: monitoring, evals on every change, model migrations, cost control and a monthly review. Identity, permissions, data and approvals stay with you. If you want to take operations in-house later, the contract provides for it: code, prompts, evals and documentation live in your repository.

### How do you handle data protection and the GDPR?

Before the first live run we document data flows, access, model providers, sub-processors and retention. Model providers are engaged as processors with an EU region; where your data calls for it, models run at European providers or on your own infrastructure. Consequential decisions keep effective human review. A blanket compliance promise does not replace that assessment, and your data protection officer is in the room from week one.

### What does the EU AI Act require from us as the deployer?

A triage or order-entry agent with human review is not a high-risk system. The general duties still apply to you as deployer: AI literacy for the people supervising the agent (Art. 4, in force since February 2025), transparency towards affected people (Art. 50, since August 2026), and classifying every use case before it is built. We supply the classification, the logs and the documentation you need for that.

### How do you integrate AI agents into existing IT?

As software connected to the systems you already have: CRM, ERP, ticketing, Microsoft 365 or Google Workspace, Slack or Teams. Where you use Langdock, ChatGPT Enterprise or Copilot, we deliver the agent into those, through MCP servers and skills. The agent gets exactly the access its job requires, and your IT keeps identity and permissions. No further platform with its own login joins the stack.

### We already use ChatGPT Enterprise, Copilot or Langdock. Do we still need an agent?

A chat seat waits for prompts; a person has to start every run. An agent has a bounded job, starts itself when a case comes in, and delivers the result into your system. The two fit together: your people stay in the interface they know, and the agent takes over the recurring part of the work inside it.

### How do you measure whether the agent works?

Before we build, we agree on a metric — time to first response, share of cases handled automatically, or cost per case. On top of that comes an eval suite of historical cases with known-good results, which runs on every change. At Projektron that meant 3.9× return in the first year and 58% less time to first response. Without a number there is no deployment.

### Do you work across Germany and Europe?

Yes. We have one office, in Berlin, and work with companies in Germany, Austria and Switzerland, mostly remotely and with on-site sessions where they speed up the sprint or the rollout. Contracts under German law, documentation and reviews in German or English, hosting in the EU. Details per region are on the Germany, DACH and Berlin pages.

### Which workflows suit a first agent?

Recurring work with a recognisable intake, a checkable result and enough volume: support triage, order entry from PDFs and free text, supplier and customer email in procurement, quote preparation, CRM upkeep, reporting. The warning signs are unclear goals, irreversible decisions without approval, and cases that come up less than a few times a day.

### What is the difference between an AI agent and a chatbot?

A chatbot is an interface: a person asks, the system answers, and the case is over. An agent is a workflow. It often starts without any chat, for instance from an incoming email, runs several steps inside your systems and leaves a result behind. An agent can have a chat window; it does not need one.

### Can we take the agent over ourselves later?

Yes. Code, prompts, evals, configuration and documentation live in your repository from day one, and identity and permissions live in your IAM. The maintenance contract settles handover and exit before the agent goes live. Switching is then a handover meeting, not a migration project.

## Our working area

- [Have an AI agent built for your company in Germany](https://www.twigbit.ai/en/ki-agenten/deutschland): twigbit builds, integrates and maintains AI agents for companies across Germany. One office in Berlin, work remote and on site, contracts under German law, hosting in the EU. Discovery sprint €2,900, deployment from €7,900, maintenance from €2,400 a month.
- [Have an AI agent built in Berlin](https://www.twigbit.ai/en/ki-agenten/berlin): twigbit is based in Berlin-Kreuzberg and builds, integrates and maintains AI agents for companies and startups in the city. Sprint sessions on site or at our office on Yorckstraße, references Projektron and Thermondo. Discovery sprint €2,900, deployment from €7,900.
- [Have an AI agent built in Germany, Austria or Switzerland](https://www.twigbit.ai/en/ki-agenten/dach): twigbit builds, integrates and maintains AI agents for companies in Germany, Austria and Switzerland. One office in Berlin, work remote, on-site sessions as needed. GDPR in Germany and Austria, the revised FADP in Switzerland. Sprint €2,900.
