Agentic Technologies GmbH

Age of
Intelligence.

Cognition is becoming infrastructure. We build the companies, products, and research that run on it.

Agents running
ZRH --:--:--
47.37°N 8.54°E
Brief 044 Kolibri: open German weights you can run

We build with agents.
For you, for us, in public.

A small Zurich team with a large agent workforce. Three lines of work, one operating system underneath.

01

Build

Agent systems for your company.

We find where your team loses hours, then design, deploy, and run the agent layer: research, content, monitoring, operations. Wired into your data and your controls.

  1. Map
  2. Pilot
  3. Run
02

Products

Tools we run on ourselves first.

Everything we build for clients runs in our own operation before it runs in theirs. What survives daily use becomes a product.

  1. Use
  2. Harden
  3. Ship
03

Research

Briefs for operators and capital.

Mechanisms, numbers, second-order effects. What changed in AI and what it does to cost, speed, and headcount. No hype feed.

  1. Read
  2. Test
  3. Publish
Read the briefs →

Thesis

Capacity used to scale with headcount.
Now it scales with design.

For a century, what a company could do was bounded by whom it could hire and coordinate. Language models cut that link. Research, drafting, monitoring, and analysis now cost compute, not calendars.

What stays scarce is architecture. Which loops run on their own, what context they share, where each decision lands and who owns it. That is the work we do, for clients and on ourselves first.

Compute

Capacity you rent, not hire

Context

Memory that compounds

Design

The part that doesn't commoditise

An operating layer,
not a chat window.

Memory, execution, feedback, and orchestration, designed as one system. Six properties everything we build is held to.

Institutional Memory

Context compounds.

Decisions, threads, and data feed a living model. Less re-briefing. More continuity across people and agents.

Execution Layer

Work moves without waiting.

Background jobs, monitors, drafts, follow-ups. The queue clears overnight, not at the next stand-up.

Thin Overhead

Signal over ceremony.

Agents coordinate at machine speed. Humans stay in the loop through tight reviews - not meetings for their own sake.

Feedback Loops

Rated work gets better.

Approve, edit, reject - then feed the signal back. The stack improves from real use, not slide decks.

Model Freedom

Best brain for the job.

Frontier APIs or open weights. Swap the model, keep memory and process. No single-vendor religion.

Multi-Agent

Specialists, one desk.

Research, ops, code, comms - separate agents with shared state. Orchestrated, not a single chat blob.

RESEARCH // BUILD // REVIEW // SHIP // COMPOUND // JUDGE // RESEARCH // BUILD // REVIEW // SHIP // COMPOUND // JUDGE // RESEARCH // BUILD // REVIEW // SHIP // COMPOUND // JUDGE // RESEARCH // BUILD // REVIEW // SHIP // COMPOUND // JUDGE // RESEARCH // BUILD // REVIEW // SHIP // COMPOUND // JUDGE // RESEARCH // BUILD // REVIEW // SHIP // COMPOUND // JUDGE // RESEARCH // BUILD // REVIEW // SHIP // COMPOUND // JUDGE // RESEARCH // BUILD // REVIEW // SHIP // COMPOUND // JUDGE // RESEARCH // BUILD // REVIEW // SHIP // COMPOUND // JUDGE // RESEARCH // BUILD // REVIEW // SHIP // COMPOUND // JUDGE // RESEARCH // BUILD // REVIEW // SHIP // COMPOUND // JUDGE // RESEARCH // BUILD // REVIEW // SHIP // COMPOUND // JUDGE //

What we're watching.

Substance briefs - mechanisms, numbers, second-order effects. Built to be useful to operators and capital, not a hype feed.

All 44 briefs →

What a token per second feels like.

Four speeds, named after things you already use. Faster rows write more, so you can actually see them.

5–10tok/s · 225–450 wpm
Like reading
Small model on a laptop CPU

40–50tok/s · 1800–2250 wpm
Local on a Mac
8B class · still readable

80–100tok/s · 3600–4500 wpm
About ChatGPT
Typical ChatGPT / Claude reply, roughly

120–150tok/s · 5400–6750 wpm
About Gemini Flash
Flash-tier APIs · a wall of text

Names are approximate - typical single-stream replies, not a lab bench. 1 token ≈ 4 characters. Not batched server throughput.

Signal

Get the briefs first.

Substance over hype. New insights land in your inbox as we publish them - mechanisms, numbers, second-order effects.

No spam. Unsubscribe anytime.