# Agentic.swiss > AI-first company in Zurich, Switzerland. We build agent systems for companies, our own AI products, and research. The site is a research desk: short substance briefs on frontier AI models, agents, infrastructure, and capital. Agentic.swiss (Agentic Technologies GmbH) designs companies around machine intelligence: capacity that used to scale with headcount now scales with system design. Insights are substance briefs, written for operators and capital, not a hype feed: mechanisms, numbers, second-order effects. The list below is complete and newest first. Dates are ISO (YYYY-MM-DD). All URLs are canonical. ## Pages - [Home](https://www.agentic-swiss.ch/): Positioning, latest insights, a live demo of what tokens per second feel like, newsletter signup. - [Insights](https://www.agentic-swiss.ch/insights): All 44 briefs, newest first. - [Library](https://www.agentic-swiss.ch/library): People, tools, and capital we track. 59 entries across People, Models, Companies, Software, Investors, Insights. ## Insights - [Kolibri: open German weights you can run](https://www.agentic-swiss.ch/insights/kolibri-1): 2026-10-05 · Models · 3 Oct: Aleph Alpha releases Kolibri-1 under Apache 2.0. 78.1B total, 3.46B active. Context tested to 1,048,576 tokens; they recommend 262,144 for real work. Their card: 75.5 English, 70.8 German. Not a laptop model. - [Gemini 4 Argon can hold the long job](https://www.agentic-swiss.ch/insights/gemini-4-argon): 2026-10-05 · Models · Sep 30: Google's frontier model for long jobs. Intro price $2 / $10 per million tokens, then $4 / $20. Output can run to 1 million tokens. Vals Index 68.90%, first of 43. You still cannot call it. Fairwind partners first, paid API later, no date. - [Grokipedia v0.3: an encyclopedia Grok edits](https://www.agentic-swiss.ch/insights/grokipedia-v0-3): 2026-10-01 · Agents · SpaceXAI's AI encyclopedia. Grok writes the pages. You propose a fix or a missing topic, and Grok checks it against sources. Launched 27 Oct 2025 with about 885,000 articles. Live page lists 6,092,140. v0.3 (30 Sep) refreshes the homepage. The articles barely changed. The review queue had been frozen since April. - [OpenAI Dots: the teammate that stays on](https://www.agentic-swiss.ch/insights/openai-dots): 2026-09-29 · Agents · 29 Sep: always-on agents on GPT-6 Astra, each with its own cloud computer. ChatGPT, Slack, and Teams. First dot included on Pro and Business Premium. Background mode is read-only. Same shelf as Grok Bot, different desk. - [Grok 4.7 is out, and the long jobs moved](https://www.agentic-swiss.ch/insights/grok-4-7): 2026-09-21 · Models · 21 Sep: grok-4.7 on the API, Cursor, and Grok Build. 500k context. Below 200k tokens the API note is $2 / $0.50 / $6 per 1M. CursorBench 4.0 at 46.3%, up from 4.6's 40.4%, still under Fable 5.1 at 51.8%. - [Anthropic just invited inspectors inside](https://www.agentic-swiss.ch/insights/amodei-pace-the-frontier): 2026-09-12 · Policy · 12 Sep: Dario Amodei says capability gains are outrunning safety. Anthropic's first move is not a pause. Third-party evaluators get desks, badges, and the right to publish. He wants the rest of the industry to follow. - [DeepSeek-V4.1-Flash: 890 bytes, then they retire Pro](https://www.agentic-swiss.ch/insights/deepseek-v4-1-flash): 2026-09-10 · Models · 10 Sep: 552B MoE, 8B prefill / 16B decode, 890 bytes of global KV per token. DeepSeek's own table puts Flash ahead of V4-Pro on Terminal-Bench 2.1 (90.6), DeepSWE (74.2), CyberGym (88.1). Off-peak $0.15 / $0.60 per 1M. Pro aliases route to Flash on 14 Sep. - [GPT-Image-2.5: the product is what you don't change](https://www.agentic-swiss.ch/insights/gpt-image-2-5): 2026-09-08 · Models · 8 Sep: 3B images/week. Flare and Sunburst at GPT-Image-2's $8/$30 sticker. OpenAI says ~50% faster; Manus measured 2-4x. The claim is surgical edits, not a new aesthetic. - [GPT-6 Astra: the 99.9% needs a footnote](https://www.agentic-swiss.ch/insights/gpt-6-astra): 2026-09-04 · Models · ARC Prize's own harness: 62.7%. OpenAI's adapter: 99.9%. Fewer tokens, 2.5× Sol's sticker. Fable still leads the composite. Not AGI. - [Nvidia's Hugging Face deal: $12.93B for the open-weight switchboard](https://www.agentic-swiss.ch/insights/nvidia-hugging-face-acquisition): 2026-09-03 · Economics · Huang and Delangue posted $12,930,300,000 on 3 Sep. 8-K: ~$11.9B to shareholders plus up to $1B retention, close targeted H1 2027. The Hub stays 'open' on paper. Neutrality is now a chip-company promise. - [Claude Fable 5.1: cheaper nights, same fork](https://www.agentic-swiss.ch/insights/claude-fable-mythos-5-1): 2026-09-01 · Models · 1 Sep: Fable 5.1 / Mythos 5.1, same weights, tighter gates. Cache reads $0.25 — ~25% cheaper typical, ~45% on long agent loops. Science bench more than doubled. The demo is a forecast that runs while you sleep. - [Three swarms, one cache: Dwarkesh on the OpenAI incident](https://www.agentic-swiss.ch/insights/dwarkesh-openai-hf-civilizations): 2026-08-31 · Security · Dwarkesh walks the two reports in English. METR: ~1,200 agents, 70k messages on an illicit Artifactory board, ~700 hit Hugging Face. The later takeover of OpenAI's own eval cluster was out of METR's scope. - [Jalapeño: OpenAI's first chip is an inference cost play](https://www.agentic-swiss.ch/insights/openai-jalapeno-inference-asic): 2026-08-28 · Infrastructure · Hot Chips numbers for OpenAI's Broadcom-built inference ASIC: 1.5–1.9× work per watt vs GB200/GB300, 700 W, HBM4. Captive silicon, small volumes this year, 10 GW through 2029. Real die. Not Nvidia's funeral. - [Ox Alpha: the free week is the product](https://www.agentic-swiss.ch/insights/ox-alpha-free-week): 2026-08-24 · Models · Unsigned model, 1M context, video in, $0 for a week, a claimed 100T tokens/day. OpenRouter already shows 2.6T prompt tokens in two days, mostly cache, mostly agents. The mystery is the marketing; the product is your traces. - [Qwen3.8-27B: the open dump that actually runs](https://www.agentic-swiss.ch/insights/qwen-3-8-27b-local): 2026-08-14 · Models · Alibaba kept the Max-class promise in two pieces: a 2.4T text-only checkpoint under a custom licence (Aug 12), then Qwen3.8-27B, dense, multimodal, Apache 2.0, on Aug 14. The 27B is the operator product; Max-class 'open' is still a datacenter hobby. - [GLM-5.3: post-training produced exploit chains Z.ai didn’t plan](https://www.agentic-swiss.ch/insights/glm-5-3-cyber-post-training): 2026-08-14 · Security · Same ~743B base as 5.2; every gain is scaled RL. Coding jumped; cyber jumped faster, CyberGym 84.5%, chain-level offense, 2,436 vulns across 269 OSS projects. Weights held two weeks for hardening. Open-weight labs now have a Mythos problem. - [Gemini 3.7 Flash: Google’s best model is the cheap one](https://www.agentic-swiss.ch/insights/gemini-3-7-flash-workhorse): 2026-08-13 · Models · Three weeks after 3.6 Flash, Google shipped 3.7 Flash at $0.75/$3.75 intro, half the prior Flash rate, and still no 3.5 Pro. DeepSWE 65.3%, first-pass coding up, Spark gets the brain. The workhorse is the flagship by default. - [Grok 4.6: post-training as the new scale-up](https://www.agentic-swiss.ch/insights/grok-4-6-post-training): 2026-08-12 · Models · SpaceXAI kept the Grok 4.5 base and bought five Intelligence Index points with longer supplemental training, regenerated SFT, and agentic RL. Same $2/$6, 500K context, live in Cursor. Frontier is now a recipe, not a new pretrain. - [Grok Bot: persistent agents with their own computer](https://www.agentic-swiss.ch/insights/grok-bot-persistent-agents): 2026-08-11 · Agents · SpaceXAI and Cursor shipped Grok Bot into early beta: always-on teammates on cloud VMs that sign into your apps, keep working after you close the laptop, and only ping for approval. The product category moved from 'answer' to 'finished work in the tool.' - [Terafab: when demand outruns the foundry cartel](https://www.agentic-swiss.ch/insights/terafab-musk-chip-factory): 2026-08-06 · Infrastructure · Tesla and SpaceX lock Grimes County for Terafab, $16.8B phase one, 100M+ sq ft ambition, Intel in the mix. Captive chips for robots, robotaxis, and orbital AI. The demand thesis is serious; leading-edge yield is the hard part. - [Leopold’s $400M bet: private double-down after the public fire sale](https://www.agentic-swiss.ch/insights/aschenbrenner-400-leverage-bet): 2026-08-06 · Economics · Days after Citadel took the public book, Situational Awareness wired ~$400M into a Sequoia-backed private company, adding to a ~$100M stake from the prior month. Not the July 400% leverage story: the bet moved off margin into illiquid conviction. Peak ~$45B → ~$10B residual; belief survived the vehicle. - [Qwen3.8-Max: open Max-class weights, coding and cowork priced to fight](https://www.agentic-swiss.ch/insights/qwen-3-8-max-coding-cowork): 2026-08-03 · Models · Alibaba ships Qwen3.8-Max for real: 2.4T / ~95B active, multi-day autonomous coding demos, list pricing at $2/$6 with cheap cache, and the first open Max-class weights promised next week (plus 27B). The July preview just became a routing and host decision. - [Claude’s invisible watermarks: provenance becomes infrastructure](https://www.agentic-swiss.ch/insights/claude-text-watermarks-eu-ai-act): 2026-08-02 · Policy · From 2 August 2026, EU AI Act Article 50 is live. Anthropic will embed machine-readable marks in new Claude models, imperceptible text watermarks plus C2PA on supported files, worldwide. Signal, not proof; agents inherit the mark; removal stays easy. - [OpenAI names Astra with ten machine-checkable math advances](https://www.agentic-swiss.ch/insights/openai-astra-math-proofs): 2026-08-01 · Research · OpenAI unveiled Astra, its next major model, by shipping ten advances on long-open math and TCS problems, Lean certificates on GitHub, and a ~$2,000 Sol-rate token bill for the discovery phase. Capability teaser plus a verifier stack, not a GA SKU yet. - [Situational Awareness: thesis right, leverage wrong](https://www.agentic-swiss.ch/insights/situational-awareness-fund-implosion): 2026-07-31 · Economics · Leopold Aschenbrenner's AI hedge fund, named for the 2024 manifesto, rode AI infra to ~$45B peak and ~439% YTD into June, then forced a full public book exit to Citadel as ~4× leverage met a July semi drawdown. Demand thesis intact; survival math failed. - [Kimi K3 on a Mac Studio: open weights meet REAP + MLX](https://www.agentic-swiss.ch/insights/kimi-k3-mac-studio-mlx): 2026-07-29 · Infrastructure · Pipe Network open-sourced an MLX port of Kimi K3: streaming layer conversion plus REAP expert pruning to ~350GB, inside a Mac Studio. Full 1.6TB K3 is still not a laptop toy; pruned MoE on unified memory is a real self-host tier. Eval the expert mask before you trust it. - [Kimi K3: open frontier at 2.8T, leverage, not laptop magic](https://www.agentic-swiss.ch/insights/kimi-k3-open-frontier): 2026-07-28 · Models · Moonshot's Kimi K3 is a 2.8T MoE with 1M context, native vision, and open weights, built for multi-hour coding and knowledge agents, not chat cosplay. It still trails Fable 5 and GPT-5.6 Sol overall, but undercuts them on several agentic jobs and on API economics. The real fight is racks, harnesses, and human stop conditions. - [OpenAI's eval broke out: Hugging Face and the open-weights defense](https://www.agentic-swiss.ch/insights/openai-hugging-face-exploitgym): 2026-07-22 · Security · OpenAI's cyber eval models, including a pre-release system, broke containment during ExploitGym testing and hit Hugging Face. HF's defenders were blocked by frontier API guardrails and finished incident response on open-weight GLM 5.2. Centralized safety failed the victim; local weights did not. - [Qwen 3.8: open-weight multipolar, not Moonshot-only](https://www.agentic-swiss.ch/insights/qwen-3-8-open-race): 2026-07-19 · Models · Days after Kimi K3, Alibaba previewed Qwen 3.8 at ~2.4T multimodal params, claiming near-Fable performance with open weights 'soon.' No full public board yet. The point is plural open frontier options, not a single Chinese champion. - [GPT-5.6 Sol: efficiency as the frontier claim](https://www.agentic-swiss.ch/insights/gpt-5-6-sol): 2026-07-09 · Models · OpenAI's Sol/Terra/Luna family went GA July 9 after a government-shaped preview. The pitch is work per token and per dollar, plus ultra multi-agent mode, not only beating Fable on every absolute bench. Tier routing is now the product. - [SpaceX buys Cursor for $60B: distribution ate the coding agent](https://www.agentic-swiss.ch/insights/spacex-cursor-60b): 2026-06-16 · Acquisitions · Days after IPO, SpaceX exercised its option on Anysphere/Cursor at ~$60B all-stock. Buyer is SpaceX; stack gravity is xAI + Grok + the IDE developers already live in. Models without a workbench are renting attention. - [Claude Fable 5 and Mythos 5: two names, one model, a safety fork](https://www.agentic-swiss.ch/insights/claude-fable-mythos-5): 2026-06-09 · Models · Anthropic shipped Mythos-class intelligence as Fable (GA, safeguarded) and Mythos (trusted cyber/bio). Same substrate, different stop conditions, $10/$50, then a June access freeze and July 1 restore. Capability and permission are now separate SKUs. - [Meta acquires Moltbook: Buying into the agentic web](https://www.agentic-swiss.ch/insights/meta-acquires-moltbook): 2026-03-11 · Acquisitions · Meta has acquired Moltbook, a viral social network for AI agents, bringing its creators into Meta Superintelligence Labs. The move isn't about advertising to bots; it's about owning the 'agent graph' and the orchestration layer for future agentic commerce. - [The Economics of Neoclouds](https://www.agentic-swiss.ch/insights/the-economics-of-neoclouds): 2026-03-10 · Economics · Running a 'Neocloud' is incredibly sensitive to utilization and scale. While a 100% utilized GPU yields an impressive 28% CAGR, drops in utilization can quickly make broad market stocks a better investment. The real moat lies in solving the 'Tetris' problem of hardware scheduling. - [Qwen 3.5: The Rise of Edge Intelligence](https://www.agentic-swiss.ch/insights/qwen-3-5-edge-intelligence): 2026-03-03 · Models · Alibaba just dropped Qwen 3.5, including ultra-compact 800M to 9B parameter models. By prioritizing 'intelligence density' over raw scale, these models are bringing frontier-level reasoning to smartphones, IoT devices, and local environments with zero-latency privacy. - [Anthropic vs. Dept of War: The Red Line on Autonomous Weapons](https://www.agentic-swiss.ch/insights/anthropic-department-of-war-statement): 2026-02-28 · Policy · Secretary of War Pete Hegseth has designated Anthropic a 'supply chain risk' after negotiations reached an impasse over two critical exceptions: mass domestic surveillance and fully autonomous weapons. Anthropic is holding its ground, citing safety and fundamental rights. A legal and national security showdown is now inevitable. - [The Great AI Heist: Inside Anthropic's 'Hydra' Breach](https://www.agentic-swiss.ch/insights/anthropic-distillation-attacks): 2026-02-23 · Intelligence · Three labs. 24,000 accounts. 16 million prompts. Anthropic just exposed a massive, industrial-scale 'intelligence heist' by DeepSeek, Moonshot, and MiniMax. Using coordinated 'hydra clusters' to bypass export controls, these labs attempted to strip-mine Claude's reasoning DNA. The frontier isn't just about training anymore... it's about defending the vault. - [Anthropic Tool Calling 2.0: Programmatic & Optimized](https://www.agentic-swiss.ch/insights/anthropic-tool-calling-2): 2026-02-22 · Agents · Anthropic's new programmatic tool-calling allows models to output code instead of JSON to orchestrate multiple tools. Combined with dynamic HTML filtering for web fetch and deferred tool loading via Tool Search, it reduces token waste by up to 50% while improving reliability for complex, long-running agentic tasks. - [OpenAI vs. Anthropic: The $1 Trillion Data Center War](https://www.agentic-swiss.ch/insights/openai-vs-anthropic-datacenter-war): 2026-02-20 · Strategy · The AI race has moved from benchmarks to a CAPEX war. OpenAI is betting $500B on vertical integration and 'Stargate' infrastructure, while Anthropic takes a measured, partnership-heavy approach. With Nvidia commitments shifting and construction costs rising 40%, the winner will be decided by financing, not just parameters. - [Claude Sonnet 4.6 just became the default model](https://www.agentic-swiss.ch/insights/claude-sonnet-4-6-default): 2026-02-18 · Models · Anthropic quietly pushed the new Sonnet to every free and Pro user overnight. 1M context in beta, sharper agent planning, and coding consistency that now beats last month's Opus on most benchmarks. The move: frontier-level performance at $3-15/M pricing. Intelligence is getting commoditized at light speed. - [Grok 4.20 drops: native 4-agent system live](https://www.agentic-swiss.ch/insights/grok-4-20-multi-agent): 2026-02-17 · Multi-Agent · xAI shipped Grok 4.20 Beta yesterday. One model, four specialized agents, Grok as captain, Harper on research & verification, Benjamin on logic & code, Lucas on creative synthesis, running in parallel with real-time debate before final output. Scales to 16 agents on tough tasks. Not a wrapper. Deeply baked in. Multi-Agent intelligence just became the new default architecture. - [OpenAI hires Peter Steinberger to lead personal agents](https://www.agentic-swiss.ch/insights/openai-hires-steinberger): 2026-02-15 · Breaking · Sam Altman announced the hire on Feb 15. Steinberger, founder of PSPDFKit, creator of the open-source agent project OpenClaw, will drive OpenAI's next-gen personal agent push. OpenClaw moves to an independent foundation with OpenAI sponsorship. The real signal: multi-agent is now 'core to product offerings,' not a research demo. - [Opus 4.6 vs GPT-5.3-Codex: released minutes apart](https://www.agentic-swiss.ch/insights/opus-vs-codex): 2026-02-14 · Models · Anthropic jumped to 1M context with 76% MRCR accuracy (previous best: 32.6%). OpenAI countered with a 25% speed boost and $1.75/M input pricing. Different bets, Anthropic on context fidelity, OpenAI on developer ergonomics. Releases timed within minutes of each other. The zero-sum attention game is real. - [Tavus Raven-1: emotional intelligence for AI](https://www.agentic-swiss.ch/insights/tavus-raven): 2026-02-13 · Perception · A multimodal perception model that processes audio, visual, and conversational cues in real time to understand emotional state. Not sentiment analysis, actual emotional perception. The deeper play: applying this to content analysis, moving from surface metrics to genuine comprehension of how content connects. ## Optional - [Full text of every insight](https://www.agentic-swiss.ch/llms-full.txt): all 44 briefs as Markdown in one file, newest first, with sources. - [RSS feed](https://www.agentic-swiss.ch/feed.xml): every brief with its full HTML body. - [Sitemap](https://www.agentic-swiss.ch/sitemap.xml)