---
schema: "swft.publication/v1"
id: "software-factories-across-industries"
title: "Software factories and agent swarms across industries"
description: "Compare software factories and agent systems across banking, trading, automotive, chemicals and research, including Jane Street, Palantir, BNY, BASF and BMW."
summary: "Agent systems are appearing in banking, trading, automotive software, industrial planning and scientific research. Their public accounts describe different kinds of use: recurring development, deployed business workflows, pilots and research experiments. Compare the work and evidence before comparing scale."
canonical: "https://swft.io/industries"
author: "SWFT Editorial"
author_type: "Organization"
published: "2026-09-08"
modified: "2026-09-08"
kind: "reference"
section: "References"
tags: ["software factories", "agent swarms", "enterprise AI adoption", "AI across industries", "Jane Street", "Palantir"]
evidence_labels: ["INFERENCE", "SELF-REPORT"]
source_ids: ["industry-map-basf", "industry-map-bmw-review", "industry-map-bny-annual", "industry-map-bny-q1", "industry-map-gemini-science", "industry-map-mercedes", "industry-map-nubank", "industry-map-palantir-ai-fde", "industry-map-siemens", "scale-klarna-alphaevolve"]
authorship_disclosure: "AI-drafted from the cited public sources and independently checked by a second AI editorial-review agent (Codex) for source fit, claim boundaries, overlap, and reader utility. SWFT Editorial is responsible for corrections."
---

# Software factories and agent swarms across industries

What banks, trading firms, manufacturers and research teams have put into practice, and what their public accounts reveal.

> **Authorship:** AI-drafted from the cited public sources and independently checked by a second AI editorial-review agent (Codex) for source fit, claim boundaries, overlap, and reader utility. SWFT Editorial is responsible for corrections.

## Quick answer

Agent systems are appearing in banking, trading, automotive software, industrial planning and scientific research. Their public accounts describe different kinds of use: recurring development, deployed business workflows, pilots and research experiments. Compare the work and evidence before comparing scale.

Banks, trading firms and manufacturers are publishing detailed accounts of agents that write, test, review and improve software. Scientific teams are applying similar methods to experiments and proofs. The useful comparison starts with the work: a code migration, a reviewed change, a planning algorithm or a research finding.

This reference follows those uses across industries. The [company library](/companies) examines individual operating systems in depth. The [scale reference](/scale) records numerical observations, including the largest concurrent runs in this collection. Together they show where industrialized agent work is appearing and what its public evidence can establish.

## Industry overview

The examples below were checked on September 8, 2026. They are selected public accounts, not a representative survey. Their presence establishes documented activity; it cannot tell us what percentage of an industry has adopted software factories.

| Industry | Organizations and work | What the evidence establishes |
| --- | --- | --- |
| Trading | [Jane Street](/companies/jane-street-aide): internal coding tools and agent feedback | An internal harness used to build tools; no public swarm-size or firm-wide adoption figure |
| Banking | [Capital One](/companies/capital-one-agents), [Itaú](/companies/itau-devin), [Nubank](https://devin.ai/customers/nubank): security repair, development and migrations | Named operating workflows; metrics cover different populations and tasks |
| Financial operations | [BNY](https://www.bny.com/assets/corporate/documents/pdf/investor-relations/earnings/quarterly-update-presentation-1q-2026.pdf): Eliza and digital employees | A reported portfolio of deployed AI solutions, including multi-agent systems |
| Automotive | [BMW](https://www.press.bmwgroup.com/usa/article/detail/T0460064EN_US/bmw-i-ventures-in-coderabbit-to-advance-independent-ai-review-in-software-development), [Mercedes-Benz](https://cognition.com/blog/mercedes-benz-cognition): code review and modernization | Deployed review at BMW; a vendor-reported modernization pilot and rollout at Mercedes-Benz |
| Chemicals and supply chains | [BASF](https://cloud.google.com/blog/products/ai-machine-learning/how-basf-manages-thousands-of-supply-chain-decisions-with-alphaevolve): evolved planning algorithms | Initial historical simulations with a measured objective; live network-wide autonomy is not established |
| Payments and machine learning | [Klarna](https://engineering.klarna.com/beyond-prompting-how-algorithmic-evolution-doubled-our-training-speed-8f874af3080d): training-code optimization | Thousands of candidate programs evaluated against speed and reproducibility constraints |
| Enterprise platforms | [Palantir](https://www.palantir.com/docs/foundry/ai-fde/overview): AI FDE | A documented application-building agent; Palantir's internal adoption and swarm size remain undisclosed here |
| Industrial engineering | [Siemens](https://blogs.sw.siemens.com/art-of-the-possible/from-outdated-to-up-to-date-modernizing-deprecated-code-with-multi-agents/): simulation-code maintenance | A research prototype with specialist agent roles |
| Life sciences | [Daiichi Sankyo and Bayer Crop Science](https://blog.google/innovation-and-ai/technology/research/gemini-for-science-io-2026/): Co-Scientist | Named private-preview research use, reported by the provider |
| Scientific research | [OpenAI and other research systems](/scale): proofs, analysis and hypothesis search | Documented research runs; scientific acceptance and business adoption require their own evidence |

## Banking and trading: agents inside established engineering systems

Jane Street is a useful case because its development environment is unusually specific. Its [AIDE case](/companies/jane-street-aide) explains how an internal harness gives agents tools and feedback that fit the firm's software. For a product leader, the transferable question is how well an agent can work inside the systems the organization already depends on.

Capital One illustrates several task shapes within one company. Its [case](/companies/capital-one-agents) covers security investigation and repair, structured development plans, and a completed data-analysis workflow. Keeping those scopes separate makes the account more useful: repository coverage in a security system says little about the speed of an unrelated data task.

[Itaú's case](/companies/itau-devin) adds an adoption perspective. Its reported use spans development, testing, migrations and security repair. The denominator matters throughout: a share of teams is an adoption measure, while the share of scanner findings repaired describes one workflow's output.

Nubank supplies a particularly clear migration pattern. Its [customer account](https://devin.ai/customers/nubank) describes independent subtasks, examples for training and evaluation, and engineers who review and merge the resulting changes. It reports an eight-to-twelvefold improvement in engineering-time efficiency on the delegated scope. The roughly 100,000 data-class implementations describe the overall migration project, not a verified count completed by agents. This is a historical project account, with no publication date displayed.

For comparison, ask which part of the work can be repeated with little variation, how failed attempts are handled, and whether the reported savings include human review. Those questions travel well between a trading firm and a retail bank.

## BNY: counting deployed solutions

BNY's [2025 annual report](https://www.bny.com/corporate/global/en/investor-relations/annual-report-2025.html) defines its “digital employees” as multi-agent solutions. Its [Q1 2026 presentation, page 4](https://www.bny.com/assets/corporate/documents/pdf/investor-relations/earnings/quarterly-update-presentation-1q-2026.pdf), reports approximately 220 enterprise AI solutions in production and 140 digital employees. It names payments processing, anomaly detection and onboarding among the work areas.

This is evidence of agents entering business operations. The 140 figure counts systems, each of which may contain several agents and run many times. It does not disclose how many workers are active together, or how many collaborate on any one task. A deployment portfolio and a research swarm need different columns in a comparison.

## Automotive: review and modernization

BMW describes CodeRabbit supporting more than 1,000 software developers after over two years of collaboration. Its [August 2026 account](https://www.press.bmwgroup.com/usa/article/detail/T0460064EN_US/bmw-i-ventures-in-coderabbit-to-advance-independent-ai-review-in-software-development) places agent review in vehicle software development, using repository context, requirements and test results to assess proposed changes. That establishes a deployed review layer. It leaves the size of any wider coding-agent operation open.

At Mercedes-Benz, [Cognition's April 2026 account](https://cognition.com/blog/mercedes-benz-cognition) describes a four-week pilot involving more than 200,000 lines of COBOL. It reports eight days of modernization work against an eight-month estimate, followed by a wider product rollout. The estimate is a useful planning comparison, with no controlled causal result or public concurrency count attached.

These cases help a reader locate where agents enter an existing development process. BMW's evidence concerns assessing changes; the Mercedes-Benz example concerns transforming an older system. Neither requires every stage of development to become autonomous before it is worth studying.

## BASF and Klarna: searching for better algorithms

BASF and Google describe a system that repeatedly changes a planning program and tests it against historical supply-chain data. Their [May 2026 account](https://cloud.google.com/blog/products/ai-machine-learning/how-basf-manages-thousands-of-supply-chain-decisions-with-alphaevolve) reports thousands of experiments and an improvement over the initial model. The work was described as initial simulations. BASF's large production network explains the problem's importance; its size should not be read as the deployment coverage of the experiment.

Klarna describes a related loop for machine-learning training code. Over three weeks, AlphaEvolve evaluated nearly 6,000 candidate programs. The [engineering account](https://engineering.klarna.com/beyond-prompting-how-algorithmic-evolution-doubled-our-training-speed-8f874af3080d) reports throughput rising from 49 to roughly 97 samples per second under reproducibility constraints. Faster candidates that broke those constraints were rejected.

The shared pattern is a searchable space of programs and a test that can discriminate between them. A team can spend substantial compute exploring alternatives when each improvement will be reused. To judge the economics, readers still need the search cost, the cost of checking a candidate, and the number of future uses. Candidate count alone cannot supply that answer.

## Palantir: building inside the business platform

Palantir's [AI FDE documentation](https://www.palantir.com/docs/foundry/ai-fde/overview) describes an agent that builds pipelines, edits code and data models, and creates applications inside Foundry. It can run previews, inspect build checks and use the results to decide its next action. Changes are proposed through branches or pull requests by default, within the user's permissions.

This makes Palantir relevant to software factories across industries: the work can include the business's data structures and workflows alongside application code. Public product documentation establishes those capabilities. It does not establish how widely Palantir itself uses the system internally or the size of its agent teams. A detailed customer operating account would answer a different question and deserves its own evidence.

## Industrial and scientific research

Siemens' [Simcenter prototype](https://blogs.sw.siemens.com/art-of-the-possible/from-outdated-to-up-to-date-modernizing-deprecated-code-with-multi-agents/) divides code maintenance among agents that select documentation, update a simulation macro and explain the changes. The March 2025 post explicitly calls this research exploration. Three named roles do not establish three simultaneous workers or a deployed product.

Google's [May 2026 science announcement](https://blog.google/innovation-and-ai/technology/research/gemini-for-science-io-2026/) names Daiichi Sankyo, Bayer Crop Science and U.S. National Labs as Co-Scientist users in private preview. That is a meaningful sign of scientific uptake, with outcomes and scale still requiring partner-specific evidence.

Research also changes what counts as a finished result. A program may pass its tests; a proposed proof needs mathematical scrutiny; a biological hypothesis needs appropriate experimental validation. The [scale reference](/scale) preserves the reported task and outcome beside each number so these differences survive comparison.

## How to follow industry adoption

Track the depth of use within each organization: access to tools, a recurring workflow, multiple teams using it, and a documented share of work passing through it. Keep research experiments and commercial product capabilities labeled alongside operating accounts. These descriptions can coexist without implying a maturity ranking.

To estimate penetration across an industry, a study would need a defined population, a sampling method and a consistent definition of qualifying use. SWFT's public-source collection supplies examples and mechanisms. It also identifies the missing evidence that a broader study would need: adoption denominators, sustained output, rejected work, human effort and total cost.

The next useful step is to [compare the company workflows](/companies), inspect [reported scale](/scale), or use the [measurement guide](/software-factory-metrics) to evaluate one inside your own organization.

## How we know

- **First-party report (SELF-REPORT)** BNY describes digital employees as multi-agent systems and reports their deployment separately from the wider AI-solution portfolio. Sources: [Annual Report 2025](https://www.bny.com/corporate/global/en/investor-relations/annual-report-2025.html); [First Quarter 2026 Financial Results, page 4](https://www.bny.com/assets/corporate/documents/pdf/investor-relations/earnings/quarterly-update-presentation-1q-2026.pdf).
- **First-party report (SELF-REPORT)** BMW, Mercedes-Benz, Nubank, BASF and Klarna publish accounts of specific review, migration or optimization workflows. Sources: [BMW i Ventures invests in CodeRabbit to Advance Independent AI Review in Software Development](https://www.press.bmwgroup.com/usa/article/detail/T0460064EN_US/bmw-i-ventures-in-coderabbit-to-advance-independent-ai-review-in-software-development); [Engineering in the fast lane: Mercedes-Benz partners with Cognition](https://cognition.com/blog/mercedes-benz-cognition); [Nubank and Devin](https://devin.ai/customers/nubank); [How BASF manages thousands of supply chain decisions with AlphaEvolve’s agentic algorithms](https://cloud.google.com/blog/products/ai-machine-learning/how-basf-manages-thousands-of-supply-chain-decisions-with-alphaevolve); [Beyond Prompting: How Algorithmic Evolution Doubled our Training Speed](https://engineering.klarna.com/beyond-prompting-how-algorithmic-evolution-doubled-our-training-speed-8f874af3080d).
- **First-party report (SELF-REPORT)** Palantir documents application-building capabilities; Siemens labels its macro-maintenance prototype as research; Google names scientific partners in private preview. Sources: [AI FDE overview](https://www.palantir.com/docs/foundry/ai-fde/overview); [From outdated to up-to-date: Modernizing deprecated code with multi-agents](https://blogs.sw.siemens.com/art-of-the-possible/from-outdated-to-up-to-date-modernizing-deprecated-code-with-multi-agents/); [Gemini for Science: AI experiments and tools for a new era of discovery](https://blog.google/innovation-and-ai/technology/research/gemini-for-science-io-2026/).
- **Analysis (INFERENCE)** Selected operating accounts establish documented presence and reusable mechanisms; they cannot estimate industry penetration without a defined population and sampling method.

## Sources

- **First-party report (SELF-REPORT)** [How BASF manages thousands of supply chain decisions with AlphaEvolve’s agentic algorithms](https://cloud.google.com/blog/products/ai-machine-learning/how-basf-manages-thousands-of-supply-chain-decisions-with-alphaevolve) — BASF / Google Cloud; published 2026-05-07; accessed 2026-09-08. Joint operator/vendor account of initial historical simulations. Network size is context; the article does not establish a live autonomous planning deployment across that network.
- **First-party report (SELF-REPORT)** [BMW i Ventures invests in CodeRabbit to Advance Independent AI Review in Software Development](https://www.press.bmwgroup.com/usa/article/detail/T0460064EN_US/bmw-i-ventures-in-coderabbit-to-advance-independent-ai-review-in-software-development) — BMW Group; published 2026-08-12; accessed 2026-09-08. BMW describes deployed code review supporting over 1,000 developers. Vendor-wide review volume and future repair capabilities are separate claims.
- **First-party report (SELF-REPORT)** [Annual Report 2025](https://www.bny.com/corporate/global/en/investor-relations/annual-report-2025.html) — BNY; accessed 2026-09-08. Defines digital employees as multi-agent solutions. The Q1 2026 presentation supplies the later deployment count.
- **First-party report (SELF-REPORT)** [First Quarter 2026 Financial Results, page 4](https://www.bny.com/assets/corporate/documents/pdf/investor-relations/earnings/quarterly-update-presentation-1q-2026.pdf) — BNY; published 2026-04-16; accessed 2026-09-08. Q1 2026 deployment snapshot: approximately 220 enterprise AI solutions and 140 digital employees. These are solution counts, not simultaneous workers.
- **First-party report (SELF-REPORT)** [Gemini for Science: AI experiments and tools for a new era of discovery](https://blog.google/innovation-and-ai/technology/research/gemini-for-science-io-2026/) — Google; published 2026-05-19; accessed 2026-09-08. Names industrial and scientific partners in private preview. It does not provide per-partner swarm sizes or independently validated outcomes.
- **First-party report (SELF-REPORT)** [Engineering in the fast lane: Mercedes-Benz partners with Cognition](https://cognition.com/blog/mercedes-benz-cognition) — Cognition; published 2026-04-27; accessed 2026-09-08. Vendor account of a four-week pilot and wider rollout. Eight months is an estimated modernization baseline, not a measured control group.
- **First-party report (SELF-REPORT)** [Nubank and Devin](https://devin.ai/customers/nubank) — Cognition / Nubank; accessed 2026-09-08. Undated customer account of an ETL migration begun in 2023–2024. Project size is distinct from the subset delegated to agents.
- **First-party report (SELF-REPORT)** [AI FDE overview](https://www.palantir.com/docs/foundry/ai-fde/overview) — Palantir; accessed 2026-09-08. Product documentation for branch-based application and data development. Does not establish Palantir's own internal adoption or concurrent agent count.
- **First-party report (SELF-REPORT)** [From outdated to up-to-date: Modernizing deprecated code with multi-agents](https://blogs.sw.siemens.com/art-of-the-possible/from-outdated-to-up-to-date-modernizing-deprecated-code-with-multi-agents/) — Siemens; published 2025-03-13; accessed 2026-09-08. Simcenter research prototype with three agent roles; explicitly not a product delivery commitment or a concurrency measurement.
- **First-party report (SELF-REPORT)** [Beyond Prompting: How Algorithmic Evolution Doubled our Training Speed](https://engineering.klarna.com/beyond-prompting-how-algorithmic-evolution-doubled-our-training-speed-8f874af3080d) — Klarna Engineering; published 2026-03-30; accessed 2026-09-08. Operator account of candidate-program search, measured training speed, reproducibility constraints, and a prolonged plateau.

## Read next

- [Agent swarm scale: documented runs and adoption](/scale)
- [Jane Street AIDE: making agent work easier to verify](/companies/jane-street-aide)
- [Capital One: choosing the right checks for agent work](/companies/capital-one-agents)
