Orrey
A Proving Ground for Drone Swarm Doctrine
Our Team
John Apessos
Mechanical Engineering
Kritanu Saha
Economics & History
David Diao
Public Policy
James Xiao
Mechanical Engineering & Computer Science
Orchestration is Becoming the Weapon System
Value is moving from the platforms to the layer that coordinates them, and no one currently owns that layer.
Key Insights:
Orchestration is vendor-neutral
Sources: CSIS, DIU.mil, DoW Directives, nso.nato,
Orchestration is underdeveloped
Auditable orchestration is novel
An Integrated Framework
We believe war in the future will be defined by the intersection of robust engineering, military doctrine, and policy insight.
New Doctrine
Cognitive Bias
Systems Thinking
True Agency
Introducing Orrey
Orrey aims to give battlespace leaders the ability to experiment with and learn from mission scenarios under their control.
Pre vs Post Training Comparison
Over multiple iterations, Orrey refines drone swarm strategies to adapt to hostile environments and available resources.
Pre-Training Iteration
Post-Training Iteration
LLMs in looped iterative learning
A structured learning loop: the model writes the policy in code, the simulator scores it, and every change arrives with the reasoning that produced it.
01 · POLICY
Deterministic policy, written as explicit Python Code, not weights.
02 · SIMULATION
Executes the policy and returns metrics: objective, assets lost, time, mesh integrity.
03 · LLM AGENT
Reasons over the metrics and writes the next policy, articulating why the last one failed and the new hypothesis.
04 · REASONING LOG
Every change stored with its rationale, versioned alongside the code it produced.
Each iteration is fed the full history of policies, parameters, and results
WHY NOT REINFORCEMENT LEARNING
WHAT THIS BUYS
Orrey implements a unique approach that utilizes LLM reasoning to communicate with a human decision maker.
Challenging Stable Assumptions
We search doctrine space with a language model and let physics grade it. The human has final authority over implemented doctrine.
OUTSIDE THE LOOP · THE HUMAN IS THE JUDGE
INSIDE THE LOOP · THE SIMULATOR IS THE JUDGE
LLM · STRATEGIST
Proposes the next formation or role allocation, and explains in language why the last one failed.
SIMULATOR · JUDGE
Ground truth from sim: objective met, assets lost, time, mesh held. The fitness function, fixed before any run.
hypothesis
verdict
PERSISTENT SEARCH LOG
Every formation, parameter, and result is fed back each iteration. Without it, this is hill climbing with amnesia, not search.
OUTPUT · CANDIDATE DOCTRINE
Not doctrine. A candidate, plus the rationale that generated it and the verdict that kept or killed it.
COMMANDER · THE JUDGE
Adopts, rejects, or bounds it. The machine never holds authority: integration produces capability, nothing produces authority.
The Audit Trail
Both plaintext reasoning records and direct python code decisions are logged and verified
Implementation and Feasibility
Adoption Rollout Plan
Phase 1: Prepare
Month
3
Phase 2: Pilot
Month
12
Phase 3: Scale
Month
24
Phase 4: Expand
Month
36
Customer and scope set
Paid pilot underway
Initial deployment live
Three teams, renewal
Dates and customer counts are proposed planning targets, not government commitments. Sources: DoDI 5000.61; DoDI 5000.97; Mar 2025 software acquisition directive; Jul 2025 Software Engineering Guide; Nov 2025 Acquisition Transformation Strategy; Jan 2026 AI Strategy and DETECT; AFWERX; DIU; SAM.gov; DFARS.
COST
INTEGRATION
REGULATORY
WHO BUYS
A 36-month rollout from one paid pilot to paying DoW programs, and what it takes to get there.
Orrey
Testable Swarm Doctrine | Judgement Driven Results