Nexos ResearchThe applied-research arm of Nexos

Applied research on agents, automation, and private AI.

We study how AI systems behave inside real businesses — and publish what we learn. What survives the lab becomes a Nexos service.

Programs

Research agenda

Four programs. Publications appear under the program they belong to.

01

Agent architectures for real workflows

How multi-agent systems should be structured to run actual business processes — delegation, tool use, and hand-offs that hold up outside a demo.

OutputReference architectures + open benchmarks

Active
02

Approval & autonomy design

When an agent should act, ask, or defer — and how to make human-in-the-loop feel like one clear decision instead of a growing queue.

OutputDesign patterns + a scored eval set

Active
03

Private & on-prem AI

Making open-weight models genuinely usable in-house: hardware sizing, tuning, and the honest trade-offs against hosted frontier models.

OutputSizing guides + deployment playbooks

Active
04

Automation reliability

Treating agent automations like production software — tracing, retries, self-evaluation, and the failure modes that only appear at scale.

OutputTooling + a reliability checklist

Ongoing