Ops & Back Office
Run reliable multi-step AI workflows across business tools
Cross-tool agent orchestration. This desk starts with the buyer's job and the evidence needed to make a defensible shortlist.
Buyer decision
Start with the outcome, not the logo.
The category only matters when it helps complete this specific buyer job.
This is the constraint statement every candidate must answer directly.
Cross-tool agent orchestration. Adjacent use cases belong in related desks when the likely winner changes.
Category-defining reference set
The tools this decision cannot ignore.
These products enter the research desk because they define the buyer's consideration set. Fame earns evaluation, not rank. Every verdict still needs niche-specific evidence and limitations.
Claude
Claude Code is an agentic coding tool that lives in your terminal, understands your codebase, and helps you code faster by executing routine tasks, explaining complex code, and handling git workflows - all through natural language commands.
OpenAI
A major generative video reference.
CrewAI
A major multi-agent framework.
LangGraph
A defining code-first agent orchestration framework.
Microsoft AutoGen
A prominent open source agent framework.
Discovered research set
11 more products under evaluation here.
These arrived through source ingestion rather than editorial selection. 13 records carry a dated metric. None of them is ranked, and inclusion is not endorsement.
haystack
Open-source AI orchestration framework for building context-engineered, production-ready LLM applications. Design modular pipelines and agent workflows with explicit control over retrieval, routing, memory, and generation. Built for scalable agents, RAG, multimodal applications, semantic search, and conversational systems.
open-multi-agent
TypeScript AI agent orchestration framework with dynamic workflows. Describe the goal, not the graph: a coordinator plans the task DAG at runtime and runs it on any LLM (Claude, ChatGPT, Gemini, DeepSeek, or local models).
FastGPT
FastGPT is a knowledge-based platform built on the LLMs, offers a comprehensive suite of out-of-the-box capabilities such as data processing, RAG retrieval, and visual AI workflow orchestration, letting you easily develop and deploy complex question-answering systems without the need for extensive setup or configuration.
ruflo
🌊 The leading agent meta-harness. Deploy intelligent multi-player swarms, coordinate autonomous workflows, and build conversational AI systems. Features adaptive memory, self-learning intelligence, RAG integration, and native Claude Code / Codex / Hermes and many more Integrated
dify
Build Agentic workflows, RAG pipelines, with rich AI model and tool support on one collaborative workspace. Deploy on cloud, VPC, or self-hosted, so teams move from prototype to production without rebuilding the stack.
hatchet
🪓 An orchestration engine for background tasks, AI agents, and durable workflows
golutra
Multi-agent AI orchestration platform for automation, workflows, and developer tools. Golutra transforms Codex, Claude Code, and OpenClaw into a unified agent system with parallel execution, task orchestration, long-running workflows, and AI productivity workspace.

Agent-MCP
Agent-MCP is a framework for creating multi-agent systems that enables coordinated, efficient AI collaboration through the Model Context Protocol (MCP). The system is designed for developers building AI applications that benefit from multiple specialized agents working in parallel on different aspects of a project.
agent-framework
A framework for building, orchestrating and deploying AI agents and multi-agent workflows with support for Python and .NET.
flowgram.ai
FlowGram is an extensible workflow development framework with built-in canvas, form, variable, and materials that helps developers build AI workflow platforms faster and simpler.
astron-agent
Enterprise-grade, commercial-friendly agentic workflow platform for building next-generation SuperAgents.
Evaluation framework
What a useful shortlist must prove.
These checks keep the desk useful before enough comparable, source-backed products are ready for a ranked recommendation.
- 01Process coverage
Walk through the full recurring process, including approvals, exceptions, reconciliation, and audit history.
- 02Reliability
Verify retries, failure alerts, recovery paths, support response, and accountability when an operation breaks.
- 03Governance
Review roles, approvals, document retention, compliance boundaries, and a durable record of changes.
- 04Total administrative cost
Count setup, maintenance, expert help, transaction fees, and manual cleanup alongside the plan price.
The job, intent, scope, and evaluation criteria are defined for this category.
Pricing, product constraints, buyer outcomes, and dated traction must be reviewed on the same basis.
A ranked recommendation appears only when the evidence supports a meaningful comparison.