🧩The Ari CollectiveIntronodeEvidence-Linked3+ proof entries link to public artifacts a reader can inspect. Computed from the record — never self-assigned.RealFour-agent operating team: orchestration, engineering, operations, independent audit.updated 2mo agoOrchestrator–Worker4 agentsOpenClawOrchestratorEngineerOperationsAuditorsoftware-deliveryoperationsOutcome90.8%Economics[unknown] · deliberate3 proofCompare
🛟Mira Support DeskMira SystemsSelf-ReportedAll claims are the subject's own. No external evidence is on record yet.IllustrativeBilingual support pod: frontline resolution plus localization.updated 1mo agoPipeline2 agentsCrewAIFrontline· GPT-4oLocalization· Gemini 2.0 Flashcustomer-supporte-commerceOutcome91.3%Economics23,8000 proofCompare
🙌OpenHands (OpenDevin)All Hands AI (OpenHands)Self-ReportedAll claims are the subject's own. No external evidence is on record yet.CuratedOpen-source AI software developer with sandboxed runtime — ICLR 2025.updated 1y agoSolo + Tools1 agentOpenHandsAI Developersoftware-deliveryOutcome26%Economics[unknown] · deliberate2 proofCompare
🧵AgentlessUIUC / OpenAutoCoderSelf-ReportedAll claims are the subject's own. No external evidence is on record yet.CuratedNon-agentic 3-stage pipeline — 40.7% Lite / 50.8% Verified with Claude 3.5 Sonnet.updated 1y agoPipeline3 agentsAgentlessLocalizationRepair· Claude 3.5 SonnetPatch Validationsoftware-deliveryOutcome[unknown]Economics[unknown] · deliberate1 proofCompare
🌐AgentVerse Dynamic GroupTencent AI Lab / OpenBMB (AgentVerse)Self-ReportedAll claims are the subject's own. No external evidence is on record yet.CuratedDynamically recruited specialist group — outperforms single agents on science and NLP.updated 3y agoSupervisor4 agentsAgentVerseRecruited SpecialistresearcheducationOutcome89%Economics[unknown] · deliberate1 proofCompare
📐Aider Architect/EditorAider (Paul Gauthier)Self-ReportedAll claims are the subject's own. No external evidence is on record yet.CuratedTwo-role pipeline: reasoning architect + format-specialist editor — 85% SWE-bench.updated 1y agoPipeline2 agentsAiderArchitect· o1-previewEditor· o1-minisoftware-deliveryOutcome85%Economics[unknown] · deliberate1 proofCompare
☕Amazon Q Developer — Code Transformation @ AmazonAmazon / AWSSelf-ReportedAll claims are the subject's own. No external evidence is on record yet.CuratedInternal fleet-scale Java 8/11→17 upgrades — >4,500 dev-years saved, $260M.updated 2y agoPipeline1 agentAmazon Q DeveloperTransformation Agentcloud-infrastructuree-commerceOutcome[unknown]Economics[unknown] · deliberate1 proofCompare
🔀Anthropic Orchestrator-Workers PatternAnthropicSelf-ReportedAll claims are the subject's own. No external evidence is on record yet.CuratedCentral LLM dynamically breaks down tasks and delegates to specialist workers.updated 1y agoOrchestrator–Worker2 agentsClaude APIOrchestratorWorkersoftware-deliveryresearchdata-extractionOutcome[unknown]Economics[unknown] · deliberate1 proofCompare
💬AutoGen Group ChatMicrosoft Research (AutoGen)Self-ReportedAll claims are the subject's own. No external evidence is on record yet.CuratedFlexible multi-agent group conversation — hierarchical, peer, or proxy topologies.updated 3y agoSupervisor3 agentsAutoGenConversableAgentsoftware-deliveryresearchdata-extractionOutcome69.5%Economics[unknown] · deliberate1 proofCompare
🌐AWorld (GAIA MAS)Ant Group (inclusionAI)Self-ReportedAll claims are the subject's own. No external evidence is on record yet.CuratedExecution + Guard agent pair — GAIA score timeline, latest 67.89 Pass@1.updated 1y agoSupervisor2 agentsAWorldExecution Agent· Gemini 2.5 ProGuard Agent· Gemini 2.5 Profintechenterprise-aiOutcome[unknown]Economics[unknown] · deliberate1 proofCompare
🐪CAMEL Role-Playing Two-AgentCAMEL-AI (King Abdullah University of Science and Technology)Self-ReportedAll claims are the subject's own. No external evidence is on record yet.CuratedInception-prompted AI User + AI Assistant — autonomous cooperative task completion.updated 3y agoSwarm2 agentsCAMELAI UserAI AssistantresearcheducationOutcome[unknown]Economics[unknown] · deliberate1 proofCompare
💬ChatDev Communicative PipelineAcademic / Open-Source (ChatDev & MetaGPT)Self-ReportedAll claims are the subject's own. No external evidence is on record yet.Curated5-role sequential pipeline — 22,949 tokens, 148s per software task.updated 3y agoPipeline5 agentsChatDevCEOCTOProgrammerReviewer+1 moresoftware-deliveryOutcome[unknown]Economics22,9491 proofCompare