• Home
  • Search
  • OmniNova: A General Multimodal Multi-Agent Framework
  • https://doi.org/10.1109/trustcom66490.2025.00393Copy DOI Icon

OmniNova: A General Multimodal Multi-Agent Framework

  • Nov 14, 2025
  • Bingzhen Li +3 more
Show More
  • Abstract
  • Literature Map
  • References
  • Similar Papers
Abstract

The integration of Large Language Models (LLMs) with external tools enables intelligent automation far beyond text generation, yet coordinating multiple LLM-based agents remains difficult due to interaction overhead, resource waste, and fragile information flow. This paper introduces OmniNova, a modular and hierarchical multi-agent framework that unifies language models with capabilities for web search, browser automation, and code execution. OmniNova advances the state of the art through a hierarchical architecture that separates coordination, planning, supervision, and specialization; a dynamic routing mechanism that activates agents according to task complexity and state; and a multi-layered LLM integration strategy that allocates high-capability reasoning models only where they are cognitively necessary while assigning routine work to lighter models. Across 50 complex tasks in research, data analysis, and web interaction, OmniNova improves task completion (87% versus a 62% baseline), reduces token usage by 41%, and delivers higher human-rated quality (4.2/5 versus 3.1/5). The contribution includes both a principled system design and an open-source implementation intended to support research and practical deployment. Code is available at https://github.com/Superagentsys/OmniNoval.git.

Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.