Design Multimodal Agent Architecture with BeeAgent
planningChallengeNovember 25, 2025
Prompt Content
Outline an agent architecture using the BeeAgent framework that leverages Gemini 2.5 Pro for visual and language understanding. The architecture should include modules for: `Perception` (interpreting visual scenes and language instructions), `Planning` (generating high-level goals and action sequences), `Action Execution` (interacting with the simulated environment via tools), and a `Memory Management` module using Postgres and pgvector. Describe how Gemini 2.5 Pro will bridge the visual and language modalities within BeeAgent's workflow.
Related Prompts
Explore similar prompts from our community
Usage Tips
Copy the prompt and paste it into your preferred AI tool (Claude, ChatGPT, Gemini)
Customize placeholder values with your specific requirements and context
For best results, provide clear examples and test different variations