prompt
FreeExpert prompt for building embodied AI systems with VLA pipelines and world models.
About prompt
This is a comprehensive system prompt designed for an Embodied AI Developer AI assistant. It provides a detailed framework for building Vision-Language-Action (VLA) systems, robotic agents, and world-model-driven embodied intelligence. The prompt covers core principles such as perception-action grounding, world models for foresight, modularity, and sim-to-real transfer. It defines architecture patterns including a VLA pipeline (perceive, understand, act), world-model-augmented planning, and conversational workflow execution. It also specifies skill action design with parameterized primitives like pick, place, navigate, push, and discusses cross-embodiment transfer via abstract action representations. The prompt is intended to guide an AI in generating accurate, practical responses for developers working on embodied AI and robotics.
Key Features
Pros & Cons
- Provides a comprehensive, expert-level framework for embodied AI
- Encourages grounded perception-action loops for robust behavior
- Supports modular design allowing easy swapping of components
- Includes practical skill primitives and cross-embodiment transfer concepts
- Promotes conversational interaction for task specification and reporting
- Requires deep background knowledge in robotics and VLA systems to use effectively
- May be overly complex for simple or non-robotic tasks
- Lacks concrete code implementations, only high-level guidance
- Assumes access to advanced simulation environments and hardware