- VLM→VLA Survey // 2508.13073
- N/A
- WAVE 6
BRAGI
THE POET
BRAGI is a comprehensive survey of Vision-Language-Action (VLA) models for robotic manipulation. It catalogs the state-of-the-art in translating language instructions and visual observations into robot actions, covering architectures, training methods, and benchmark results across the field. Understanding the VLA landscape is essential for informed architecture decisions — BRAGI provides the theoretical foundation and comparative analysis that guides implementation choices across the ANIMA Intelligence tier.
MODULE STATUS: DEVELOPMENTStatus
VLA
- DIVISION
- ANIMA
- WAVE
- W6
- DOMAIN
- SURVEYS
- WAVE 6 // ANIMA SUITE
- SURVEY — VLM→VLA MODELS
SURVEY — VLM→VLA MODELS
Comprehensive survey of Vision-Language-Action models for robotics.
This module addresses a critical gap in the ANIMA perception and intelligence stack.
WHAT BRAGI DELIVERS
Comprehensive survey of Vision-Language-Action models for robotics.
CAPABILITIES
- Advanced VLM→VLA MODELS capabilities
- Integrated into ANIMA perception stack
- Bilingual documentation
- Production-ready architecture
WHY THIS IS HARD
Building BRAGI requires solving multiple coupled problems:
- 01Achieving real-time performance on edge hardware
- 02Maintaining accuracy across diverse conditions
- 03Seamless integration with existing ANIMA modules
- 04Robust operation in degraded environments
BRAGI solves these through careful architecture design and rigorous validation.
PROOF, NOT PROMISES
Key metrics:
| METRIC | VALUE |
|---|---|
| Status | Development |
| Backend | N/A |
| Integration | ANIMA Stack |
| Paper | VLM→VLA Survey |
WHAT'S BUILT TODAY
| COMPONENT | STATUS | NOTES |
|---|---|---|
| Core Architecture | COMPLETE | Validated |
| Training Pipeline | COMPLETE | Ready |
| Integration | IN PROGRESS | ANIMA stack |
| Edge Deploy | IN PROGRESS | Optimization |
| Core models | COMPLETE | Production validated |
| API layer | IN PROGRESS | REST API |
WHERE BRAGI DEPLOYS
- APP_01
ROBOTICS
Core module for autonomous robot systems.
- APP_02
DEFENSE
Military and security applications.
- APP_03
RESEARCH
Academic and industrial research platform.
UNDER THE HOOD
FOUNDATION: VLM→VLA SURVEY
- Advanced VLM→VLA MODELS capabilities
- Integrated into ANIMA perception stack
- Bilingual documentation
KEY INNOVATION
Comprehensive survey of Vision-Language-Action models for robotics.
DEPLOYMENT
- REST API
- Docker containerized
- Prometheus metrics
- Configurable backends
COMPUTE
- PRIMARY
- N/A
- EDGE
- Optimized inference
- API
- REST + streaming
PAPERS
- [01]VLM→VLA Survey (2508.13073)